AIBase
Home
AI NEWS
AI Tools
GEO & AEO
MCP
AI Models
EN

AI News

View More

Qwen-Audio-3.0-TTS: A Voice Synthesis Large Model from Qwen - Direct Command with Natural Language, Real-time Version First Packet Delay Reduced to 300 Milliseconds

Qwen released Qwen-Audio-3.0-TTS, a text-to-speech model featuring free-style natural language instruction control. Users can simply describe desired tone, pace, and style in everyday language, eliminating the need to adjust engineering parameters and drastically lowering the usage barrier.....

10.1k 3 hours ago
Qwen-Audio-3.0-TTS: A Voice Synthesis Large Model from Qwen - Direct Command with Natural Language, Real-time Version First Packet Delay Reduced to 300 Milliseconds

First Packet Delay 300ms, Supports 20 Dialects: Tongyi Qianwen Qwen-Audio-3.0-TTS Officially Opened

Alibaba's Tongyi Qianwen launched Qwen-Audio-3.0-TTS, a real-time TTS model advancing from speech to expression. Its Plus version tops Artificial Analysis Speech Arena, beating Gemini 3.1 TTS. Flash version offers low-latency (~300ms first packet) interaction; Plus focuses on high naturalness and timbre fidelity.....

17.4k 1 hours ago
First Packet Delay 300ms, Supports 20 Dialects: Tongyi Qianwen Qwen-Audio-3.0-TTS Officially Opened
AIBase
Empowering the future, your artificial intelligence solution think tank
English简体中文繁體中文にほんご
FirendLinks:
AI Newsletters AI ToolsMCP ServersAI NewsAI MarketingLLM LeaderboardAI Ranking
© 2026AIBase
Business CooperationSite Map