Alibaba has launched Qwen-Audio-3.0-TTS, a new text-to-speech model that offers significant improvements in control, multilingual support, and acoustic robustness. The model comes in two versions: Flash for real-time interaction with low latency and Plus for high-quality generation. Qwen-Audio-3.0-TTS-Plus has achieved the top position on the Artificial Analysis Speech Arena leaderboard, outperforming competitors like Google and Speechify. AI
IMPACT Sets a new benchmark for TTS quality and control, potentially influencing enterprise adoption and further research in expressive speech synthesis.
RANK_REASON New model release from a major tech company (Alibaba) with performance claims and benchmark results.
- Alibaba
- Artificial Analysis
- Fun-Realtime-TTS
- Qwen-Audio-3.0-TTS
- Qwen-Audio-3.0-TTS-Plus
- Alibaba Cloud Model Studio
- DashScope
- Tongyi Lab
- Speechify
AI-generated summary · Google Gemini · from 4 sources. How we write summaries →