PulseAugur
EN
LIVE 13:53:39
中文(ZH) AI语音进入“表演时代”:阿里Qwen-Audio-3.0-TTS登顶全球权威榜单

Alibaba's Qwen-Audio-3.0-TTS leads global speech synthesis benchmarks

Alibaba has released its Qwen-Audio-3.0-TTS, a new large model for speech synthesis that offers significant improvements in fine-grained control, instruction following, multilingual support, and acoustic robustness. The model comes in two versions: Qwen-Audio-3.0-TTS-Plus for high-quality generation and Fun-Realtime-TTS for real-time interaction. Qwen-Audio-3.0-TTS-Plus has achieved the top position on the Artificial Analysis global leaderboard, with its preview version also previously topping the chart. AI

IMPACT Sets a new standard for expressive and multilingual speech synthesis, potentially impacting voice acting, gaming, and real-time communication applications.

RANK_REASON New model release from a major AI lab (Alibaba) with benchmark performance claims. [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on 量子位 (QbitAI) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Alibaba's Qwen-Audio-3.0-TTS leads global speech synthesis benchmarks

COVERAGE [1]

  1. 量子位 (QbitAI) TIER_1 中文(ZH) · 梦晨 ·

    AI Voice Enters the "Performance Era": Alibaba's Qwen-Audio-3.0-TTS Tops Global Authoritative Rankings

    细粒度标签+ 20 种方言