EN Submit a tool
Model

Tongyi Lab Releases Qwen-Audio-3.0-TTS Real-Time Speech Synthesis Model

Published: Source: WeChat Official Account: Tongyi Lab (Qwen)

ShareXFacebookTelegramWhatsApp

Tongyi Lab releases Qwen-Audio-3.0-TTS, including Flash (first packet delay ~300ms) and Plus versions. The Plus version tops the Artificial Analysis leaderboard, supports 16 languages and 20 Chinese dialects, with average WER/CER as low as 3.87 (Flash) and speaker similarity up to 82.75 (Plus).

Read the original (opens in a new tab)

News stream data aggregated by AI HOT

Related newsLatest in this category