Tagged articles

Qwen-Audio-3.0-TTS

1 articles · Page 1 of 1
Alibaba Cloud Developer
Alibaba Cloud Developer
Jul 21, 2026 · Artificial Intelligence

Qwen-Audio-3.0-TTS: From Speaking to Expressive Voice Synthesis

Qwen-Audio-3.0-TTS launches two variants—Flash with ~300 ms latency and Plus with higher naturalness—offering multilingual support for 16 languages, superior WER/CER and speaker similarity scores, free‑style natural‑language control, fine‑grained tag editing, and robust performance in noisy environments, all backed by benchmark results that crown Plus as the top performer on the Artificial Analysis leaderboard.

AI modelQwen-Audio-3.0-TTSmultilingual TTS
0 likes · 7 min read
Qwen-Audio-3.0-TTS: From Speaking to Expressive Voice Synthesis