Tagged articles

multilingual TTS

3 articles · Page 1 of 1
Alibaba Cloud Developer
Alibaba Cloud Developer
Jul 21, 2026 · Artificial Intelligence

Qwen-Audio-3.0-TTS: From Speaking to Expressive Voice Synthesis

Qwen-Audio-3.0-TTS launches two variants—Flash with ~300 ms latency and Plus with higher naturalness—offering multilingual support for 16 languages, superior WER/CER and speaker similarity scores, free‑style natural‑language control, fine‑grained tag editing, and robust performance in noisy environments, all backed by benchmark results that crown Plus as the top performer on the Artificial Analysis leaderboard.

AI modelQwen-Audio-3.0-TTSmultilingual TTS
0 likes · 7 min read
Qwen-Audio-3.0-TTS: From Speaking to Expressive Voice Synthesis
Xiaomi Tech
Xiaomi Tech
Mar 18, 2026 · Artificial Intelligence

Xiaomi Unveils MiMo-V2-TTS: Giving Agents a Voice with Soul

Xiaomi introduces MiMo-V2-TTS, a self‑developed speech‑synthesis large model that combines a custom audio tokenizer, multi‑codebook architecture, massive pre‑training on over a hundred million hours of data and multi‑dimensional reinforcement learning to deliver fine‑grained style control, dialect support, role‑play and high‑quality singing, aiming to give AI agents expressive, human‑like voices.

audio tokenizerlarge modelmultilingual TTS
0 likes · 6 min read
Xiaomi Unveils MiMo-V2-TTS: Giving Agents a Voice with Soul
Ubuntu
Ubuntu
Jan 24, 2026 · Artificial Intelligence

Deploy Alibaba’s Qwen3‑TTS on Ubuntu and Clone Your Voice in 3 Seconds

This guide walks through installing the open‑source Qwen3‑TTS model on Ubuntu, covering environment setup, GPU requirements, package installation, model variants, and hands‑on Python scripts for ultra‑low‑latency voice cloning and text‑driven voice design.

AI speech synthesisPyTorchPython
0 likes · 9 min read
Deploy Alibaba’s Qwen3‑TTS on Ubuntu and Clone Your Voice in 3 Seconds