Tagged articles

Multilingual speech synthesis

2 articles · Page 1 of 1
Weekly Large Model Application
Weekly Large Model Application
Jun 10, 2026 · Artificial Intelligence

OmniVoice: A Zero‑Shot TTS Paradigm Covering 600+ Languages

OmniVoice introduces a single‑stage, diffusion‑style language model that maps text directly to multi‑codebook acoustic tokens, achieving zero‑shot voice cloning for over 600 languages with high intelligibility and real‑time factor as low as 0.025, making it suitable for large‑scale multilingual deployment.

Acoustic tokenDiffusion Language ModelMultilingual speech synthesis
0 likes · 8 min read
OmniVoice: A Zero‑Shot TTS Paradigm Covering 600+ Languages
Xiaomi Tech
Xiaomi Tech
May 7, 2026 · Artificial Intelligence

OmniVoice: Open‑Source TTS Model Clones Voices in 600+ Languages with a Single Architecture

OmniVoice, an open‑source TTS system from Xiaomi AI Lab, uses a minimalist bidirectional Transformer and LLM‑enhanced pre‑training to synthesize high‑quality speech in over 600 languages, outperforming commercial systems while offering fine‑grained control and fully public code and models.

Multilingual speech synthesisOmniVoiceOpen-source
0 likes · 8 min read
OmniVoice: Open‑Source TTS Model Clones Voices in 600+ Languages with a Single Architecture