Surprising Discovery: A Chinese embodied‑AI company solves the distributed Muon bottleneck
The article analyzes how the Muon optimizer, adopted by DeepSeek‑V4 and Kimi‑K2, suffers a 2.2× overhead in distributed training, and how the DMuon system from Zibian Robot reduces that overhead to near‑AdamW levels, achieving up to 97.4× speedup and only 2% slower end‑to‑end training than AdamW.
