Tagged articles

expert routing

3 articles · Page 1 of 1
Cambridge Mofang Notes
Cambridge Mofang Notes
Sep 1, 2026 · Artificial Intelligence

From Dense to MoE: Decoding Total vs. Activated Parameters

This article explains the distinction between total and activated parameters in Mixture-of-Experts (MoE) models, contrasting dense and sparse architectures, detailing expert routing mechanisms, and analyzing memory and compute implications across model loading, prefill, and decode stages.

Mixture of ExpertsMoEactivated parameters
0 likes · 15 min read
From Dense to MoE: Decoding Total vs. Activated Parameters
Machine Heart
Machine Heart
Jul 31, 2026 · Artificial Intelligence

A²-Edit: Overcoming Object Category and Mask Precision Limits for Precise Reference-Guided Image Editing

A²-Edit introduces a unified framework that handles arbitrary object categories and coarse masks through mixed‑Transformer expert routing, mask‑annealing training, and a 500K multi‑category dataset, achieving robust, high‑quality reference‑guided edits across e‑commerce, virtual try‑on, and visual effects scenarios.

UniEdit-500Kcross-category modelingexpert routing
0 likes · 10 min read
A²-Edit: Overcoming Object Category and Mask Precision Limits for Precise Reference-Guided Image Editing
Machine Heart
Machine Heart
Jun 7, 2026 · Artificial Intelligence

FusionRoute: Token-Level Expert Routing and Self-Correction for Multi-LLM Collaboration

FusionRoute introduces a token‑level routing framework that dynamically selects the most suitable expert LLM for each token and adds a complementary generation step, enabling fine‑grained, stable multi‑model collaboration that outperforms existing sequence‑level and expert‑selection methods across diverse benchmarks.

AI researchLarge Language ModelsModel Merging
0 likes · 11 min read
FusionRoute: Token-Level Expert Routing and Self-Correction for Multi-LLM Collaboration