2026
Teacher-Guided Routing for Sparse Vision Mixture-of-Experts
CVPR 2026
Recent progress in deep learning has been driven by increasingly large-scale models, but the resulting computational cost has become a critical bottleneck. Sparse Mixture of Experts (MoE) offers an effective solution by activating only a small subset of experts for each input, achieving high scalabi