← Search

Xuange Gao

2 accepted papers

2025

Ada-K Routing: Boosting the Efficiency of MoE-based LLMs

ICLR 2025poster

In the era of Large Language Models (LLMs), Mixture-of-Experts (MoE) architectures offer a promising approach to managing computational costs while scaling up model parameters. Conventional MoE-based LLMs typically employ static Top-K routing, which activates a fixed and equal number of experts for…

Cited by 1SourcePDFScholar