← Search

Hongrui Zhan

1 accepted papers

2025

BigMac: A Communication-Efficient Mixture-of-Experts Model Structure for Fast Training and Inference

AAAI 2025technical

The Mixture-of-Experts (MoE) structure scales the Transformer-based large language models (LLMs) and improves their performance with only the sub-linear increase in computation resources. Recently, a fine-grained DeepSeekMoE structure is proposed, which can further improve the computing efficiency o…