2026
Reasoning-preserved Efficient Distillation of Large Language Models via Activation-aware Initialization
ICML 2026poster
Efficient Distillation (EDistill) compresses large language models (LLMs) by structured pruning parameters and tuning lightweight modules with high training efficiency. Although these EDistilled LLMs achieve state-of-the-art (SOTA) performance on general ability benchmarks relative to similarly size…