← Search

Binyu Yang

1 accepted papers

2026

Reasoning-preserved Efficient Distillation of Large Language Models via Activation-aware Initialization

ICML 2026poster

Efficient Distillation (EDistill) compresses large language models (LLMs) by structured pruning parameters and tuning lightweight modules with high training efficiency. Although these EDistilled LLMs achieve state-of-the-art (SOTA) performance on general ability benchmarks relative to similarly size…

Cited by 0SourceScholar