2026
Efficient Plug-and-Play Weight Refinement for Sparse Large Models
AAAI 2026technical
One-shot pruning efficiently compresses Large Language Models but produces coarse sparse weights, causing significant performance degradation. Traditional fine-tuning approaches to refine these weights are prohibitively expensive for large models. This highlights the need for a training-free weight