2025
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models
EMNLP 2025
Post-training pruning, as one of the key techniques for compressing large language models (LLMs), plays a vital role in lightweight model deployment and model sparsity. However, current mainstream pruning methods dependent on the Hessian matrix face significant limitations in both pruning speed and