← Search

Mei Wen

1 accepted papers

2025

SwiftPrune: Hessian-Free Weight Pruning for Large Language Models

EMNLP 2025

Post-training pruning, as one of the key techniques for compressing large language models (LLMs), plays a vital role in lightweight model deployment and model sparsity. However, current mainstream pruning methods dependent on the Hessian matrix face significant limitations in both pruning speed and

Cited by 0SourcePDFScholar