2025
Bypass Back-propagation: Optimization-based Structural Pruning for Large Language Models via Policy Gradient
ACL 2025long
Recent Large-Language Models (LLMs) pruning methods typically operate at the post-training phase without the expensive weight finetuning, however, their pruning criteria often rely on **heuristically hand-crafted metrics**, potentially leading to suboptimal performance. We instead propose a novel **…