2025
MeMoTune: A Measure and Moment-Driven Fine-Tuning Framework for Quantized Large Language Models
ACL 2025finding
Quantizing large language models (LLMs) is essential for reducing memory and computational costs in natural language processing. Existing methods combine quantization with parameter-efficient fine-tuning but often fail to meet practical performance requirements. This paper introduces MeMoTune, a nov…