2025
Mixed-Precision Graph Neural Quantization for Low Bit Large Language Models
ICASSP 2025accepted
Post-Training Quantization (PTQ) is pivotal for deploying large language models (LLMs) within resource-limited settings by significantly reducing resource demands. However, existing PTQ strategies underperform at low bit levels (< 3 bits) due to the significant difference between the quantized and o…