← Search

Hongyang Zhao

1 accepted papers

2025

Mixed-Precision Graph Neural Quantization for Low Bit Large Language Models

ICASSP 2025accepted

Post-Training Quantization (PTQ) is pivotal for deploying large language models (LLMs) within resource-limited settings by significantly reducing resource demands. However, existing PTQ strategies underperform at low bit levels (< 3 bits) due to the significant difference between the quantized and o…

Cited by 0SourceScholar