2026
ImpQuant: Fine-Grained Importance-Aware Quantization for Large Vision-Language Models
ICML 2026poster
Large Vision–Language Models (LVLMs) have demonstrated remarkable capabilities across diverse multimodal tasks, yet their high inference costs necessitate low-bit deployment. Existing post-training quantization (PTQ) pipelines primarily adopt methodologies from text-only LLMs by treating multimodal …