2025
MBQ: Modality-Balanced Quantization for Large Vision-Language Models
CVPR 2025poster
Vision-Language Models (VLMs) have already enabled a variety of real-world applications. The large parameter size of VLMs brings large memory and computation overhead which poses significant challenges for deployment. Post-Training Quantization (PTQ) is an effective technique to reduce the memory an…