2026
SeGO: Sensitivity-Aware Golden Optimization for Large-Scale VLM Quantization
IJCAI 2026
The deployment of Vision-Language Models (VLMs) faces memory and computational bottlenecks because of the massive parameters and intensive computations. While Post-Training Quantization (PTQ) can reduce these costs, existing methods often overlook the heterogeneity of multimodal input when applied t