Adaptive Residual-Update Steering for Low-Overhead Hallucination Mitigation in Large Vision-Language Models
Large Vision-Language Models (LVLMs) typically process visual inputs as a prefix to the language decoder. As the model autoregressively generates text, this initial visual information inevitably undergoes ``dilution'', leading the model to over-rely on language priors and hallucinate objects. Existi…