← Search

Shi Liu*

1 accepted papers

2024

Paying More Attention to Images: A Training-Free Method for Alleviating Hallucination in LVLMs

ECCV 2024poster

"Existing Large Vision-Language Models (LVLMs) primarily align image features of vision encoder with Large Language Models (LLMs) to leverage their superior text generation capabilities. However, the scale disparity between vision encoder and language model may led to LLMs assuming a predominant rol…