← Search

Vishnu Prabhakaran

2 accepted papers

2025

VADE: Visual Attention Guided Hallucination Detection and Elimination

ACL 2025finding

Vision Language Models (VLMs) have achieved significant advancements in complex visual understanding tasks. However, VLMs are prone to hallucinations—generating outputs that lack alignment with visual content. This paper addresses hallucination detection in VLMs by leveraging the visual grounding in…

Cited by 0SourcePDFScholar
2025

VIT-Pro: Visual Instruction Tuning for Product Images

NAACL 2025industry

General vision-language models (VLMs) trained on web data struggle to understand and converse about real-world e-commerce product images. We propose a cost-efficient approach for collecting training data to train a generative VLM for e-commerce product images. The key idea is to leverage large-scale…

Cited by 0SourcePDFScholar