← Search

Jinlun Ye

2 accepted papers

2026

Instruction Lens Score: Your Instruction Contributes a Powerful Object Hallucination Detector for Multimodal Large Language Models

ICML 2026poster

Multimodal large language models (MLLMs) have achieved remarkable progress, yet the object hallucination remains a critical challenge for reliable deployment. In this paper, we present an in-depth analysis of instruction token embeddings and reveal that they implicitly encode visual information whil…

Cited by 0SourceScholar
2026

TTL: Test-time Textual Learning for OOD Detection with Pretrained Vision-Language Models

CVPR 2026

Vision-language models (VLMs) such as CLIP exhibit strong Out-of-distribution (OOD) detection capabilities by aligning visual and textual representations. Recent CLIP-based test-time adaptation methods further improve detection performance by incorporating external OOD labels. However, such labels a

Cited by 0SourcecodeScholar