← Search

Nicholas Jiang

2 accepted papers

2025

Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

ICLR 2025poster

We investigate the internal representations of vision-language models (VLMs) to address hallucinations, a persistent challenge despite advances in model size and training. We project VLMs’ internal image representations to their language vocabulary and observe more confident output probabilities on…

2025

Vision Transformers Don't Need Trained Registers

NeurIPS 2025spotlight

We investigate the mechanism underlying a previously identified phenomenon in Vision Transformers -- the emergence of high-norm tokens that lead to noisy attention maps (Darcet et al., 2024). We observe that in multiple models (e.g., CLIP, DINOv2), a sparse set of neurons is responsible for concentr…

Cited by 0SourcecodeScholar