← Search

Manish Bhattarai

4 accepted papers

2026

PAS: Prelim Attention Score for Detecting Object Hallucinations in Large Vision-Language Models

CVPR 2026

Large vision-language models (LVLMs) are powerful, yet they remain unreliable due to object hallucinations. In this work, we show that in many hallucinatory predictions the LVLM effectively ignores the image and instead relies on previously generated output ("prelim") tokens to infer new objects. We

Cited by 0SourcecodeScholar
2025

Benchmarking Large Language Models with Integer Sequence Generation Tasks

NeurIPS 2025poster

We present a novel benchmark designed to rigorously evaluate the capabilities of large language models (LLMs) in mathematical reasoning and algorithmic code synthesis tasks. The benchmark comprises integer sequence generation tasks sourced from the Online Encyclopedia of Integer Sequences (OEIS), te…

Cited by 0SourceScholar
2025

LoRID: Low-Rank Iterative Diffusion for Adversarial Purification

AAAI 2025technical

This work presents an information-theoretic examination of diffusion-based purification methods, the state-of-the-art adversarial defenses that utilize diffusion models to remove malicious perturbations in adversarial examples. By theoretically characterizing the inherent purification errors associa…

Cited by 2SourcePDFScholar
2025

Topological Signatures of Adversaries in Multimodal Alignments

ICML 2025poster

Multimodal Machine Learning systems, particularly those aligning text and image data like CLIP/BLIP models, have become increasingly prevalent, yet remain susceptible to adversarial attacks. While substantial research has addressed adversarial robustness in unimodal contexts, defense strategies for…

Cited by 0SourcePDFScholar