2026
Reasoning-Driven Anomaly Detection and Localization with Image-Level Supervision
CVPR 2026
Multimodal large language models (MLLMs) have recently demonstrated remarkable reasoning and perceptual abilities for anomaly detection. However, most approaches remain confined to image-level anomaly detection and textual reasoning, while pixel-level localization still relies on external vision mod