← Search

Maan Qraitem

3 accepted papers

2025

KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models

ICLR 2025poster

This paper investigates visual analogical reasoning in large multimodal models (LMMs) compared to human adults and children. A “visual analogy” is an abstract rule inferred from one image and applied to another. While benchmarks exist for testing visual reasoning in LMMs, they require advanced skill…

2025

Web Artifact Attacks Disrupt Vision Language Models

ICCV 2025poster

Vision-language models (VLMs) (e.g., CLIP, LLaVA) are trained on large-scale, lightly curated web datasets, leading them to learn unintended correlations between semantic concepts and unrelated visual signals. These associations degrade model accuracy by causing predictions to rely on incidental pat…

2023

Bias Mimicking: A Simple Sampling Approach for Bias Mitigation

CVPR 2023poster

Prior work has shown that Visual Recognition datasets frequently underrepresent bias groups B (e.g. Female) within class labels Y (e.g. Programmers). This dataset bias can lead to models that learn spurious correlations between class labels and bias groups such as age, gender, or race. Most recent m…