← Search

Sennay Ghebreab

2 accepted papers

2026

Same Content, Different Answers: Cross-Modal Inconsistency in MLLMs

CVPR 2026

We introduce two new benchmarks REST and REST+ (Render-Equivalence Stress Tests) to enable systematic evaluation of cross-modal inconsistency in multimodal large language models (MLLMs). MLLMs are trained to represent vision and language in the same embedding space, yet they cannot perform the same

Cited by 0SourceScholar