← Search

Mohammad Shahab Sepehri

2 accepted papers

2025

Hyperphantasia: A Benchmark for Evaluating the Mental Visualization Capabilities of Multimodal LLMs

NeurIPS 2025poster

Mental visualization, the ability to construct and manipulate visual representations internally, is a core component of human cognition and plays a vital role in tasks involving reasoning, prediction, and abstraction. Despite the rapid progress of Multimodal Large Language Models (MLLMs), current be…

Cited by 0SourceScholar
2025

MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation models

ICLR 2025poster

Multimodal Large Language Models (MLLMs) have tremendous potential to improve the accuracy, availability, and cost-effectiveness of healthcare by providing automated solutions or serving as aids to medical professionals. Despite promising first steps in developing medical MLLMs in the past few years…

Cited by 152SourcePDFScholar