2025
MAPS: Advancing Multi-Modal Reasoning in Expert-Level Physical Science
ICLR 2025poster
Pre-trained on extensive text and image corpora, current Multi-Modal Large Language Models (MLLM) have shown strong capabilities in general visual reasoning tasks. However, their performance is still lacking in physical domains that require understanding diagrams with complex physical structures an…