2025
CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models
COLING 2025main
With the rapid advancements in multimodal large language models, evaluating their multimodal mathematical capabilities continues to receive wide attention. Although datasets such as MathVista have been introduced for evaluating mathematical capabilities in multimodal scenarios, there remains a lack…