← Search

Lorenzo Tordi

1 accepted papers

2025

Can Large Language Models Win the International Mathematical Games?

EMNLP 2025

Recent advances in large language models (LLMs) have demonstrated strong mathematical reasoning abilities, even in visual contexts, with some models surpassing human performance on existing benchmarks. However, these benchmarks lack structured age categorization, clearly defined skill requirements,