← Search

Antonella Carbonaro

2 accepted papers

2025

Can Large Language Models Win the International Mathematical Games?

EMNLP 2025

Recent advances in large language models (LLMs) have demonstrated strong mathematical reasoning abilities, even in visual contexts, with some models surpassing human performance on existing benchmarks. However, these benchmarks lack structured age categorization, clearly defined skill requirements,

2022

[RETRACTED] NLG-Metricverse: An End-to-End Library for Evaluating Natural Language Generation

COLING 2022main

Driven by deep learning breakthroughs, natural language generation (NLG) models have been at the center of steady progress in the last few years, with a ubiquitous task influence. However, since our ability to generate human-indistinguishable artificial text lags behind our capacity to assess it, it…

Cited by 21SourcePDFScholar