← Search

Viktor Kun{\v{c}}ak

2 accepted papers

2025

Reliable Evaluation and Benchmarks for Statement Autoformalization

EMNLP 2025

Evaluating statement autoformalization, translating natural language mathematics into formal languages like Lean 4, remains a significant challenge, with few metrics, datasets, and standards to robustly measure progress. In this work, we present a comprehensive approach combining improved metrics, r