← Search

Tommaso Bonomo

2 accepted papers

2025

BOOKCOREF: Coreference Resolution at Book Scale

ACL 2025long

Coreference Resolution systems are typically evaluated on benchmarks containing small- to medium-scale documents.When it comes to evaluating long texts, however, existing benchmarks, such as LitBank, remain limited in length and do not adequately assess system capabilities at the book scale, i.e., w…

2025

LiteraryQA: Towards Effective Evaluation of Long-document Narrative QA

EMNLP 2025

Question Answering (QA) on narrative text poses a unique challenge to current systems, requiring a deep understanding of long, complex documents. However, the reliability of NarrativeQA, the most widely used benchmark in this domain, is hindered by noisy documents and flawed QA pairs. In this work,