← Search

Ilia Kuznetsov

10 accepted papers

2025

STRICTA: Structured Reasoning in Critical Text Assessment for Peer Review and Beyond

ACL 2025long

Critical text assessment is at the core of many expert activities, such as fact-checking, peer review, and essay grading. Yet, existing work treats critical text assessment as a black box problem, limiting interpretability and human-AI collaboration. To close this gap, we introduce Structured Reason…

Cited by 0SourcePDFScholar
2024

Are Large Language Models Good Classifiers? A Study on Edit Intent Classification in Scientific Document Revisions

EMNLP 2024main

Classification is a core NLP task architecture with many potential applications. While large language models (LLMs) have brought substantial advancements in text generation, their potential for enhancing classification tasks remains underexplored. To address this gap, we propose a framework for thor…

2024

M2QA: Multi-domain Multilingual Question Answering

EMNLP 2024finding

Generalization and robustness to input variation are core desiderata of machine learning research. Language varies along several axes, most importantly, language instance (e.g. French) and domain (e.g. news). While adapting NLP models to new languages within a single domain, or to new domains within…

2024

Re3: A Holistic Framework and Dataset for Modeling Collaborative Document Revision

ACL 2024long

Collaborative review and revision of textual documents is the core of knowledge work and a promising target for empirical analysis and NLP assistance. Yet, a holistic framework that would allow modeling complex relationships between document revisions, reviews and author responses is lacking. To add…

2024

Systematic Task Exploration with LLMs: A Study in Citation Text Generation

ACL 2024long

Large language models (LLMs) bring unprecedented flexibility in defining and executing complex, creative natural language generation (NLG) tasks. Yet, this flexibility brings new challenges, as it introduces new degrees of freedom in formulating the task inputs and instructions and in evaluating mod…

2023

CiteBench: A Benchmark for Scientific Citation Text Generation

EMNLP 2023long main

Science progresses by building upon the prior body of knowledge documented in scientific publications. The acceleration of research makes it hard to stay up-to-date with the recent developments and to summarize the ever-growing body of prior work. To address this, the task of citation text generatio…

Cited by 28SourcecodeScholar
2023

NLPeer: A Unified Resource for the Computational Study of Peer Review

ACL 2023long

Peer review constitutes a core component of scholarly publishing; yet it demands substantial expertise and training, and is susceptible to errors and biases. Various applications of NLP for peer reviewing assistance aim to support reviewers in this complex process, but the lack of clearly licensed d…

2022

Yes-Yes-Yes: Proactive Data Collection for ACL Rolling Review and Beyond

EMNLP 2022finding

The shift towards publicly available text sources has enabled language processing at unprecedented scale, yet leaves under-serviced the domains where public and openly licensed data is scarce. Proactively collecting text data for research is a viable strategy to address this scarcity, but lacks syst…