← Search

Pranay Kumar Yelugam

2 accepted papers

2024

Every Answer Matters: Evaluating Commonsense with Probabilistic Measures

ACL 2024long

Large language models have demonstrated impressive performance on commonsense tasks; however, these tasks are often posed as multiple-choice questions, allowing models to exploit systematic biases. Commonsense is also inherently probabilistic with multiple correct answers. The purpose of “boiling wa…

2022

DISAPERE: A Dataset for Discourse Structure in Peer Review Discussions

NAACL 2022long

At the foundation of scientific evaluation is the labor-intensive process of peer review. This critical task requires participants to consume vast amounts of highly technical text. Prior work has annotated different aspects of review argumentation, but discourse relations between reviews and rebutta…

Cited by 27SourcePDFScholar