← Search

Jay DeYoung

3 accepted papers

2026

Improving Attributed Long-form Question Answering with Intent Awareness

ICLR 2026poster

Large language models (LLMs) are increasingly being used to generate comprehensive, knowledge-intensive reports. However, while these models are trained on diverse academic papers and reports, they are not exposed to the reasoning processes and intents that guide authors in crafting these documents.…

Cited by 0SourceScholar
2023

Automated Metrics for Medical Multi-Document Summarization Disagree with Human Evaluations

ACL 2023long

Evaluating multi-document summarization (MDS) quality is difficult. This is especially true in the case of MDS for biomedical literature reviews, where models must synthesize contradicting evidence reported across different documents. Prior work has shown that rather than performing the task, models…

2021

MSˆ2: Multi-Document Summarization of Medical Studies

EMNLP 2021main

To assess the effectiveness of any medical intervention, researchers must conduct a time-intensive and manual literature review. NLP systems can help to automate or assist in parts of this expensive process. In support of this goal, we release MSˆ2 (Multi-Document Summarization of Medical Studies),…