← Search

Lydia Chilton

3 accepted papers

2024

STORYSUMM: Evaluating Faithfulness in Story Summarization

EMNLP 2024main

Human evaluation has been the gold standard for checking faithfulness in abstractive summarization. However, with a challenging source domain like narrative, multiple annotators can agree a summary is faithful, while missing details that are obvious errors only once pointed out. We therefore introdu…

2023

StoryWars: A Dataset and Instruction Tuning Baselines for Collaborative Story Understanding and Generation

ACL 2023long

Collaborative stories, which are texts created through the collaborative efforts of multiple authors with different writing styles and intentions, pose unique challenges for NLP models. Understanding and generating such stories remains an underexplored area due to the lack of open-domain corpora. To…

2022

SafeText: A Benchmark for Exploring Physical Safety in Language Models

EMNLP 2022main

Understanding what constitutes safe text is an important issue in natural language processing and can often prevent the deployment of models deemed harmful and unsafe. One such type of safety that has been scarcely studied is commonsense physical safety, i.e. text that is not explicitly violent and…