← Search

Katherine Thai

7 accepted papers

2024

One Thousand and One Pairs: A “novel” challenge for long-context language models

EMNLP 2024main

Synthetic long-context LLM benchmarks (e.g., “needle-in-the-haystack”) test only surface-level retrieval capabilities; but how well can long-context LLMs retrieve, synthesize, and reason over information across book-length inputs? We address this question by creating NoCha, a dataset of 1,001 minima…

2022

DEMETR: Diagnosing Evaluation Metrics for Translation

EMNLP 2022main

While machine translation evaluation metrics based on string overlap (e.g., BLEU) have their limitations, their computations are transparent: the BLEU score assigned to a particular candidate translation can be traced back to the presence or absence of certain words. The operations of newer learned…

2022

Exploring Document-Level Literary Machine Translation with Parallel Paragraphs from World Literature

EMNLP 2022main

Literary translation is a culturally significant task, but it is bottlenecked by the small number of qualified literary translators relative to the many untranslated works published around the world. Machine translation (MT) holds potential to complement the work of human translators by improving bo…