← Search

Nishant Raj

1 accepted papers

2022

DEMETR: Diagnosing Evaluation Metrics for Translation

EMNLP 2022main

While machine translation evaluation metrics based on string overlap (e.g., BLEU) have their limitations, their computations are transparent: the BLEU score assigned to a particular candidate translation can be traced back to the presence or absence of certain words. The operations of newer learned…