← Search

Max Grusky

1 accepted papers

2023

Rogue Scores

ACL 2023long

Correct, comparable, and reproducible model evaluation is essential for progress in machine learning. Over twenty years, thousands of language and vision models have been evaluated with a popular metric called ROUGE. Does this widespread benchmark metric meet these three evaluation criteria? This sy…

Cited by 24SourcePDFScholar