← Search

Roland Daynauth

2 accepted papers

2025

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

ACL 2025long

Evaluating large language model (LLM) is a complex task. Pairwise ranking has emerged as state-of-the-art method to evaluate human preferences by having humans compare pairs of LLM outputs based on predefined criteria, enabling ranking across multiple LLMs by aggregating pairwise results through alg…

2024

GuyLingo: The Republic of Guyana Creole Corpora

NAACL 2024short

While major languages often enjoy substantial attention and resources, the linguistic diversity across the globe encompasses a multitude of smaller, indigenous, and regional languages that lack the same level of computational support. One such region is the Caribbean. While commonly labeled as “Engl…