← Search

John Mendonça

1 accepted papers

2024

Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs

EMNLP 2024finding

Although human evaluation remains the gold standard for open-domain dialogue evaluation, the growing popularity of automated evaluation using Large Language Models (LLMs) has also extended to dialogue. However, most frameworks leverage benchmarks that assess older chatbots on aspects such as fluency…