← Search

Timothy Leffel

2 accepted papers

2024

Leveraging LLMs for Dialogue Quality Measurement

NAACL 2024industry

In task-oriented conversational AI evaluation, unsupervised methods poorly correlate with human judgments, and supervised approaches lack generalization. Recent advances in large language models (LLMs) show robust zero- and few-shot capabilities across NLP tasks. Our paper explores using LLMs for au…

Cited by 4SourcePDFScholar
2023

Toward More Accurate and Generalizable Evaluation Metrics for Task-Oriented Dialogs

ACL 2023industry

Measurement of interaction quality is a critical task for the improvement of large-scale spoken dialog systems. Existing approaches to dialog quality estimation either focus on evaluating the quality of individual turns, or collect dialog-level quality measurements from end users immediately followi…

Cited by 3SourcePDFScholar