← Search

Eric Davis

3 accepted papers

2024

TelBench: A Benchmark for Evaluating Telco-Specific Large Language Models

EMNLP 2024industry

The telecommunications industry, characterized by its vast customer base and complex service offerings, necessitates a high level of domain expertise and proficiency in customer service center operations. Consequently, there is a growing demand for Large Language Models (LLMs) to augment the capabil…

Cited by 0SourcePDFScholar
2023

What, When, and How to Ground: Designing User Persona-Aware Conversational Agents for Engaging Dialogue

ACL 2023industry

This paper presents a method for building a personalized open-domain dialogue system to address the WWH (WHAT, WHEN, and HOW) problem for natural response generation in a commercial setting, where personalized dialogue responses are heavily interleaved with casual response turns. The proposed approa…

Cited by 12SourcePDFScholar
2022

KoBEST: Korean Balanced Evaluation of Significant Tasks

COLING 2022main

A well-formulated benchmark plays a critical role in spurring advancements in the natural language processing (NLP) field, as it allows objective and precise evaluation of diverse models. As modern language models (LMs) have become more elaborate and sophisticated, more difficult benchmarks that req…