← Search

Chieh-Yang Huang

3 accepted papers

2025

Using Contextually Aligned Online Reviews to Measure LLMs’ Performance Disparities Across Language Varieties

NAACL 2025short

A language can have different varieties. These varieties can affect the performance of natural language processing (NLP) models, including large language models (LLMs), which are often trained on data from widely spoken varieties. This paper introduces a novel and cost-effective approach to benchmar…

2023

GPT-4 as an Effective Zero-Shot Evaluator for Scientific Figure Captions

EMNLP 2023short findings

There is growing interest in systems that generate captions for scientific figures. However, assessing these systems' output poses a significant challenge. Human evaluation requires academic expertise and is costly, while automatic evaluation depends on often low-quality author-written captions. Thi…

Cited by 0SourceScholar