← Search

Gen Suzuki

1 accepted papers

2025

Aligning Black-box Language Models with Human Judgments

NAACL 2025findings

Large language models (LLMs) are increasingly used as automated judges to evaluate recommendation systems, search engines, and other subjective tasks, where relying on human evaluators can be costly, time-consuming, and unscalable. LLMs offer an efficient solution for continuous, automated evaluatio…

Cited by 0SourcePDFScholar