← Search

Vivian Yvonne Nastl

3 accepted papers

2025

Limits to scalable evaluation at the frontier: LLM as judge won’t beat twice the data

ICLR 2025oral

High quality annotations are increasingly a bottleneck in the explosively growing machine learning ecosystem. Scalable evaluation methods that avoid costly annotation have therefore become an important research ambition. Many hope to use strong existing models in lieu of costly labels to provide che…

Cited by 7SourcePDFScholar