← Search

Serena Lutong Wang

1 accepted papers

2025

Metritocracy: Representative Metrics for Lite Benchmarks

NeurIPS 2025poster

A common problem in LLM evaluation is how to choose a subset of metrics from a full suite of possible metrics. Subset selection is usually done for efficiency or interpretability reasons, and the goal is often to select a "representative" subset of metrics. However, "representative" is rarely clearl…

Cited by 0SourceScholar