← Search

Qiyao Wei

6 accepted papers

2026

Position: The AI Imperative: Scaling High-Quality Peer Review in Machine Learning

ICML 2026oral

Peer review, the bedrock of scientific advancement in machine learning (ML), is strained by a crisis of scale. Exponential growth in manuscript submissions to premier ML venues such as NeurIPS, ICML, and ICLR is outpacing the finite capacity of qualified reviewers, leading to concerns about review q…

Cited by 0SourceScholar
2025

Semantic-KG: Using Knowledge Graphs to Construct Benchmarks for Measuring Semantic Similarity

NeurIPS 2025poster

Evaluating the open-form textual responses generated by Large Language Models (LLMs) typically requires measuring the semantic similarity of the response to a (human generated) reference. However, there is evidence that current semantic similarity methods may capture syntactic or lexical forms over…

Cited by 0SourcecodeScholar
2025

Statistical Hypothesis Testing for Auditing Robustness in Language Models

ICML 2025poster

Consider the problem of testing whether the outputs of a large language model (LLM) system change under an arbitrary intervention, such as an input perturbation or changing the model variant. We cannot simply compare two LLM outputs since they might differ due to the stochastic nature of the system,…

Cited by 0SourcePDFScholar
2024

Defining Expertise: Applications to Treatment Effect Estimation

ICLR 2024poster

Decision-makers are often experts of their domain and take actions based on their domain knowledge. Doctors, for instance, may prescribe treatments by predicting the likely outcome of each available treatment. Actions of an expert thus naturally encode part of their domain knowledge, and can help ma…

Cited by 1SourcePDFScholar