← Search

Yaowen Ye

2 accepted papers

2026

Auditing Black-Box LLM APIs with a Rank-Based Uniformity Test

ICLR 2026poster

As API access becomes a primary interface to large language models (LLMs), users often interact with black-box systems that offer little transparency into the deployed model. To reduce costs or maliciously alter model behaviors, API providers may discreetly serve quantized or fine-tuned variants, wh…

Cited by 0SourcecodeScholar
2025

Iterative Label Refinement Matters More than Preference Optimization under Weak Supervision

ICLR 2025spotlight

Language model (LM) post-training relies on two stages of human supervision: task demonstrations for supervised finetuning (SFT), followed by preference comparisons for reinforcement learning from human feedback (RLHF). As LMs become more capable, the tasks they are given become harder to supervise.…