← Search

Stanley Wei

6 accepted papers

2026

Improved high-dimensional estimation with Langevin dynamics and stochastic weight averaging

ICLR 2026poster

Significant recent work has studied the ability of gradient descent to recover a hidden planted direction $\theta^\star \in S^{d-1}$ in different high-dimensional settings, including tensor PCA and single-index models. The key quantity that governs the ability of gradient descent to traverse these l…

Cited by 0SourceScholar
2025

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

NeurIPS 2025poster

Recent reports claim that large language models (LLMs) now outperform elite humans in competitive programming. Drawing on knowledge from a group of medalists in international algorithmic contests, we revisit this claim, examining how LLMs differ from human experts and where limitations still remain.…

Cited by 0SourceScholar
2025

Provable unlearning in topic modeling and downstream tasks

ICLR 2025poster

Machine unlearning algorithms are increasingly important as legal concerns arise around the provenance of training data, but verifying the success of unlearning is often difficult. Provable guarantees for unlearning are often limited to supervised learning settings. In this paper, we provide the fir…

Cited by 1SourcePDFScholar
2025

What Makes a Reward Model a Good Teacher? An Optimization Perspective

NeurIPS 2025spotlight

The success of Reinforcement Learning from Human Feedback (RLHF) critically depends on the quality of the reward model. However, while this quality is primarily evaluated through accuracy, it remains unclear whether accuracy fully captures what makes a reward model an effective teacher. We address t…

Cited by 0SourcecodeScholar
2024

Transformers Provably Learn Sparse Token Selection While Fully-Connected Nets Cannot

ICML 2024poster

The transformer architecture has prevailed in various deep learning settings due to its exceptional capabilities to select and compose structural information. Motivated by these capabilities, Sanford et al. (2023) proposed the *sparse token selection* task, in which transformers excel while fully-co…

Cited by 14SourcePDFScholar