← Search

Ben Rank

2 accepted papers

2026

PostTrainBench: Can LLM Agents Automate LLM Post-Training?

ICML 2026poster

Given the recent rapid progress of LLM agents like Claude Code or Codex CLI for software engineering, an important next question is whether they can automate AI research itself. In this paper, we study *post-training*, which is the critical step that turns base LLMs into useful assistants. We introd…

Cited by 0SourceScholar
2024

Performative Reinforcement Learning in Gradually Shifting Environments

UAI 2024poster

When Reinforcement Learning (RL) agents are deployed in practice, they might impact their environment and change its dynamics. We propose a new framework to model this phenomenon, where the current environment depends on the deployed policy as well as its previous dynamics. This is a generalization…

Cited by 6SourcePDFScholar