2026
Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy
ICML 2026poster
One goal in reinforcement learning (RL) research is to understand general purpose sequential decision-making, using benchmark simulators as a proxy for learning in a deployment setting. When running experiments, however, the goal of achieving high performance in the simulator can mutate into focusin…