2026
ARLArena: Demystifying Policy Gradient Stability in Agentic Reinforcement Learning
ICML 2026poster
Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interactive tasks. In this paper, we first propose $\textbf{ARLArena}$, a fair and systematic analysis framework that encompasses a broad spectrum of ARL algorit…