2025
Optimal Dynamic Regret by Transformers for Non-Stationary Reinforcement Learning
NeurIPS 2025poster
Transformers have demonstrated exceptional performance across a wide range of domains. While their ability to perform reinforcement learning in-context has been established both theoretically and empirically, their behavior in non-stationary environments remains less understood. In this study, we ad…