← Search

Rich Sutton

2 accepted papers

2026

Intentional Updates for Streaming Reinforcement Learning

ICML 2026poster

In gradient-based learning, a step size chosen in parameter units does not produce a predictable per-step change in the function output. This may lead to instability in the streaming setting (i.e., batch size=1), where stochasticity is not averaged out and update magnitudes can momentarily become ar…

Cited by 0SourceScholar