2026
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning
ICLR 2026poster
The use of target networks is a popular approach for estimating value functions in deep Reinforcement Learning (RL). While effective, the target network remains a compromise solution that preserves stability at the cost of slowly moving targets, thus delaying learning. Conversely, using the online n…