2026
Bridging the performance-gap between target-free and target-based reinforcement learning
ICLR 2026poster
The use of target networks in deep reinforcement learning is a widely popular solution to mitigate the brittleness of semi-gradient approaches and stabilize learning. However, target networks notoriously require additional memory and delay the propagation of Bellman updates compared to an ideal targ…