2025
Temporal Difference Learning: Why It Can Be Fast and How It Will Be Faster
ICLR 2025poster
Temporal difference (TD) learning represents a fascinating paradox: It is the prime example of a divergent algorithm that has not vanished after its instability was proven. On the contrary, TD continues to thrive in reinforcement learning (RL), suggesting that it provides significant compensatory be…