← Search

Patrick Schnell

3 accepted papers

2025

Temporal Difference Learning: Why It Can Be Fast and How It Will Be Faster

ICLR 2025poster

Temporal difference (TD) learning represents a fascinating paradox: It is the prime example of a divergent algorithm that has not vanished after its instability was proven. On the contrary, TD continues to thrive in reinforcement learning (RL), suggesting that it provides significant compensatory be…

Cited by 0SourcePDFScholar