2019
Two-Timescale Networks for Nonlinear Value Function Approximation
ICLR 2019poster
A key component for many reinforcement learning agents is to learn a value function, either for policy evaluation or control. Many of the algorithms for learning values, however, are designed for linear function approximation---with a fixed basis or fixed representation. Though there have been a few…