← Search

Ajin Joseph

1 accepted papers

2019

Two-Timescale Networks for Nonlinear Value Function Approximation

ICLR 2019poster

A key component for many reinforcement learning agents is to learn a value function, either for policy evaluation or control. Many of the algorithms for learning values, however, are designed for linear function approximation---with a fixed basis or fixed representation. Though there have been a few…

Cited by 57SourcePDFScholar