2019
Sample-Efficient Deep Reinforcement Learning via Episodic Backward Update
NeurIPS 2019poster
We propose Episodic Backward Update (EBU) – a novel deep reinforcement learning algorithm with a direct value propagation. In contrast to the conventional use of the experience replay with uniform random sampling, our agent samples a whole episode and successively propagates the value of a state to…