2017
Learning to Play in a Day: Faster Deep Reinforcement Learning by Optimality Tightening
ICLR 2017poster
We propose a novel training algorithm for reinforcement learning which combines the strength of deep Q-learning with a constrained optimization approach to tighten optimality and encourage faster reward propagation. Our novel technique makes deep reinforcement learning more practical by drastically…