ICLR 2017poster98 citations
Learning to Play in a Day: Faster Deep Reinforcement Learning by Optimality Tightening
Frank S.He, Yang Liu, Alexander G. Schwing, Jian Peng
Abstract
We propose a novel training algorithm for reinforcement learning which combines the strength of deep Q-learning with a constrained optimization approach to tighten optimality and encourage faster reward propagation. Our novel technique makes deep reinforcement learning more practical by drastically reducing the training time. We evaluate the performance of our approach on the 49 games of the challenging Arcade Learning Environment, and report significant improvements in both training time and accuracy.
Reinforcement LearningOptimizationGames
BibTeX
@inproceedings{
s.he2017learning,
title={Learning to Play in a Day: Faster Deep Reinforcement Learning by Optimality Tightening},
author={Frank S.He and Yang Liu and Alexander G. Schwing and Jian Peng},
booktitle={International Conference on Learning Representations},
year={2017},
url={https://openreview.net/forum?id=rJ8Je4clg}
}