2018
Regret Minimization for Partially Observable Deep Reinforcement Learning
ICLR 2018workshop
Deep reinforcement learning algorithms that estimate state and state-action value functions have been shown to be effective in a variety of challenging domains, including learning control strategies from raw image pixels. However, algorithms that estimate state and state-action value functions typic…