2023
Adaptive Estimation Q-learning with Uncertainty and Familiarity
IJCAI 2023poster
One of the key problems in model-free deep reinforcement learning is how to obtain more accurate value estimations. Current most widely-used off-policy algorithms suffer from over- or underestimation bias which may lead to unstable policy. In this paper, we propose a novel method, Adaptive Estimatio…