← Search

Xiaoyu Gong

1 accepted papers

2023

Adaptive Estimation Q-learning with Uncertainty and Familiarity

IJCAI 2023poster

One of the key problems in model-free deep reinforcement learning is how to obtain more accurate value estimations. Current most widely-used off-policy algorithms suffer from over- or underestimation bias which may lead to unstable policy. In this paper, we propose a novel method, Adaptive Estimatio…