← Search

Mihail Bîrsan

1 accepted papers

2025

ADDQ: Adaptive distributional double Q-learning

ICML 2025poster

Bias problems in the estimation of Q-values are a well-known obstacle that slows down convergence of Q-learning and actor-critic methods. One of the reasons of the success of modern RL algorithms is partially a direct or indirect overestimation reduction mechanism. We introduce an easy to implement…