← Search

Stepan Martyanov

1 accepted papers

2022

Continuous Deep Q-Learning in Optimal Control Problems: Normalized Advantage Functions Analysis

NeurIPS 2022accept

One of the most effective continuous deep reinforcement learning algorithms is normalized advantage functions (NAF). The main idea of NAF consists in the approximation of the Q-function by functions quadratic with respect to the action variable. This idea allows to apply the algorithm to continuous…

Cited by 1SourcePDFScholar