2024
Towards Optimal Adversarial Robust Q-learning with Bellman Infinity-error
ICML 2024oral
Establishing robust policies is essential to counter attacks or disturbances affecting deep reinforcement learning (DRL) agents. Recent studies explore state-adversarial robustness and suggest the potential lack of an optimal robust policy (ORP), posing challenges in setting strict robustness constr…