← Search

Eric R Chen

1 accepted papers

2022

Redeeming intrinsic rewards via constrained optimization

NeurIPS 2022accept

State-of-the-art reinforcement learning (RL) algorithms typically use random sampling (e.g., $\epsilon$-greedy) for exploration, but this method fails on hard exploration tasks like Montezuma's Revenge. To address the challenge of exploration, prior works incentivize exploration by rewarding the age…