← Search

Chris Dann

3 accepted papers

2022

Guarantees for Epsilon-Greedy Reinforcement Learning with Function Approximation

ICML 2022spotlight

Myopic exploration policies such as epsilon-greedy, softmax, or Gaussian noise fail to explore efficiently in some reinforcement learning tasks and yet, they perform well in many others. In fact, in practice, they are often selected as the top choices, due to their simplicity. But, for what tasks do…

Cited by 81SourcePDFScholar