← Search

Pierre Thodoroff

1 accepted papers

2018

Temporal Regularization for Markov Decision Process

NeurIPS 2018poster

Several applications of Reinforcement Learning suffer from instability due to high variance. This is especially prevalent in high dimensional domains. Regularization is a commonly used technique in machine learning to reduce variance, at the cost of introducing some bias. Most existing regularizatio…