← Search

Alexandre Pouget

2 accepted papers

2020

A Local Temporal Difference Code for Distributional Reinforcement Learning

NeurIPS 2020poster

Recent theoretical and experimental results suggest that the dopamine system implements distributional temporal difference backups, allowing learning of the entire distributions of the long-run values of states rather than just their expected values. However, the distributional codes explored so far…

Cited by 37SourcePDFScholar
2020

Dynamic allocation of limited memory resources in reinforcement learning

NeurIPS 2020poster

Biological brains are inherently limited in their capacity to process and store information, but are nevertheless capable of solving complex tasks with apparent ease. Intelligent behavior is related to these limitations, since resource constraints drive the need to generalize and assign importance d…