2022
Action-modulated midbrain dopamine activity arises from distributed control policies
NeurIPS 2022accept
Animal behavior is driven by multiple brain regions working in parallel with distinct control policies. We present a biologically plausible model of off-policy reinforcement learning in the basal ganglia, which enables learning in such an architecture. The model accounts for action-related modulatio…