Policy Complexity Suppresses Dopamine Responses

S Samuel J. Gershman (Department of Psychology) A Armin Lak

Abstract

Limits on information processing capacity impose limits on task performance. We show that male and female mice achieve performance on a perceptual decision task that is near-optimal given their capacity limits, as measured by policy complexity (the mutual information between states and actions). This behavioral profile could be achieved by reinforcement learning with a penalty on high complexity policies, realized through modulation of dopaminergic learning signals. In support of this hypothesis, we find that policy complexity suppresses midbrain dopamine responses to reward outcomes. Furthermore, neural and behavioral reward sensitivity were positively correlated across sessions. Our results suggest that policy compression shapes basic mechanisms of reinforcement learning in the brain.

Article Details

Volume / Issue Vol. 45, Issue 9
Published February 26, 2025
Pages e1756242024
ISSN 0270-6474
Publisher Society for Neuroscience

Journal Info

Journal of Neuroscience

Society for Neuroscience

ISSN: 0270-6474 Life Sciences

Authors (2)

S

Samuel J. Gershman

Department of Psychology

A

Armin Lak