bioRxiv · 10.1101/2024.09.15.613150
Policy complexity suppresses dopamine responses
Abstract
Limits on information processing capacity impose limits on task performance. We show that animals achieve performance on a perceptual decision task that is near-optimal given their capacity limits, as measured by policy complexity (the mutual information between states and actions). This behavioral profile could be achieved by reinforcement learning with a penalty on high complexity policies, realized through modulation of dopaminergic learning signals. In support of this hypothesis, we find that policy complexity suppresses midbrain dopamine responses to reward outcomes, thereby reducing behavioral sensitivity to these outcomes. Our results suggest that policy compression shapes basic mechanisms of reinforcement learning in the brain.
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Gershman, S. J., Lak, A.. 2024-09-16. Policy complexity suppresses dopamine responses. https://doi.org/10.1101/2024.09.15.613150
Cite the original work for its findings. Save a collection to share your selection of sources.