A reinforcement learning algorithm shapes maternal care in mice
Xie, Y.; Huang, L.; Corona, A.; Pagliaro, A. H.; Shea, S. D.
Show abstract
The neural substrates for processing classical rewards such as food or drugs of abuse are well-understood. In contrast, the mechanisms by which organisms perceive social contact as rewarding and subsequently modify their interactions are unclear. Here we tracked the gradual emergence of a repetitive and highly-stereotyped parental behavior and show that trial-by-trial performance correlates with the history of midbrain dopamine (DA) neuron activity. We used a novel behavior paradigm to manipulate the subjects expectation of imminent pup contact and show that DA signals conform to reward prediction error, a fundamental component of reinforcement learning (RL). Finally, closed-loop optogenetic inactivation of DA neurons at the onset of pup contact dramatically slowed emergence of parental care. We conclude that this prosocial behavior is shaped by an RL mechanism in which social contact itself is the primary reward. One-Sentence SummaryMaternal interactions with offspring are shaped by a dopaminergic reinforcement learning mechanism.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Sex differences in neural representations of social and nonsocial reward in the medial prefrontal cortex. 97%
- Distinct spatially organized striatum-wide acetylcholine dynamics for the learning and extinction of Pavlovian associations 97%
- Temporal Dynamics of Nucleus Accumbens Neurons in Male Mice During Reward Seeking 97%
Similar papers in this journal
- The rostral intralaminar nuclear complex of the thalamus supports striatally-mediated action reinforcement 97%
- Acetylcholine modulates prefrontal outcome coding during threat learning under uncertainty 97%
- Mixed representations of choice direction and outcome by GABA/glutamate cotransmitting neurons in the entopeduncular nucleus 97%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.