Electrophysiological Correlates of Reinforcement Learning in the Human Ventral Tegmental Area
Ramaswamy, A.; Steele, D.; Roiser, J. P.; Simmonds, L.; Lagrata, S.; Matharu, M. S.; Vivekananda, U.; Akram, H.; Zrinzo, L.; Litvak, V.
Show abstract
The ventral tegmental area is the primary source of dopaminergic input to the human prefrontal cortex and plays a central role in reinforcement learning. Although animal studies have established that dopaminergic neurons encode reward prediction error signals, direct electrophysiological evidence in humans is scarce. Understanding these mechanisms is clinically relevant because of their involvement in disorders of motivation and reward processing. In this cross-sectional study, we recorded local field potentials from the ventral tegmental area in fourteen patients (nine male; mean age 46 years, range 30-62) undergoing deep brain stimulation surgery for chronic cluster headache. During temporary electrode externalisation, participants performed a probabilistic instrumental learning task comprising reward, loss and neutral trials. Behaviour was modelled using a hierarchical Rescorla-Wagner framework with separate learning rates for rewards and losses. Electrophysiological responses were analysed using Statistical Parametric Mapping and linear mixed-effects models testing sensitivity to outcome, expected value and reward prediction error. Clear evoked responses were observed for stimulus and outcome events in thirteen subjects. Responses reflecting the contrast between outcomes that delivered a gain, a loss or a neutral signal and those delivering no outcome were significantly larger for gains than for losses or neutral signals (paired t-tests: gain versus loss, t(12)=2.60, p=0.023, Cohens d=0.72; gain versus neutral, t(12)=4.49, p<0.001, d=1.25), while loss and neutral contrasts did not differ (p=0.218). Nine of fourteen subjects developed a clear preference for the high-reward option; two exhibited gradual learning, while others adopted a win-stay strategy. In eight subjects with clear evoked responses, activity around the button press correlated with the expected value of the chosen option (peak at 0.02s, p=0.013, corrected). In a further subset of two subjects who explored the low-value option, single-trial responses peaked closer to the button press during low-value choices. No compelling evidence was found for a distinct reward prediction error signal beyond outcome and value. Clinical covariates and smoking status did not significantly modulate electrophysiological responses. Human ventral tegmental area activity is selectively tuned to rewarding outcomes and, under conditions of effective learning, reflects the expected value of chosen options during decision-making. These findings align with reinforcement learning principles established in animal models and suggest that local field potentials primarily represent inputs to reward prediction error computation rather than its output. This work provides the first direct electrophysiological evidence of reward-related signalling in the human ventral tegmental area, supporting its translational relevance for understanding motivation and guiding neuromodulatory interventions.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Timing along the cardiac cycle modulates neural signals of reward-based learning 96%
- Lateral Prefrontal Theta Oscillations Causally Drive a Computational Mechanism Underlying Conflict Expectation and Adaptation 96%
- Electrophysiological population dynamics reveal context dependencies during decision making in human frontal cortex 95%
Similar papers in this journal
- Cortical sensorimotor activity in the execution and suppression of discrete and rhythmic movements 96%
- Decomposing Simon task BOLD activation using a drift-diffusion model framework 95%
- Bayesian Prior Uncertainty and Surprisal Elicit Distinct Neural Patterns During Sound Localization in Dynamic Environments 95%
Similar papers in this journal
- The role of alpha oscillations in a premotor-cerebellar loop in modulation of motor learning: insights from transcranial alternating current stimulation 96%
- tDCS modulates effective connectivity during motor command following; a potential therapeutic target for disorders of consciousness 95%
- The neural activity of auditory conscious perception 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.