Inverse reinforcement learning reveals action-oriented value signals in naturalistic decision making
Lee, S. H.; Chung, C.; Oh, M.-h.; Ahn, W.-Y.
Show abstract
A major challenge for cognitive neuroscience is to explain how value of a goal-directed behavior is computed in complex and naturalistic environments. Standard computational models of decision making have been highly successful in controlled, trial-based paradigms, but they are often ill-suited to real-time behavior unfolding in naturalistic paradigms. Inverse reinforcement learning (IRL) offers a way to infer latent evaluative state from observed behavior in naturalistic environments, but its neural interpretability remains largely unknown. Here, we investigated whether moment-to-moment reward trajectories derived from IRL map onto value signals in the brain during a real-time driving task performed during fMRI scanning. IRL-derived reward trajectories were most robustly associated with activity in the dorsal striatum, a region often linked to value-guided action selection. They also showed associations with distributed regions supporting additional processes, including cognitive control and sensorimotor processing. This pattern suggests that IRL reward captures distributed neural activity centered on the reward circuitry, potentially reflecting how valuation interacts with other processes. Together, these findings suggest that IRL reward provides a behaviorally grounded, temporally resolved proxy for action-oriented valuation during naturalistic decision making.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Neural signatures of attentional engagement during narratives and its consequences for event memory 95%
- Striatal and cerebellar interactions during reward-based motor performance 95%
- Beta and theta oscillations track effort and previous reward in human basal ganglia and prefrontal cortex during decision making 95%
Similar papers in this journal
- Neuro-computational account of arbitration between imitation and emulation during human observational learning 97%
- Human hippocampus and dorsomedial prefrontal cortex infer and update latent causes during social interaction 97%
- Joint representation of working memory and uncertainty in human cortex 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.