Back

Brain-AI Alignment in Naturalistic Movies

Li, M.

2025-12-08 neuroscience
10.64898/2025.12.03.692164 bioRxiv
Show abstract

Naturalistic paradigms offer a powerful window into human cognition, but it remains difficult to link rich, continuous movie content to distributed brain activity in an interpretable way. In this study, I use a multimodal large language model (Gemini) as an automated "semantic annotator" to bridge naturalistic movie stimuli, brain responses, and cognitive performance. Using the Human Connectome Project movie-watching dataset, I segmented the film into 293 overlapping clips, prompted Gemini to rate each clip on 11 psychologically interpretable dimensions, and simultaneously extracted clip-wise BOLD activation patterns from the fMRI images in 360 cortical parcels. In this way, the AI and the brain effectively "watch" the same movies in parallel. For each parcel, I then fit linear regression models to predict clip-to-clip variation in movie-evoked responses from these features. Gemini-derived features robustly predicted movie-evoked responses in temporal, medial parietal, and lateral frontal association cortex, but explained little variance in unimodal somatosensory, dorsal parietal, insular, and piriform regions. Feature-weight maps recapitulated known functional specializations, and features with the largest global influence overlapped with the most explainable parcels. Partial least squares analysis revealed that individual differences in resting-state connectivity strength and semantic explainability covaried along an asymmetric intrinsic axis: strongly integrated sensory-opercular systems at rest were associated with poorer AI predictability, whereas a smaller set of dorsal and medial association regions showed enhanced alignment. Finally, regional AI explainability in medial parietal and left perisylvian association areas was positively related to fluid and crystallized cognitive abilities. Together, these findings demonstrate that prompt-defined, interpretable features from foundation models provide a simple and scalable framework for quantifying brain-AI alignment in naturalistic settings, offering a practical bridge between biological and artificial semantic representations.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.