Brain–AI Alignment in Naturalistic Movies
preprint
OA: closed
Abstract
ABSTRACT Naturalistic paradigms offer a powerful window into human cognition, but it remains difficult to link rich, continuous movie content to distributed brain activity in an interpretable way. In this study, I use a multimodal large language model (Gemini) as an automated “semantic annotator” to bridge naturalistic movie stimuli, brain responses, and cognitive performance. Using the Human Connectome Project movie-watching dataset, I segmented the film into 293 overlapping clips, prompted Gemini to rate each clip on 11 psychologically interpretable dimensions, and simultaneously extracted clip-wise BOLD activation patterns from the fMRI images in 360 cortical parcels. In this way, the AI and the brain effectively “watch” the same movies in parallel. For each parcel, I then fit linear regression models to predict clip-to-clip variation in movie-evoked responses from these features. Gemini-derived features robustly predicted movie-evoked responses in temporal, medial parietal, and lateral frontal association cortex, but explained little variance in unimodal somatosensory, dorsal parietal, insular, and piriform regions. Feature-weight maps recapitulated known functional specializations, and features with the largest global influence overlapped with the most explainable parcels. Partial least squares analysis revealed that individual differences in resting-state connectivity strength and semantic explainability covaried along an asymmetric intrinsic axis: strongly integrated sensory–opercular systems at rest were associated with poorer AI predictability, whereas a smaller set of dorsal and medial association regions showed enhanced alignment. Finally, regional AI explainability in medial parietal and left perisylvian association areas was positively related to fluid and crystallized cognitive abilities. Together, these findings demonstrate that prompt-defined, interpretable features from foundation models provide a simple and scalable framework for quantifying brain–AI alignment in naturalistic settings, offering a practical bridge between biological and artificial semantic representations.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2025) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.
Source provenance
- europepmc
- last seen: 2026-05-20T01:45:00.602351+00:00