Overlooked features lead to divergent neurobiological interpretations of brain-based machine learning biomarkers
Adkinson, B. D.; Rosenblatt, M.; Sun, H.; Dadashkarimi, J.; Tejavibulya, L.; Horien, C.; Westwater, M. L.; Noble, S.; Scheinost, D.
Show abstract
A central objective in human neuroimaging is to understand the neurobiology underlying cognition and mental health. Machine learning models trained on brain connectivity data are increasingly used as tools for predicting behavioral phenotypes 1,2, enhancing precision medicine 3,4, and improving generalizability compared to traditional MRI studies 5. However, the high dimensionality of brain connectivity data makes model interpretation challenging 6. Prevailing practices within the field rely on sparsely selected brain connectivity features, implicitly interpreting identified feature networks as uniquely representative of a given phenotype while overlooking others. Here, we show that commonly overlooked brain connectivity features can achieve similar prediction accuracies while yielding markedly different neurobiological interpretations. Using four large-scale neuroimaging datasets spanning over 12,000 participants and 13 outcomes, we demonstrate that this phenomenon is widespread across cognitive, developmental, and psychiatric phenotypes. It extends to both functional connectivity (fMRI) and structural (DTI) connectomes and remains evident even in external validation. These findings suggest that common practices may lead to feature overinterpretation and a misrepresentation of the neurobiological bases of brain-behavior associations. Such interpretations present only the "tip of the iceberg" when certain disregarded features may be just as meaningful, potentially contributing to ongoing issues surrounding reproducibility within the field. More broadly, our results point to the possibility that multiple neurobiologically distinct models may exist for the same phenotype, with implications for identifying meaningful subtypes within clinical and research populations.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Reconfigurations of cortical manifold structure during reward-based motor learning 96%
- No effect of additional education on long-term brain structure: a preregistered natural experiment in thousands of individuals. 96%
- Neurotransmitter Transporter/Receptor Co-Expression Shares Organizational Traits With Brain Structure And Function 96%
Similar papers in this journal
- White matter connections of human ventral temporal cortex are organized by cytoarchitecture, eccentricity, and category-selectivity from birth 95%
- Higher rostral locus coeruleus integrity is associated with better memory performance in older adults 94%
- The genetic architecture of structural left-right asymmetry of the human brain 94%
Similar papers in this journal
- Evaluating the sensitivity of functional connectivity measures to motion artifact in resting-state fMRI data 96%
- Functional brain network community structure in childhood: Unfinished territories and fuzzy boundaries 95%
- Disentangling cortical functional connectivity strength and topography reveals divergent roles of genes and environment 95%
Similar papers in this journal
- Functional Parcellation of the Default Mode Network: A Large-Scale Meta-Analysis 96%
- Experience sampling reveals the role that covert goal states play in task-relevant behavior. 95%
- No evidence for a relationship between social closeness and similarity in resting-state functional brain connectivity in schoolchildren 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.