Back

Generalizability of Risk Models for Treatment-Resistant Depression Across Three Health Systems

Walsh, C. G.; Ripperger, M.; McCoy, T. H.; Castro, V.; Hu, Y.; Kirchner, L.; Ruderfer, D. M.; Perlis, R. H.

2025-05-27 psychiatry and clinical psychology
10.1101/2025.05.21.25328089 medRxiv
Show abstract

BackgroundAs multiple strategies have emerged for managing treatment-resistant major depressive disorder, efficient identification of individuals at elevated risk for this outcome earlier in their illness course remains essential. MethodWe extracted electronic health records data for all individuals with a diagnosis of major depressive disorder who received an index antidepressant prescription in the clinical networks of three geographically-distinct health systems - Mass General-Brigham (MGB), Vanderbilt University Medical Center (VUMC), and Geisinger Clinic (GC) - between April 1, 2004, and March 30, 2022. The primary outcome, treatment resistant depression, was defined as provision of electroconvulsive therapy, transcranial magnetic stimulation, vagus nerve stimulation, prescription of either ketamine or esketamine or monoamine oxidase inhibitors (MAOIs), or failed trials of more than two antidepressants. We applied L1-regularized regression to sociodemographic features, medications, and ICD10 diagnostic code counts to fit a model of treatment resistance in each of the three cohorts. For each, we then estimated generalizable model performance, aka external validity, across the other two cohorts. Model concordance was measured with Concordance Correlation Coefficients (CCCs) and random forest regression analyses were used to estimate importance of features predicting discordance. ResultsAcross sites, discrimination performance ranged from Area Under the Receiver Operating Characteristic curves (AUROCs) 0.58 - 0.64 on internal validation and 0.51 - 0.58 on external validation. Area Under the Precision-Recall curve (AUPRC) ranged from 0.1-0.13 on internal validation and averaged 0.07-0.13 in external validation on the same test sets held out at each site. On the same testing set, CCCs were 0.13 for the VUMC<-> MGB models, 0.18 for VUMC<->GC models, and 0.38 for MGB<-> GC models. These results indicate the MGB and GC models were better correlated, but none were well correlated. Important features predicting discordance were dominated primarily by age and secondarily coded sex. ConclusionThese linear models demonstrated consistent aggregate performance and discordant individual performance across three, disparate major health systems. The inclusion of large and heterogeneous samples suggest that further improvement may require incorporation of data types beyond those readily available in EHR. Close attention to performance by key subgroups is indicated to ensure models do not perform disparately or unfairly. Prospective studies to evaluate the extent to which clinical models might improve early identification and outcomes are warranted.

Matching journals

The top 11 journals account for 50% of the predicted probability mass.

1
PLOS ONE
5266 papers in training set
Top 25%
6.6%
2
BMJ Mental Health
15 papers in training set
Top 0.1%
6.6%
3
Psychological Medicine
88 papers in training set
Top 0.4%
5.3%
4
Acta Psychiatrica Scandinavica
10 papers in training set
Top 0.1%
5.3%
5
Journal of Affective Disorders
92 papers in training set
Top 0.5%
5.0%
6
American Journal of Psychiatry
24 papers in training set
Top 0.1%
4.7%
7
npj Digital Medicine
118 papers in training set
Top 1%
4.3%
8
JAMA Network Open
130 papers in training set
Top 0.8%
4.2%
9
Biological Psychiatry: Cognitive Neuroscience and Neuroimaging
71 papers in training set
Top 0.5%
3.3%
10
BMJ Open
601 papers in training set
Top 6%
3.3%
11
Translational Psychiatry
260 papers in training set
Top 2%
3.2%
50% of probability mass above
12
The British Journal of Psychiatry
23 papers in training set
Top 0.2%
3.1%
13
JAMA Psychiatry
15 papers in training set
Top 0.1%
3.0%
14
BMC Psychiatry
25 papers in training set
Top 0.4%
2.3%
15
BJPsych Open
29 papers in training set
Top 0.3%
2.3%
16
JAMIA Open
42 papers in training set
Top 0.8%
2.1%
17
PLOS Medicine
110 papers in training set
Top 2%
1.9%
18
Psychiatry Research
41 papers in training set
Top 0.7%
1.7%
19
Frontiers in Psychiatry
87 papers in training set
Top 1%
1.7%
20
JMIR Formative Research
33 papers in training set
Top 0.8%
1.6%
21
Frontiers in Digital Health
24 papers in training set
Top 0.9%
1.5%
22
European Psychiatry
11 papers in training set
Top 0.2%
1.4%
23
Epidemiology and Psychiatric Sciences
11 papers in training set
Top 0.2%
1.3%
24
Schizophrenia Bulletin
32 papers in training set
Top 0.3%
1.1%
25
BMC Health Services Research
51 papers in training set
Top 2%
1.1%
26
Biological Psychiatry
137 papers in training set
Top 2%
1.1%
27
Acta Neuropsychiatrica
14 papers in training set
Top 0.5%
1.0%
28
Computational Psychiatry
12 papers in training set
Top 0.1%
1.0%
29
JMIRx Med
32 papers in training set
Top 2%
1.0%
30
Communications Medicine
113 papers in training set
Top 4%
1.0%