Back

Selective prediction as a triage gate for primary-care depression screening: quantifying and mitigating selection bias in CHARLS-2011

Wang, Z.; liu, y.

2026-07-20 health informatics
10.64898/2026.07.17.26357845 medRxiv
Show abstract

Background Primary care in China lacks structured mental-health assessment, and the machine-learning models that could support such screening are typically developed on heavily selected samples. Cumulative inclusion and exclusion criteria, though usually treated as neutral data-cleaning steps, can create heterogeneity in predictive reliability among retained participants. Using the China Health and Retirement Longitudinal Study (CHARLS) 2011 baseline, we quantified how selection funnels distort epidemiological associations and inflate machine-learning metrics, and tested selective prediction as mitigation. Methods Using the CHARLS 2011 baseline with temporal external validation in CHARLS-2018, we built a four-level selection funnel (L0-L3), evaluated five classifiers with nested cross-validation and SMOTE, and compared model-embedded uncertainty with a decoupled predictor-selector framework; XGBoost cross-validation residuals drove risk stratification and classification and regression tree (CART) rules. Results Sample sizes fell from L0 n=17,705 to L3 n=4,256 (24.0%). The cancer-depression odds ratio attenuated from 1.78 (95% CI 1.32-2.41) to 1.39 (0.74-2.63), losing significance. AUC rose with selection but not after multiple-comparison correction, whereas calibration error increased for four of five models. Model-embedded uncertainty succeeded only for XGBoost; with the decoupled XGBoost residual selector, all five models achieved selective prediction at approximately 20% coverage (test AUC 0.90, 95% CI 0.85-0.95), abstaining on approximately 80% of cases for individual safety. Risk stratification was stable (residual Spearman correlations >0.95; multi-seed Jaccard 0.88), and CART rules used self-rated health, education, pain, and marital status. Conclusions The findings support a deployable primary-care triage pathway: a four-variable rule identifies patients suitable for algorithm-assisted scoring (approximately 20% coverage) and routes the remainder to human evaluation. Methodologically, cumulative selection bias produces a dual distortion: epidemiological associations are compressed and machine-learning metrics inflated. Selective prediction is limited mainly by uncertainty-indicator design. Performance metrics should be reported with selection level, coverage, and calibration trajectory. Decoupled selective prediction with CART rule extraction provides an actionable framework for quality-controlled, tiered-care deployment. Keywords: selective prediction, selection bias, CHARLS, depression, predictor-selector decoupling, uncertainty quantification, classification and regression tree, triage, clinical decision support, health management.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

1
npj Digital Medicine
118 papers in training set
Top 0.4%
14.8%
2
Journal of the American Medical Informatics Association
71 papers in training set
Top 0.2%
14.6%
3
BMC Medical Research Methodology
47 papers in training set
Top 0.1%
9.6%
4
Communications Medicine
113 papers in training set
Top 0.6%
4.8%
5
PLOS ONE
5266 papers in training set
Top 33%
4.3%
6
Journal of Biomedical Informatics
47 papers in training set
Top 0.4%
4.0%
50% of probability mass above
7
BMJ Open
601 papers in training set
Top 7%
3.2%
8
PLOS Digital Health
106 papers in training set
Top 2%
3.2%
9
Journal of Clinical Epidemiology
31 papers in training set
Top 0.3%
2.6%
10
BMC Medical Informatics and Decision Making
43 papers in training set
Top 0.7%
2.6%
11
The Lancet Digital Health
25 papers in training set
Top 0.2%
2.6%
12
International Journal of Medical Informatics
26 papers in training set
Top 0.5%
2.4%
13
JMIR Medical Informatics
18 papers in training set
Top 0.4%
2.1%
14
JAMA Network Open
130 papers in training set
Top 2%
2.1%
15
BMJ
51 papers in training set
Top 0.4%
1.9%
16
Frontiers in Digital Health
24 papers in training set
Top 0.9%
1.4%
17
JMIR Public Health and Surveillance
45 papers in training set
Top 1%
1.1%
18
Scientific Reports
3612 papers in training set
Top 66%
1.1%
19
Journal of Medical Internet Research
87 papers in training set
Top 2%
1.0%
20
Medical Decision Making
12 papers in training set
Top 0.3%
1.0%
21
eClinicalMedicine
77 papers in training set
Top 2%
0.9%
22
JAMIA Open
42 papers in training set
Top 2%
0.8%
23
Nature Communications
5641 papers in training set
Top 57%
0.8%
24
DIGITAL HEALTH
17 papers in training set
Top 0.9%
0.8%
25
Emergency Medicine Journal
21 papers in training set
Top 0.5%
0.6%
26
BMJ Health & Care Informatics
15 papers in training set
Top 1%
0.6%
27
BMJ Public Health
25 papers in training set
Top 2%
0.6%
28
BJPsych Open
29 papers in training set
Top 0.9%
0.6%
29
Nature Medicine
125 papers in training set
Top 4%
0.6%