Why Risk Factors Minimally Change the ROC Curve AUC
Stern, R. H.
Show abstract
Novel risk factors that improve statistical measures of fit on addition to established clinical prediction models often minimally change measures of discrimination (ROC curve AUC or c-index). As a result, measures of discrimination have been suggested to be insensitive in the evaluation of such models. To understand this phenomenon, it is necessary to focus on the population risk distributions produced by models with and without the risk factor. This is because these risk distribution fully determines the risk distributions of cases/patients and controls/nonpatients, which in turn fully determine the ROC curve and its AUC. Broader population risk distributions result in larger ROC curve AUCs. A fully independent risk factor with a relative risk of 2 added to the standard cardiovascular risk model produces risk distributions of those with or without the risk that are clearly different (which is evaluated by statistical measures of fit), while minimally broadening the population risk distributions (which is evaluated by measures of discrimination). The reason for this is that although addition of the risk factor replaces every risk stratum with higher and lower risk strata, this depopulated risk stratum is largely repopulated by similar splitting in neighboring risk strata. The interweaving of the the up and down migration paths to and from every point on the risk distribution results in a largely compensatory shuffling of risk assignments with minimal changes in the ROC curve AUC.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Time-to-event estimation of birth year prevalence trends: a method to enable investigating the etiology of childhood disorders including autism 94%
- ChatGPT Provides Inconsistent Risk-Stratification of Patients With Atraumatic Chest Pain 93%
- Introducing riskCommunicator: an R package to obtain interpretable effect estimates for public health 93%
Similar papers in this journal
- Causal attribution of smoking and BMI to the landscape of disease incidence in UK Biobank 93%
- Can machine learning improve risk prediction of incident hypertension? An internal method comparison and external validation of the Framingham risk model using HUNT Study data 91%
- Rapid Clinical Screening and Staging for COVID-19 Severe Outcome A Hospitalization Study in New York City 91%
Similar papers in this journal
- Negative Control Exposures: Causal effect Identifiability and Use in Probabilistic-Bias and Bayesian Analyses with Unmeasured Confounders 91%
- Sensitivity and Uncertainty Analysis for Two-Stream Capture-Recapture Methods in Disease Surveillance 89%
- The Epidemiological Implications of Jails for Community, Corrections Officer, and Incarcerated Population Risks from COVID-19 89%
Similar papers in this journal
- External control arm analysis: an evaluation of propensity score approaches, G-computation, and doubly debiased machine learning 93%
- Using linear and natural cubic splines, SITAR, and latent trajectory models to characterise nonlinear longitudinal growth trajectories in cohort studies 91%
- Comparing methods to predict baseline mortality for excess mortality calculations 91%
Similar papers in this journal
- Nonspecific blood tests as proxies for COVID-19 hospitalization: are there plausible associations after excluding noisy predictors? 93%
- Estimation of case-fatality rate in COVID-19 patients with hypertension and diabetes mellitus in the New York State 92%
- Excess Mortality in the United States During the First Three Months of the COVID-19 Pandemic 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.