Diagnosing Rejection Collapse via Uncertainty Decomposition
Endrizzi, W.; Ragni, F.; Bovo, S.; Moroni, M.; Jurman, G.; Osmani, V.
Show abstract
Standard uncertainty-informed rejection can unexpectedly trigger severe performance collapse, exposing localized vulnerabilities that common machine learning metrics typically do not show. We systematically diagnose this failure dynamic using Levodopa-Induced Dyskinesia prediction in Parkinson's Disease as a proof-of-concept. By training a heterogeneous ML ensemble, decomposing Aleatoric and Epistemic uncertainty and applying unsupervised subgroup discovery, we isolated the precise drivers of these atypical errors. Stratified error analysis revealed two divergent predictive regimes previously hidden by a global evaluation. While the models successfully extracted a predictive signal for one subgroup, the baseline features of a second subgroup lacked discriminative capacity, resulting in a high rate of confident misclassifications. Operating entirely below rejection thresholds, this single subgroup flatlined predictive metrics, driving the collapse of the global rejection curve. Ultimately, we demonstrate that atypical rejection failures stem from subgroup-specific data ambiguity rather than algorithmic deficiencies, making localized uncertainty-aware evaluation a critical methodological requirement prior to real-world deployment.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Generation of realistic synthetic data using multimodal neural ordinary differential equations 93%
- Quantifying Device Type and Handedness Biases in a Remote Parkinson’s Disease AI-Powered Assessment 93%
- Crowdsourcing digital health measures to predict Parkinson's disease severity: the Parkinson's Disease Digital Biomarker DREAM Challenge 93%
Similar papers in this journal
- Generalising uncertainty improves accuracy and safety of deep learning analytics applied to oncology 94%
- Mitigating Machine Learning Bias Between High Income and Low-Middle Income Countries for Enhanced Model Fairness and Generalizability 94%
- Predicting bloodstream infection outcome using machine learning 93%
Similar papers in this journal
- Optimal policy determination in sequential systemic and locoregional therapy of oropharyngeal squamous carcinomas: A patient-physician digital twin dyad with deep Q-learning for treatment selection 93%
- Empirical Sample Size Determination for Popular Classification Algorithms in Clinical Research 93%
- An Interpretable Machine Learning Framework for Accurate Severe vs Non-severe COVID-19 Clinical Type Classification 93%
Similar papers in this journal
- A deep learning model for clinical outcome prediction using longitudinal inpatient electronic health records 93%
- Characterizing subgroup performance of probabilistic phenotype algorithms within older adults: A case study for dementia, mild cognitive impairment, and Alzheimer’s and Parkinson’s diseases 93%
- Modeling physician variability to prioritize relevant medical record information 92%
Similar papers in this journal
- Addressing Label Noise for Electronic Health Records: Insights from Computer Vision for Tabular Data 94%
- OASIS+: leveraging machine learning to improve the prognostic accuracy of OASIS severity score for predicting in-hospital mortality 92%
- Confidence-based laboratory test reduction recommendation algorithm 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.