When clinical prediction models do not generalize: a simulation study in liver transplantation
Brulhart, D.; Magini, G.; Schafer, A.; Schwab, S.; Held, U.
Show abstract
Objectives: Clinical prediction models estimate the risk of a future outcome in patients. Such models are often externally validated using independent datasets; however, even when a model has been rigorously validated in a new setting and patient population, its performance across other clinical settings remains unclear. Therefore, we systematically evaluated model performance and clinical utility across diverse patient populations to quantify the limits of transportability. Methods: Using liver transplantation as an example, we used the UK donation-after-circulatory-death (DCD) risk score and descriptive statistics from Swiss DCD liver transplant populations to simulate realistic target populations with varying donor and recipient characteristics. The risk score's ability to predict one-year graft failure was evaluated using calibration intercept, calibration slope, area under the receiver operating characteristic (ROC) curve, and net benefit. Results: The UK DCD Risk Score's performance depended heavily on the simulated population characteristics. While the score performed adequately in settings similar to those where it was derived, it was not satisfactory in others. Discussion: The study showed, using a risk score in liver transplantation as an example, that the application of a prediction model can be limited in certain external populations when they differ, and that its transportability in new settings is not guaranteed. Conclusion: This study highlights the importance of external validation of clinical prediction models to determine transportability to various target populations. Their application requires careful consideration and potential model re-estimation.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Trends in underlying causes of death in solid organ transplant recipients between 2010 and 2020: Using the CLASS method for determining specific causes of death 93%
- Return to work outcomes in solid organ transplant recipients: a protocol for a global scoping review 93%
- Patient Perceptions by Race of Educational Animations About Living Kidney Donation Made for a Diverse Population 92%
Similar papers in this journal
- Impact of porcine cytomegalovirus on long-term orthotopic cardiac xenotransplant survival 92%
- Application of physiological network mapping in the prediction of survival in critically ill patients with acute liver failure 91%
- Randomized controlled trial of convalescent plasma therapy against standard therapy in patients with severe COVID-19 disease 91%
Similar papers in this journal
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 91%
- Development and Validation of ‘Patient Optimizer’ (POP) Algorithms for Predicting Surgical Risk with Machine Learning 91%
- Optimized Feature Selection and Advanced Machine Learning for Stroke Risk Prediction in Revascularized Coronary Artery Disease Patients 90%
Similar papers in this journal
- Opt-out policies capacity to increase organ donors is limited 95%
- Rehabilitation interventions to modify physical frailty in adults before lung transplantation: A systematic review protocol 91%
- A Delphi study to establish a consensus definition and clinical reporting guidelines for Mesenchymal Stromal Cells 90%
Similar papers in this journal
- Balancing Equity and HLA Matching in Deceased-Donor Kidney Allocation with Eplet Mismatch 94%
- Machine learning-supported interpretation of kidney graft elementary lesions in combination with clinical data 93%
- Direct and indirect impact of the COVID-19 pandemic on the survival of kidney transplant recipients: a national observational study in France 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.