Enhancing Cause of Death Prediction: Development and Validation of ML Models Using Multimodal Data Across Multiple Healthcare Sites
Al-Garadi, M.; Desai, R. J.; Ngan, K.; LeNoue-Newton, M.; Reeves, R. M.; Park, D.; Hernandez-Munoz, J. J.; Wang, S. V.; Maro, J. C.; Fuller, C. C.; Kueiyu, J. L.; Kuzucan, A.; Coughlin, K.; Pillai, H.; McPheeters, M.; Whitaker, J.; Deere, J. A.; McLemore, M. F.; Westerman, D. M.; Morrow, J. A.; Adgent, M. A.; Matheny, M. E.
Show abstract
ImportanceTimely and accurate determination of causes of death (CoD) is essential for public health surveillance, epidemiological research, and healthcare policy development. However, obtaining up-to-date and detailed CoD information is challenging due to delays in official death records and inconsistencies in data reporting across institutions. ObjectiveTo develop and validate machine learning (ML) models capable of predicting probable CoD by integrating comprehensive features from structured electronic health record (EHR) data, unstructured clinical notes, and publicly available data. Design, Setting, and ParticipantsThis multi-institutional retrospective cohort study was conducted at Vanderbilt University Medical Center (VUMC) and Massachusetts General Brigham (MGB). Deceased patients were included if they had at least one inpatient or outpatient encounter between October 1, 2015, and January 1, 2021, with corresponding death records from state health departments and the National Death Index. The study was comprised of 13,708 deceased patients from VUMC and 34,839 from MGB. ExposuresIntegration of structured EHR data, unstructured clinical notes processed using advanced language models, and publicly available data into machine learning models to predict CoD. Main Outcomes and MeasuresThe primary outcome was the underlying CoD, classified into one of the top 15 National Center for Health Statistics (NCHS) rankable CoD categories, with all other causes grouped into an "Other" category. Model performance was evaluated using weighted area under the receiver operating characteristic curve (AUC) and weighted F-measure. ResultsThe XGBoost model using structured EHR data alone achieved weighted AUCs of 0.86 (95% CI, 0.84-0.88) at VUMC and 0.80 (95% CI, 0.79-0.80) at MGB. Adding unstructured notes improved performance, with weighted AUCs of 0.90 (95% CI, 0.88-0.93) at VUMC and 0.92(95% CI, 0.91-0.92) at MGB. Adding publicly available data did not further improve performance. Cross-institutional validation revealed significant performance degradation. Conclusions and RelevanceML models integrating EHR structured and unstructured data to predict underlying CoD at the time of the most recent encounter among deceased patients achieved excellent performance within individual institutions. The inclusion of publicly available data did not improve performance, and all versions had poor portability between institutions. Healthcare institutions may benefit from adopting robust processes for locally tailored models, and future research should focus on enhancing model generalizability while addressing unique institutional data environments.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Real-Time Electronic Health Record Mortality Prediction During the COVID-19 Pandemic: A Prospective Cohort Study 97%
- Machine Learning Approaches for Electronic Health Records Phenotyping: A Methodical Review 96%
- Use of unstructured text in prognostic clinical prediction models: a systematic review 96%
Similar papers in this journal
- Machine Learning Generalizability Across Healthcare Settings: Insights from multi-site COVID-19 screening 95%
- Predicting critical state after COVID-19 diagnosis: Model development using a large US electronic health record dataset 95%
- Evaluating large language model workflows in clinical decision support: referral, triage, and diagnosis 94%
Similar papers in this journal
- A deep learning model for clinical outcome prediction using longitudinal inpatient electronic health records 95%
- Modeling physician variability to prioritize relevant medical record information 94%
- Evaluation of Patient-Level Retrieval from Electronic Health Record Data for a Cohort Discovery Task 93%
Similar papers in this journal
- Natural language processing for scalable feature engineering and ultra-high-dimensional confounding adjustment in healthcare database studies 95%
- A scoping review of fair machine learning techniques when using real-world data 95%
- Developing A Deep Learning Natural Language Processing Algorithm For Automated Reporting Of Adverse Drug Reactions 94%
Similar papers in this journal
- A Deep Learning Method to Detect Opioid Prescription and Opioid Use Disorder from Electronic Health Records 93%
- Predicting Prognosis in COVID-19 Patients using Machine Learning and Readily Available Clinical Data 93%
- Image and structured data analysis for prognostication of health outcomes in patients presenting to the Emergency Department during the COVID-19 pandemic 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.