Data Imputation for Clinical Trial Emulation: A Case Study on Impact of Intracranial Pressure Monitoring for Traumatic Brain Injury
Zhao, Z.; Liu, R.; Groner, J. I.; Xiang, H.; Zhang, P.
Show abstract
Randomized clinical trial emulation using real-world data is significant for treatment effect evaluation. Missing values are common in the observational data. Handling missing data improperly would cause biased estimations and invalid conclusions. However, discussions on how to address this issue in causal analysis using observational data are still limited. Multiple imputation by chained equations (MICE) is a popular approach to fill in missing data. In this study, we combined multiple imputation with propensity score weighted model to estimate the average treatment effect (ATE). We compared various multiple imputation (MI) strategies and a complete data analysis on two benchmark datasets. The experiments showed that data imputations had better performances than completely ignoring the missing data, and using different imputation models for different covariates gave a high precision of estimation. Furthermore, we applied the optimal strategy on a medical records data to evaluate the impact of ICP monitoring on inpatient mortality of traumatic brain injury (TBI). The experiment details and code are available at https://github.com/Zhizhen-Zhao/IPTW-TBI.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- External control arm analysis: an evaluation of propensity score approaches, G-computation, and doubly debiased machine learning 93%
- Comparing randomized trial designs to estimate treatment effect in rare diseases with longitudinal models: a simulation study showcased by Autosomal Recessive Cerebellar Ataxias using the SARA score 92%
- Comparison of Bayesian networks, G-estimation and linear models to estimate causal treatment effects in aggregated N-of-1 trials 91%
Similar papers in this journal
- Selecting the most important self-assessed features for predicting conversion to Mild Cognitive Impairment with Random Forest and Permutation-based methods 93%
- Using explainable machine learning to identify patients at risk of reattendance at discharge from emergency departments 92%
- Widely accessible prognostication using medical history for fetal growth restriction and small for gestational age in nationwide insured women 92%
Similar papers in this journal
- A scoping review of fair machine learning techniques when using real-world data 93%
- Automated Interpretable Discovery of Heterogeneous Treatment Effectiveness: A Covid-19 Case Study 92%
- Natural language processing for scalable feature engineering and ultra-high-dimensional confounding adjustment in healthcare database studies 91%
Similar papers in this journal
- Automated stratification of trauma injury severity across multiple body regions using multi-modal, multi-class machine learning models 93%
- Temporally-Informed Random Forests for Suicide Risk Prediction 92%
- Learning from local to global - an efficient distributed algorithm for modeling time-to-event data 92%
Similar papers in this journal
- Sepsis prediction via the clinical data integration system in the ICU 92%
- OASIS+: leveraging machine learning to improve the prognostic accuracy of OASIS severity score for predicting in-hospital mortality 91%
- On the predictability of postoperative complications for cancer patients: a Portuguese cohort study 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.