Population-split-based risk assessment model of venous thromboembolism in Chinese medical inpatients
Wang, X.; yang, y.; Hong, X.; Liu, S.; Li, J.; Chen, T.; Shi, J.
Show abstract
AbstractsO_ST_ABSObjectiveC_ST_ABSInpatients with high risk of venous thromboembolism (VTE) usually face serious threats to their health and economic conditions. Many studies using machine learning (ML) models to predict VTE risk neglected an important statistical phenomenon, fuzzy feature, and achieved inferior results. Considering the effect of fuzzy feature, our study aims to develop a VTE risk assessment model suitable for Chinese medical inpatients. Materials and MethodsInpatients in the medical department of Peking Union Medical College Hospital (PUMCH) from January 2014 to June 2016 were collected. A new ML VTE risk assessment model was built through population splitting. First patients were classified into different groups based on values of VTE risk factors, then trustless groups were filtered out, and finally ML models were built on training data in unit of groups. Predictive performances of our method, five traditional ML models, and the Padua model were compared. ResultsThe fuzzy feature was verified on the whole dataset. Compared with the Padua model, the proposed model showed higher sensitivities and specificities on training data, and higher specificities and similar sensitivities on test data. Standard deviations of predictive validity of five ML models were larger than the proposed model. DiscussionThe proposed model was the only one which showed advantages on both sensitivity and specificity over Padua model. Its robustness was better than traditional ML models. ConclusionThis study built a population-split-based ML model of VTE for Chinese medical inpatients and it may help clinicians stratify VTE risk and guide prevention more efficiently.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- The PBL teaching method in Neurology Education in the Traditional Chinese Medicine undergraduate students: An Observational Study 94%
- SGCI (Smartwatch Gait Coordination Index): New Measure for Human Gait Utilizing Smartwatch Sensor 93%
- Upregulation of ARHGAP9 is correlated with poor prognosis and immune infiltration in clear cell renal cell carcinoma 93%
Similar papers in this journal
- Regional medical inter-institutional cooperation in medical provider network constructed using patient claims data from Japan 95%
- Machine learning based prediction of recurrence after curative resection for rectal cancer 95%
- Leveraging Machine Learning for Enhanced and Interpretable Risk Prediction of Venous Thromboembolism in Acute Ischemic Stroke Care 95%
Similar papers in this journal
- Unsupervised Discovery of Risk Profiles on Negative and Positive COVID-19 Hospitalized Patients 96%
- Identification of Myocardial Infarction (MI) Probability from Imbalanced Medical Survey Data: An Artificial Neural Network (ANN) with Explainable AI (XAI) Insights 95%
- Biomechanical stress analysis of Type-A aortic dissection at pre-dissection, post-dissection, and post-repair states 93%
Similar papers in this journal
- Comparing protein-protein interaction networks of SARS-CoV-2 and (H1N1) influenza using topological features 94%
- Rapid Clinical Screening and Staging for COVID-19 Severe Outcome A Hospitalization Study in New York City 94%
- Classification models for Invasive Ductal Carcinoma Progression, based on gene expression data-trained supervised machine learning 93%
Similar papers in this journal
- MEAHNE: MiRNA-disease association prediction based on semantic information in heterogeneous networks 90%
- Generating complex explanations for artificial intelligence models: an application to clinical data on severe mental illness 90%
- Single-cell transcriptome profiling simulation reveals the impact of sequencing parameters and algorithms on clustering 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.