A hybrid-computer vision model to predict lung cancer in diverse patient populations
Zakkar, A.; Perwaiz, N.; Zhong, W.; Krule, A.; Burrage-Burton, M.; Kim, D.; Miglani, M.; Narra, V.; Yousef, F.; Gadi, V.; Korpics, M. C.; Kim, S. J.; Khan, A. A.; Molina, Y.; Dai, Y.; Marai, E.; Meidani, H.; Nguyen, R.; Salahudeen, A. A.
Show abstract
PURPOSEDisparities of lung cancer incidence exist in Black populations and screening criteria underserve Black populations due to disparately elevated risk in the screening eligible population. Prediction models that integrate clinical and imaging-based features to individualize lung cancer risk is a potential means to mitigate these disparities. PATIENTS AND METHODSThis Multicenter (NLST) and catchment population based (UIH, urban and suburban Cook County) cross-sectional study utilized participants at risk of lung cancer with available lung CT imaging and follow up between the years 2015 and 2024. 53,452 in NLST and 11,654 in UIH were included based on age and tobacco use based risk factors for lung cancer. Cohorts were used for training and testing of deep and machine learning models using clinical features alone or combined with CT image features (hybrid computer vision). RESULTSAn optimized 7 clinical feature model achieved ROC-AUC values ranging 0.64-0.67 in NLST and 0.60-0.65 in UIH cohorts across multiple years. Incorporation of imaging features to form a hybrid computer vision model significantly improved ROC-AUC values to 0.78-0.91 in NLST but deteriorated in UIH with ROC-AUC values of 0.68-0.80, attributable to Black participants where ROC-AUC values ranged from 0.63-0.72 across multiple years. Retraining the hybrid computer vision model by incorporating Black and other participants from the UIH cohort improved performance with ROC-AUC values of 0.70-0.87 in a held out UIH test set. CONCLUSIONHybrid computer vision predicted risk with improved accuracy compared to clinical risk models alone. However, potential biases in image training data reduced model generalizability in Black participants. Performance was improved upon retraining with a subset of the UIH cohort, suggesting that inclusive training and validation datasets can minimize racial disparities. Future studies incorporating vision models trained on representative data sets may demonstrate improved health equity upon clinical use.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- An ML prediction model based on clinical parameters and automated CT scan features for COVID-19 patients 95%
- Opportunistic Assessment of Ischemic Heart Disease Risk Using Abdominopelvic Computed Tomography and Medical Record Data: a Multimodal Explainable Artificial Intelligence Approach 94%
- High-Dimensional Multinomial Multiclass Severity Scoring of COVID-19 Pneumonia Using CT Radiomics Features and Machine Learning Algorithms 93%
Similar papers in this journal
- Evaluation of an artificial intelligence model for detection of pneumothorax and tension pneumothorax on chest radiograph 90%
- A Crowdsourcing Approach to Develop Machine Learning Models to Quantify Radiographic Joint Damage in Rheumatoid Arthritis 90%
- Missing data in the medical record for oncology patients: prevalence and association with outcomes 90%
Similar papers in this journal
- Histology-based Prediction of Therapy Response to Neoadjuvant Chemotherapy for Esophageal and Esophagogastric Junction Adenocarcinomas Using Deep Learning 92%
- Towards Predicting 30-Day Readmission among Oncology Patients: Identifying Timely and Actionable Risk Factors 91%
- Simple Linear Cancer Risk Prediction Models with Novel Features Outperform Complex Approaches 91%
Similar papers in this journal
- Integration of clinical characteristics, lab tests and a deep learning CT scan analysis to predict severity of hospitalized COVID-19 patients 95%
- The Impact of Digital Histopathology Batch Effect on Deep Learning Model Accuracy and Bias 93%
- Artificial intelligence-based histopathology image analysis identifies a novel subset of endometrial cancers with distinct genomic features and unfavourable outcome 91%
Similar papers in this journal
- Volumetric lung cancer screening reduces unnecessary low-dose computed tomography scans: results from a single-centre prospective trial on 4,119 subjects 92%
- Deep learning models for poorly differentiated colorectal adenocarcinoma classification in whole slide images using transfer learning 92%
- Auto-detection of motion artifacts on CT pulmonary angiograms with a physician-trained AI algorithm 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.