Performance and Robustness of Machine Learning-based Radiomic COVID-19 Severity Prediction
Yip, S. S. F.; Klanecek, Z.; Naganawa, S.; Kim, J.; Studen, A.; Rivetti, L.; Jeraj, R.
Show abstract
ObjectivesThis study investigated the performance and robustness of radiomics in predicting COVID-19 severity in a large public cohort. MethodsA public dataset of 1110 COVID-19 patients (1 CT/patient) was used. Using CTs and clinical data, each patient was classified into mild, moderate, and severe by two observers: (1) dataset provider and (2) a board-certified radiologist. For each CT, 107 radiomic features were extracted. The dataset was randomly divided into a training (60%) and holdout validation (40%) set. During training, features were selected and combined into a logistic regression model for predicting severe cases from mild and moderate cases. The models were trained and validated on the classifications by both observers. AUC quantified the predictive power of models. To determine model robustness, the trained models was cross-validated on the inter-observers classifications. ResultsA single feature alone was sufficient to predict mild from severe COVID-19 with [Formula] and [Formula] (p<< 0.01). The most predictive features were the distribution of small size-zones (GLSZM-SmallAreaEmphasis) for providers classification and linear dependency of neighboring voxels (GLCM-Correlation) for radiologists classification. Cross-validation showed that both [Formula]. In predicting moderate from severe COVID-19, first-order-Median alone had sufficient predictive power of [Formula]. For radiologists classification, the predictive power of the model increased to [Formula] as the number of features grew from 1 to 5. Cross-validation yielded [Formula] and [Formula]. ConclusionsRadiomics significantly predicted different levels of COVID-19 severity. The prediction was moderately sensitive to inter-observer classifications, and thus need to be used with caution. Key pointsO_LIInterpretable radiomic features can predict different levels of COVID-19 severity C_LIO_LIMachine Learning-based radiomic models were moderately sensitive to inter-observer classifications, and thus need to be used with caution C_LI
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Assessing GPT-4 Multimodal Performance in Radiological Image Analysis 94%
- First-generation clinical dual-source photon-counting CT: ultra-low dose quantitative spectral imaging 93%
- From Community Acquired Pneumonia to COVID-19: A Deep Learning Based Method for Quantitative Analysis of COVID-19 on thick-section CT Scans 93%
Similar papers in this journal
- Fully Automated Explainable Abdominal CT Contrast Media Phase Classification Using Organ Segmentation and Machine Learning 95%
- Phase Recognition in Contrast-Enhanced CT Scans based on Deep Learning and Random Sampling 92%
- SCU-Net: A deep learning method for segmentation and quantification of breast arterial calcifications on mammograms 91%
Similar papers in this journal
- Inconsistency of AI in Intracranial Aneurysm Detection with Varying Dose and Image Reconstruction 94%
- High-Dimensional Multinomial Multiclass Severity Scoring of COVID-19 Pneumonia Using CT Radiomics Features and Machine Learning Algorithms 93%
- Tracking And Predicting COVID-19 Radiological Trajectory Using Deep Learning On Chest X-Rays: Initial Accuracy Testing 93%
Similar papers in this journal
- Enhancing Semantic Segmentation in Chest X-Ray Images through Image Preprocessing: ps-KDE for Pixel-wise Substitution by Kernel Density Estimation 93%
- ai-corona : Radiologist-Assistant Deep Learning Framework for COVID-19 Diagnosis in Chest CT Scans 93%
- tbiExtractor: A framework for Extracting Traumatic Brain Injury Common Data Elements from Radiology Reports 92%
Similar papers in this journal
- Effects of contrast-medium and vertebral measurement level on computed tomography-based body composition parameters of skeletal muscle and adipose tissue 94%
- “This is a quiz” Premise Input: A Key to Unlocking Higher Diagnostic Accuracy in Large Language Models 93%
- Benchmarking Deep Learning-based Image Retrieval of Oral Tumor Histology 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.