Whole-Lung Ct Radiomics-Based Machine Learning Classification of Nontuberculous Mycobacteria Lung Disease Across Geographically Distinct Cohort
Kanagala, A.; Garcia, B.; Dutt, T. S.; Aguilera, S. M.; Pudhota, A. S.; Panjwani, D. D.; Dukkipati, N.; Gaggar, A.; Naidoo, T.; Jololian, L.; Bhatt, S. P.; Margaroli, C.; Bodduluri, S.
Show abstract
BACKGROUND Nontuberculous mycobacterial lung disease (NTM-LD) is highly heterogenous, geographically and etiologically, hindering effective timely identification. Prior CT radiomics studies require manual segmentation of pathology. We developed a whole-lung CT radiomics-based machine learning approach and identified common features across two geographically distinct NTM-LD cohorts. STUDY DESIGN AND METHODS 1,300 chest CT scans from China (871 TB; 429 NTM, Dataset 1) and 173 independent NTM cohort from UAB, US. Whole-lung regions were automatically segmented on each scan, and 85 quantitative radiomic features were extracted using a standardized image-processing pipeline. We evaluated two frameworks to assess model performance and generalizability: (1) training on Dataset 1 with external validation on Dataset 2, and (2) training on the combined cohort. Linear discriminant analysis (LDA) was used as the primary classification method. Cross-cohort concordance analysis was performed to evaluate the reproducibility of radiomic features across datasets. RESULTS In Scenario 1, the LDA classifier trained on Dataset 1 achieved an AUC of 0.79 (95% CI, 0.73-0.84) with high specificity (0.91). On the external UAB cohort, the model achieved an AUC of 0.94 (95% CI, 0.90-0.97). In Scenario 2, the combined cohort model achieved an AUC of 0.81 (95% CI, 0.76-0.85) with improved sensitivity (0.61) and precision (0.82). Feature importance analysis identified 16 features consistently ranked among the top 20 in both scenarios, predominantly texture-based descriptors reflecting distinct parenchymal patterns between mycobacterial species. CONCLUSION Whole-lung CT radiomics enables interpretable NTM-LD classification across geographically distinct populations without manual annotation. Suggesting population-independent parenchymal signatures of NTM-LD.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Automated airway quantification associates with mortality in idiopathic pulmonary fibrosis 94%
- Observer agreement and clinical significance of chest CT reporting in patients suspected of COVID-19 92%
- From Community Acquired Pneumonia to COVID-19: A Deep Learning Based Method for Quantitative Analysis of COVID-19 on thick-section CT Scans 92%
Similar papers in this journal
- An ML prediction model based on clinical parameters and automated CT scan features for COVID-19 patients 95%
- High-Dimensional Multinomial Multiclass Severity Scoring of COVID-19 Pneumonia Using CT Radiomics Features and Machine Learning Algorithms 94%
- Immune, metabolic, anatomical, and functional features of people after successful tuberculosis treatment: an exploratory analysis 94%
Similar papers in this journal
- Annotation-free multi-organ anomaly detection in abdominal CT using free-text radiology reports: A multi-center retrospective study 92%
- Fusing Data from CT Deep Learning, CT Radiomics and Peripheral Blood Immune profiles to Diagnose Lung Cancer in Symptomatic Patients 91%
- Diagnostic accuracy of swab-based molecular tests for tuberculosis using novel near point-of-care platforms: A multi-country evaluation 91%
Similar papers in this journal
- Volumetric lung cancer screening reduces unnecessary low-dose computed tomography scans: results from a single-centre prospective trial on 4,119 subjects 93%
- Auto-detection of motion artifacts on CT pulmonary angiograms with a physician-trained AI algorithm 92%
- Passive Microwave Radiometry (MWR) for diagnostics of COVID-19 lung complications in Kyrgyzstan 89%
Similar papers in this journal
- Organizing Pneumonia of COVID-19: Time-dependent Evolution and Outcome in CT Findings 93%
- The impact of respiratory gating on improving volume measurement of murine lung tumors in micro-CT imaging 92%
- Classification performance bias between training and test sets in a limited mammography dataset 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.