Buyer Beware: confounding factors and biases abound when predicting omics-based biomarkers from histological images
Dawood, M.; Branson, K.; Tejpar, S.; Rajpoot, N.; Minhas, F.
Show abstract
BackgroundRecent advancements in computational pathology have introduced deep learning methods to predict genomic, transcriptomic and molecular biomarkers from routine histology whole slide images (WSIs) for cancer diagnosis, prognosis, and treatment. However, existing methods often overlook the critical role of co-dependencies among biomarker statuses during training and inference. We hypothesize that this oversight results in models that predict the combined effect of multiple interdependent biomarkers rather than individual statuses independently, akin to attributing the quality of an orchestral symphony to a single instrument, highlighting limitations of current predictors. MethodsUsing large datasets (n = 8,221 patients), we conducted statistical co-dependence testing to demonstrate significant interdependencies among biomarker statuses in training datasets. Following standard protocols, we trained two machine learning models to predict biomarkers from WSIs achieving or matching state-of-the-art predictive performance. We then employed permutation testing and stratification analysis to evaluate their predictive quality based on the principle of conditional independence, i.e., if a model accurately captures the phenotypic influence of a specific biomarker independent of other biomarkers, its performance should remain consistent across subgroups of patients stratified by other biomarkers, aligning with its overall performance on the entire dataset. FindingsOur statistical analysis reveals significant interdependencies among biomarkers, reflecting expected co-occurrence and mutual exclusivity patterns influenced by pathological and biological processes that are consistent across datasets, as well as sampling artefacts that can be different across datasets. Our results indicate that the predictive quality of an image-based predictor for a biomarker is contingent on the status of other biomarkers, revealing that models capture aggregated influences rather than predicting individual statuses independently. For example, mutation predictions are confounded by the overall tumour mutation burden. We also show that, due to the presence of such correlations, deep learning models may not offer significant advantages in predicting certain biomarkers in comparison to simply using pathologist-assigned grades for their prediction. InterpretationWe show that current deep learning models in computational pathology fall short in isolating individual biomarker effects, leading to confounded and less precise predictions. Our findings suggest revisiting model training protocols to recognize and adjust for biomarker interdependencies at all development stages--from problem definition to usage guidelines. This involves selecting diverse datasets to reflect clinical heterogeneity, defining prediction variables or grouping patients based on co-dependencies, designing models to disentangle complex relationships, and stringent stratification testing. Clinically, failure to account for interdependencies may lead to suboptimal decisions, necessitating appropriate usage guidelines for predictive models.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Generalizing AI-driven Assessment of Immunohistochemistry across Immunostains and Cancer Types: A Universal Immunohistochemistry Analyzer 95%
- A Deep Learning Model for Molecular Label Transfer that Enables Cancer Cell Identification from Histopathology Images 95%
- Development and validation of a gene expression score to account for tumour purity and improve prognostication in breast cancer 95%
Similar papers in this journal
- Spatial Transcriptomics Inferred from Pathology Whole-Slide Images Links Tumor Heterogeneity to Survival in Breast and Lung Cancer 95%
- Aggregation of Cohorts for Histopathological Diagnosis with Deep Morphological Analysis 95%
- Accurate Prediction of Breast Cancer Survival through Coherent Voting Networks with Gene Expression Profiling 95%
Similar papers in this journal
- Simple Linear Cancer Risk Prediction Models with Novel Features Outperform Complex Approaches 95%
- Machine learning and mechanistic modeling for prediction of metastatic relapse in breast cancer 94%
- Histology-based Prediction of Therapy Response to Neoadjuvant Chemotherapy for Esophageal and Esophagogastric Junction Adenocarcinomas Using Deep Learning 94%
Similar papers in this journal
- Breast invasive ductal carcinoma classification on whole slide images with weakly-supervised and transfer learning 94%
- Use of high-plex data reveals novel insights into the tumour microenvironment of clear cell renal cell carcinoma 94%
- Clustering Digestive Tract Tumors Using Transcriptomic and Mutation Data 92%
Similar papers in this journal
- The Impact of Digital Histopathology Batch Effect on Deep Learning Model Accuracy and Bias 96%
- Artificial intelligence-based histopathology image analysis identifies a novel subset of endometrial cancers with distinct genomic features and unfavourable outcome 96%
- Teacher-student collaborated multiple instance learning for pan-cancer PDL1 expression prediction from histopathology slides 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.