Contrastive Representation Learning for Single Cell Phenotyping in Whole Slide Imaging of Enrichment-free Liquid Biopsy
Naghdloo, A.; Tessone, D.; Mandya Nagaraju, R.; Zhang, B.; Kang, J.; Li, S.; Oberai, A.; Hicks, J. B.; Kuhn, P.
Show abstract
Tumor-associated cells derived from a liquid biopsy are promising biomarkers for cancer detection, diagnosis, prognosis, and monitoring. However, their rarity, heterogeneity and plasticity make precise identification and biological characterization challenging for clinical utility. Enrichment-free approaches using whole slide imaging of all circulating cells offer a comprehensive and unbiased strategy for capturing the full spectrum of tumor-associated cell phenotypes. However, current analysis methods often depend on engineered features and manual expert review, making them sensitive to technical variations and subjective biases. These limitations highlight the need for a better feature representation to improve performance and reproducibility of applications in large-scale patient cohort analyses. In this study, we present a deep contrastive learning framework for learning features of all circulating cells, enabling robust identification and stratification of single cells in whole slide immunofluorescence microscopy images. We demonstrate performance of learned features in classification of diverse cell phenotypes in the liquid biopsy, achieving an accuracy of 92.64%. We further demonstrate that learned features improve performance in downstream applications such as outlier detection and clustering. Lastly, our feature representation enables automated identification and enumeration of distinct rare cell phenotypes, achieving average F1-score of 0.93 across cell lines mimicking circulating tumor cells and endothelial cells in contrived samples and average F1-score of 0.858 across CTC phenotypes in clinical samples. This workflow has significant implications for scalable analysis of tumor-associated cellular biomarkers in clinical prognosis and personalized treatment strategies.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Quantitative image analysis pipeline for detecting circulating hybrid cells in immunofluorescence images with human-level accuracy 96%
- DeepIFC: virtual fluorescent labeling of blood cells in imaging flow cytometry data with deep learning 95%
- Cleanet: robust doublet detection in cytometry data based on protein expression patterns 94%
Similar papers in this journal
- A Deep Learning Model for Molecular Label Transfer that Enables Cancer Cell Identification from Histopathology Images 96%
- Generalizing AI-driven Assessment of Immunohistochemistry across Immunostains and Cancer Types: A Universal Immunohistochemistry Analyzer 95%
- Deep learning inference of cell type-specific gene expression from breast tumor histopathology 93%
Similar papers in this journal
- Obtaining Spatially Resolved Tumor Purity Maps Using Deep Multiple Instance Learning In A Pan-cancer Study 96%
- Bi-level Graph Learning Unveils Prognosis-Relevant Tumor Microenvironment Patterns in Breast Multiplexed Digital Pathology 95%
- Knowledge transfer to enhance the performance of deep learning models for automated classification of B-cell neoplasms 94%
Similar papers in this journal
- Unsupervised discovery of dynamic cell phenotypic states from transmitted light movies 95%
- The shape of cancer relapse: Topological data analysis predicts recurrence in paediatric acute lymphoblastic leukaemia 95%
- Learning unsupervised feature representations for single cell microscopy images with paired cell inpainting 94%
Similar papers in this journal
- ROSIE: AI generation of multiplex immunofluorescence staining from histopathology images 97%
- scSemiProfiler: Advancing Large-scale Single-cell Studiesthrough Semi-profiling with Deep Generative Models andActive Learning 95%
- METI: Deep profiling of tumor ecosystems by integrating cell morphology and spatial transcriptomics 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.