Pilot study demonstrating changes in DNA hydroxymethylation enable detection of multiple cancers in plasma cell-free DNA
Bergamaschi, A.; Ning, Y.; Ku, C.-J.; Ellison, C.; Collin, F.; Guler, G.; Phillips, T.; McCarthy, E.; Wang, W.; Antoine, M.; Scott, A.; Lloyd, P.; Ashworth, A.; Quake, S.; Levy, S.
Show abstract
Our study employed the detection of 5-hydroxymethyl cytosine (5hmC) profiles on cell free DNA (cfDNA) from the plasma of cancer patients using a novel enrichment technology coupled with sequencing and machine learning based classification method. These classification methods were develoiped to detect the presence of disease in the plasma of cancer and control subjects. Cancer and control patient cfDNA cohorts were accrued from multiple sites consisting of 48 breast, 55 lung, 32 prostate and 53 pancreatic cancer subjects. In addition, a control cohort of 180 subjects (non-cancer) was employed to match cancer patient demographics (age, sex and smoking status) in a case-control study design. Logistic regression methods applied to each cancer case cohort individually, with a balancing non-cancer cohort, were able to classify cancer and control samples with measurably high performance. Measures of predictive performance by using 5-fold cross validation coupled with out-of-fold area under the curve (AUC) measures were established for breast, lung, pancreatic and prostate cancer to be 0.89, 0.84, 0.95 and 0.83 respectively. The genes defining each of these predictive models were enriched for pathways relevant to disease specific etiology, notably in the control of gene regulation in these same pathways. The breast cancer cohort consisted primarily of stage I and II patients, including tumors < 2 cm and these samples exhibited a high cancer probability score. This suggests that the 5hmC derived classification methodology may yield epigenomic detection of early stage disease in plasma. Same observation was made for the pancreatic dataset where >50% of cancers were stage I and II and showed the highest cancer probability score.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genomic alterations and abnormal expression of APE2 in multiple cancers 93%
- Bioinformatic Screen with Clinical Validation for the Identification of Novel Stool Based mRNA Biomarkers for the Detection of Colorectal Lesions Including Advanced Precancerous Lesions 92%
- Multi-omic signatures identify pan-cancer classes of tumors beyond tissue of origin. 92%
Similar papers in this journal
- Predicting cancer origins with a DNA methylation-based deep neural network model 92%
- Comprehensive cancer-oriented biobanking resource of human samples for studies of post-zygotic genetic variation involved in cancer predisposition 91%
- Sparse canonical correlation to identify breast cancer related genes regulated by copy number aberrations 91%
Similar papers in this journal
- BC-Predict: Mining of signal biomarkers and multilevel validation of cascade classifier for early-stage breast cancer subtyping and prognosis 92%
- Machine-learning-based determination of sex-related bladder cancer biomarkers 92%
- Predicting GD2 expression across cancer types by the integration of pathway topology and transcriptome data 91%
Similar papers in this journal
- Longitudinal cell-free DNA methylome and fragmentome profiles in health uncover signatures of cell type and demographic origin 92%
- Neutrophil extracellular traps have auto-catabolic activity and produce mononucleosome-associated circulating DNA 91%
- DNA methylation reveals distinct cells of origin for pancreatic neuroendocrine carcinomas (PanNECs) and pancreatic neuroendocrine tumors (PanNETs) 91%
Similar papers in this journal
- A tissue specific atlas of gene promoter DNA methylation variability and the clinical value of its assessment 92%
- A transcriptomics-based meta-analysis combined with machine learning approach identifies a secretory biomarker panel for diagnosis of pancreatic adenocarcinoma 92%
- Time-series plasma cell-free DNA analysis reveals disease severity of COVID-19 patients 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.