Revisiting the cancer microbiome using PRISM
Ghaddar, B. C.; Blaser, M. J.; De, S.
Show abstract
Recent controversy around the cancer microbiome highlights the need for improved microbial analysis methods for human genomics data. We developed PRISM, a computational approach for precise microorganism identification and decontamination from low-biomass sequencing data. PRISM removes spurious signals and achieves excellent performance when benchmarked on a curated dataset of 62,006 known true- and false-positive taxa. We then use PRISM to detect microbes in 8 cancer types from the CPTAC and TCGA datasets. We identify rich microbiomes in gastrointestinal tract tumors in CPTAC and identify bacteria in a subset of pancreatic tumors that are associated with altered glycoproteomes, more extensive smoking histories, and higher tumor recurrence risk. We find relatively sparse microbes in other cancer types and in TCGA, which we demonstrate may reflect differing sequencing parameters. Overall, PRISM does not replace gold-standard controls, but it enables higher-confidence analyses and reveals tumor-associated microorganisms with potential molecular and clinical significance.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- MaAsLin 3: Refining and extending generalized multivariable linear models for meta-omic association discovery 96%
- mEnrich-seq: Methylation-guided enrichment sequencing of bacterial taxa of interest from microbiome 96%
- Bin Chicken: targeted metagenomic coassembly for the efficient recovery of novel genomes 96%
Similar papers in this journal
- Processing-bias correction with DEBIAS-M improves cross-study generalization of microbiome-based prediction models 97%
- No evidence for a common blood microbiome based on a population study of 9,770 healthy humans 97%
- A human gut metagenome-assembled genome catalogue spanning 41 countries supports genome-scale metabolic models 97%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.