AmpliconSuite: an end-to-end workflow for analyzing focal amplifications in cancer genomes
Luebeck, J.; Huang, E.; Kim, F.; Liefeld, T.; Dameracharla, B.; Ahuja, R.; Schreyer, D.; Prasad, G.; Adamaszek, M.; Kenkre, R.; Agashe, T.; Torvi, D.; Tabor, T.; Giurgiu, M.; Kim, S.; Kim, H.; Bailey, P.; Verhaak, R. G. W.; Deshpande, V. B.; Reich, M. M.; Mischel, P. S.; Mesirov, J.; Bafna, V.
Show abstract
Focal amplifications in the cancer genome, particularly extrachromosomal DNA (ecDNA) amplifications, are emerging as a pivotal event in cancer progression across diverse cancer contexts, presenting a paradigm shift in our understanding of tumor dynamics. Simultaneously, identification of the various modes of focal amplifications is bioinformatically challenging. We present AmpliconSuite, a collection of tools that enables robust identification of focal amplifications from whole-genome sequencing data. AmpliconSuite includes AmpliconSuite- pipeline; utilizing the AmpliconArchitect (AA) method, and AmpliconRepository.org; a community- editable website for the sharing of focal amplification calls. We also describe improvements made to AA since its initial release that improve its accuracy and speed. As a proof of principle, we utilized publicly available pan-cancer datasets encompassing 2,525 tumor samples hosted on AmpliconRepository.org to illustrate important properties of focal amplifications, showing ecDNA has higher copy number, and stronger oncogene enrichment, compared to other classes of focal amplifications. Finally, we illustrate how AmpliconSuite-pipeline enables delineation of the various mechanisms by which ecDNA forms.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- SVCROWS: A User-Defined Tool for Interpreting Significant Structural Variants in Heterogeneous Datasets 96%
- Methyl-CODEC enables simultaneous methylation and duplex sequencing 96%
- Precise characterization of somatic complex structural variations from paired long-read sequencing data with nanomonsv 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.