An optimised computational approach for the identification of somatic structural variants in cancer
Waise, S.; Mensah, N.; Lesluyes, T.; Demeulemeester, J.; Flanagan, A. M.; Pillay, N.; Van Loo, P.
Show abstract
Structural variants play a critical role in tumorigenesis. At present, these events are most commonly identified using short-read whole-genome sequencing data, and a number of computational tools are available for this purpose. Consensus approaches have been used to improve precision, but may reduce sensitivity. The optimal number and combination of callers remains unclear, in part due to the lack of gold standard real-world datasets for validation. Here, we benchmark the performance of Delly, GRIDSS, LUMPY, Manta and SvABA, using a validation set of consensus calls from the Pan-Cancer Analysis of Whole Genomes Consortium. Manta showed the best standalone performance, identifying 88% of the validation set calls, and was included in all of the best-performing caller combinations. A consensus approach comprising Delly, GRIDSS, Manta and SvABA was selected as the optimum approach from those tested. We provide a NextFlow implementation of our optimised consensus approach as a resource for the cancer genomics community.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- MetaRNN: Differentiating Rare Pathogenic and Rare Benign Missense SNVs and InDels Using Deep Learning 94%
- OncoGEMINI: Software for Investigating Tumor Variants From Multiple Biopsies With Integrated Cancer Annotations 94%
- Genome-Wide Sequencing as a First-Tier Screening Test for Short Tandem Repeat Expansions 94%
Similar papers in this journal
Similar papers in this journal
- Reducing Sanger Confirmation Testing through False Positive Prediction Algorithms 94%
- IGenomic answers for children: Dynamic analyses of >1000 pediatric rare disease genomes 93%
- Genome Alert!: a standardized procedure for genomic variant reinterpretation and automated genotype-phenotype reassessment in clinical routine 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.