Exploring the genetic and epigenetic underpinnings of early-onset cancers: Variant prioritization for long read whole genome sequencing from family cancer pedigrees
Kramer, M.; Goodwin, S.; Wappel, R.; Borio, M.; Offit, K.; Feldman, D. R.; Stadler, Z. K.; McCombie, W. R.
Show abstract
Despite significant advances in our understanding of genetic cancer susceptibility, known inherited cancer predisposition syndromes explain at most 20% of early-onset cancers. As early-onset cancer prevalence continues to increase, the need to assess previously inaccessible areas of the human genome, harnessing a trio or quad family-based architecture for variant filtration, may reveal further insights into cancer susceptibility. To assess a broader spectrum of variation than can be ascertained by multi-gene panel sequencing, or even whole genome sequencing with short reads, we employed long read whole genome sequencing using an Oxford Nanopore Technology (ONT) PromethION of 3 families containing an early-onset cancer proband using a trio or quad family architecture. Analysis included 2 early-onset colorectal cancer family trios and one quad consisting of two siblings with testicular cancer, all with unaffected parents. Structural variants (SVs), epigenetic profiles and single nucleotide variants (SNVs) were determined for each individual, and a filtering strategy was employed to refine and prioritize candidate variants based on the family architecture. The family architecture enabled us to focus on inapposite variants while filtering variants shared with the unaffected parents, significantly decreasing background variation that can hamper identification of potentially disease causing differences. Candidate de novo and compound heterozygous variants were identified in this way. Gene expression, in matched neoplastic and pre-neoplastic lesions, was assessed for one trio. Our study demonstrates the feasibility of a streamlined analysis of genomic variants from long read ONT whole genome sequencing and a way to prioritize key variants for further evaluation of pathogenicity, while revealing what may be missing from panel based analyses.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Discovery of an unusual high number of de novo mutations in sperm of older men using duplex sequencing 95%
- Long-read genome sequencing and variant reanalysis increase diagnostic yield in neurodevelopmental disorders 94%
- Discordant calls across genotype discovery approaches elucidate variants with systematic errors 94%
Similar papers in this journal
- Concerning the eXclusion in human genomics: The choice of sex chromosome representation in the human genome drastically affects number of identified variants 94%
- Low-pass sequencing plus imputation using avidity sequencing displays comparable imputation accuracy to sequencing by synthesis while reducing duplicates 93%
- Protein Coding Variation In Outbred Laboratory Mouse Stocks Provides A Molecular Basis For Distinct Research Applications 91%
Similar papers in this journal
Similar papers in this journal
- A Simple Deep Learning Approach for Detecting Duplications and Deletions in Next-Generation Sequencing Data 93%
- Network-based functional prediction augments genetic association to predict candidate genes for histamine hypersensitivity in mice 92%
- Analysis of independent cohorts of outbred CFW mice reveals novel loci for behavioral and physiological traits and identifies factors determining reproducibility 91%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.