Design and performance of a long-read sequencing panel for pharmacogenomics
van der Lee, M.; Busscher, L.; Menafra, R.; Zhai, Q.; van den Berg, R. R.; Kingan, S. B.; Gonzaludo, N.; Hon, T.; Han, T.; Arbiza, L.; Numanagic, I.; Kloet, S. L.; Swen, J. J.
Show abstract
Pharmacogenomics (PGx)-guided drug treatment is one of the cornerstones of personalized medicine. However, the genes involved in drug response are highly complex and known to carry many (rare) variants. Current technologies (short-read sequencing and SNP panels) are limited in their ability to resolve these genes and characterize all variants. Moreover, these technologies cannot always phase variants to their allele of origin. Recent advance in long-read sequencing technologies have shown promise in resolving these problems. Here we present a long-read sequencing panel-based approach for PGx using PacBio HiFi sequencing. A capture based approach was developed using a custom panel of clinically-relevant pharmacogenes including up- and downstream regions. A total of 27 samples were sequenced and panel accuracy was determined using benchmarking variant calls for 3 Genome in a Bottle samples and GeT-RM star(*)-allele calls for 21 samples.. The coverage was uniform for all samples with an average of 94% of bases covered at >30x. When compared to benchmarking results, accuracy was high with an average F1 score of 0.89 for INDELs and 0.98 for SNPs. Phasing was good with an average of 68% the target region phased (compared to ~20% for short-reads) and an average phased haploblock size of 6.6kbp. Using Aldy 4, we compared our variant calls to GeT-RM data for 8 genes (CYP2B6, CYP2C19, CYP2C9, CYP2D6, CYP3A4, CYP3A5, SLCO1B1, TPMT), and observed highly accurate star(*)-allele calling with 98.2% concordance (165/168 calls), with only one discordance in CYP2C9 leading to a different predicted phenotype. We have shown that our long-read panel-based approach results in high accuracy and target phasing for SNVs as well as for clinical star(*)-alleles.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Array Comparative Genomic Hybridisation and Droplet Digital PCR uncover recurrent copy number variation of the titin segmental duplication region 93%
- The FORCE panel: An all-in-one SNP marker set for confirming investigative genetic genealogy leads and for general forensic applications 93%
- VarGenius-HZD allows accurate detection of rare homozygous or hemizygous deletions in targeted sequencing leveraging breadth of coverage 92%
Similar papers in this journal
- Variant calling and genotyping accuracy of ddRAD-seq: comparison with 20X WGS in layers 93%
- Pharmacogenetic allele variant frequencies: An analysis of the VAs Million Veteran Program (MVP) as a representation of the diversity in US population. 93%
- CaBagE: a Cas9-based Background Elimination strategy for targeted, long-read DNA sequencing 93%
Similar papers in this journal
- Identification of a CCG-enriched expanded allele in DM1 patients using Amplification-free long-read sequencing 94%
- Third Generation Cytogenetic Analysis (TGCA): diagnostic application of long-read sequencing. 92%
- Clinical Validation and Diagnostic Utility of Optical Genome Mapping in Prenatal Diagnostic Testing 92%
Similar papers in this journal
- Optimised multiplex amplicon sequencing for mutation identification using the MinION nanopore sequencer 94%
- Rapid detection of G6PD deficiency SNPs using a novel amplicon-based MinION Sequencing Assay 94%
- High Precision Characterization Of Rccx Rearrangements In A 21-Hydroxylase Deficiency Latin American Cohort Using Oxford Nanopore Long Read Sequencing 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.