Reference plasmid pHXB2_D is an HIV-1 molecular clone that exhibits identical LTRs and a single integration site indicative of an HIV provirus
Gener, A. R.; Foley, B. T.; Hyink, D. P.; Klotman, P. E.
Show abstract
ObjectiveTo compare long-read nanopore DNA sequencing (DNA-seq) with short-read sequencing-by-synthesis for sequencing a full-length (e.g., non-deletion, nor reporter) HIV-1 model provirus in plasmid pHXB2_D. DesignWe sequenced pHXB2_D and a control plasmid pNL4-3_gag-pol({Delta}1443-4553)_EGFP with long- and short-read DNA-seq, evaluating sample variability with resequencing (sequencing and mapping to reference HXB2) and de novo viral genome assembly. MethodsWe prepared pHXB2_D and pNL4-3_gag-pol({Delta}1443-4553)_EGFP for long-read nanopore DNA-seq, varying DNA polymerases Taq (Sigma-Aldrich) and Long Amplicon (LA) Taq (Takara). Nanopore basecallers were compared. After aligning reads to the reference HXB2 to evaluate sample coverage, we looked for variants. We next assembled reads into contigs, followed by finishing and polishing. We hired an external core to sequence-verify pHXB2_D and pNL4-3_gag-pol({Delta}1443-4553)_EGFP with single-end 150 base-long Illumina reads, after masking sample identity. ResultsWe achieved full-coverage (100%) of HXB2 HIV-1 from 5 to 3 long terminal repeats (LTRs), with median per-base coverage of over 9000x in one experiment on a single MinION flow cell. The longest HIV-spanning read to-date was generated, at a length of 11,487 bases, which included full-length HIV-1 and plasmid backbone with flanking host sequences supporting a single HXB2 integration event. We discovered 20 single nucleotide variants in pHXB2_D compared to reference, verified by short-read DNA sequencing. There were no variants detected in the HIV-1 segments of pNL4-3_gag-pol({Delta}1443-4553)_EGFP. ConclusionsNanopore sequencing performed as-expected, phasing LTRs, and even covering full-length HIV. The discovery of variants in a reference plasmid demonstrates the need for sequence verification moving forward, in line with calls from funding agencies for reagent verification. These results illustrate the utility of long-read DNA-seq to advance the study of HIV at single integration site resolution.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Identifying the best PCR enzyme for library amplification in NGS 93%
- Nanopore and Illumina Sequencing Reveal Different Viral Populations from Human Gut Samples 93%
- Analysis of the ARTIC V4 and V4.1 SARS-CoV-2 primers and their impact on the detection of Omicron BA.1 and BA.2 lineage defining mutations 92%
Similar papers in this journal
- Stable integrant-specific differences in bimodal HIV-1 expression patterns revealed by high-throughput analysis 95%
- Genetic variation of the HIV-1 subtype C transmitted/founder viruses long terminal repeat elements and the impact on transcription activation potential and clinical disease outcomes 94%
- No more business as usual: agile and effective responses to emerging pathogen threats require open data and open analytics 93%
Similar papers in this journal
- Differential HIV-1 Proviral Defects in Children vs. Adults on Antiretroviral Therapy 94%
- High throughput Next-Generation Sequencing Respiratory Viral Panel: A Diagnostic and Epidemiologic Tool for SARS-CoV-2 and Other Viruses. 94%
- Intrahost SARS-CoV-2 k-mer identification method (iSKIM) for rapid detection of mutations of concern reveals emergence of global mutation patterns 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.