scPathoQuant: A tool for efficient alignment and quantification of pathogen sequence reads from 10x single cell sequencing data sets
Whitmore, L. S.; Tisoncik-Go, J.; Gale, M.
Show abstract
Currently there is a lack of efficient computational pipelines/tools for conducting simultaneous genome mapping of pathogen-derived and host reads from single cell RNA sequencing (scRNAseq) output from pathogen-infected cells. Contemporary options include processes involving multiple steps and/or running multiple computational tools, increasing user operations time. To address the need for new tools to directly map and quantify pathogen and host sequence reads from within an infected cell from scRNAseq data sets in a single operation, we have built a python package, called scPathoQuant. scPathoQuant extracts sequences that were not aligned to the primary host genome, maps them to a pathogen genome of interest, here as demonstrated for viral pathogens, quantifies total reads mapping to the entire pathogen, quantifies reads mapping to individual pathogen genes, and finally reintegrates pathogen sequence counts into matrix files that are used by standard single cell pipelines for downstream analyses with only one command. We demonstrate that scPathoQuant provides a scRNAseq viral and host genome-wide sequence read abundance analysis that can differentiate and define multiple viruses in a single sample scRNAseq output.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Selective Ablation of 3' RNA ends and Processive RTs Facilitate Direct cDNA Sequencing of Full-length Host Cell and Viral Transcripts 96%
- Co-variation of viral recombination with single nucleotide variants during virus evolution revealed by CoVaMa 96%
- HIV-PULSE: A long-read sequencing assay for high-throughput near full-length HIV-1 proviral genome characterization 94%
Similar papers in this journal
- Kmerator Suite: design of specific k-mer signatures andautomatic metadata discovery in large RNA-Seq datasets. 94%
- High-resolution HIV-1 m6A epitranscriptome reveals isoform-dependent methylation clusters and unique 2-LTR transcript modifications 93%
- Repeat Detector: versatile sizing of expanded tandem repeats and identification of interrupted alleles from targeted DNA sequencing 92%
Similar papers in this journal
- A Human H5N1 Influenza Virus Expressing Bioluminescence for Evaluating Viral Infection and Identifying Therapeutic Interventions 92%
- Delta-Omicron recombinant escapes therapeutic antibody neutralization 92%
- Mycobacterium tuberculosis infection associated immune perturbations correlate with antiretroviral immunity 91%
Similar papers in this journal
- LINE1-mediated reverse transcription and genomic integration of SARS-CoV-2 mRNA detected in virus-infected but not in viral mRNA-transfected cells 94%
- Protocol and reagents for pseudotyping lentiviral particles with SARS-CoV-2 Spike protein for neutralization assays 93%
- Differential HIV-1 Proviral Defects in Children vs. Adults on Antiretroviral Therapy 93%
Similar papers in this journal
- A Short Plus Long-Amplicon Based Sequencing Approach Improves Genomic Coverage and Variant Detection In the SARS-CoV-2 Genome 95%
- Nanopore Sequencing of SARS-CoV-2: Comparison of Short and Long PCR-tiling Amplicon Protocols 94%
- Deficient uracil base excision repair leads to persistent dUMP in HIV proviruses during infection of monocytes and macrophages 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.