XenoCP: Cloud-based BAM cleansing tool for RNA and DNA from Xenograft
arunachalam, s.; Zhang, J.; Rusch, M.; Ding, L.; Thrasher, A.; Dyer, M.; Baker, S.; Jin, H.; Macias, M.; Kasper, L.
Show abstract
SummaryXenografts are important models for cancer research and the presence of mouse reads in xenograft next generation sequencing data can potentially confound interpretation of experimental results. We present an efficient, cloud-based BAM-to-BAM cleaning tool called XenoCP to remove mouse reads from xenograft BAM files. We show application of XenoCP in obtaining accurate gene expression quantification in RNA-seq and tumor heterogeneity in WGS of xenografts derived from brain and solid tumors. Availability and ImplementationSt. Jude Cloud (https://pecan.stjude.cloud/permalink/xenocp) and St. Jude Github (https://github.com/stjude/XenoCP)
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Uniform Genomic Data Analysis in the NCI Genomic Data Commons 95%
- Partner-independent fusion gene detection by multiplexed CRISPR/Cas9 enrichment and long-read Nanopore sequencing 94%
- Rescuing Low Frequency Variants within Intra-Host Viral Populations directly from Oxford Nanopore sequencing data 94%
Similar papers in this journal
- Short and long-read genome sequencing methodologies for somatic variant detection; genomic analysis of a patient with diffuse large B-cell lymphoma 93%
- GASOLINE: detecting germline and somatic structural variants from long-reads data. 93%
- Reliable variant calling during runtime of Illumina sequencing 92%
Similar papers in this journal
- Samplot: A Platform for Structural Variant Visual Validation and Automated Filtering 93%
- MINTIE: identifying novel structural and splice variants in transcriptomes using RNA-seq data 93%
- Comprehensive characterization of single cell full-length isoforms in human and mouse with long-read sequencing 92%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.