EdiTyper: a high-throughput tool for analysis of targeted sequencing data from genome editing experiments
Yahi, A.; Hoffman, P.; Brandt, M.; Mohammadi, P.; Tatonetti, N. P.; Lappalainen, T.
Show abstract
Genome editing experiments are generating an increasing amount of targeted sequencing data with specific mutational patterns indicating the success of the experiments and genotypes of clonal cell lines. We present EdiTyper, a high-throughput command line tool specifically designed for analysis of sequencing data from polyclonal and monoclonal cell populations from CRISPR gene editing. It requires simple inputs of sequencing data and reference sequences, and provides comprehensive outputs including summary statistics, plots, and SAM/BAM alignments. Analysis of simulated data showed that EdiTyper is highly accurate for detection of both single nucleotide mutations and indels, robust to sequencing errors, as well as fast and scalable to large experimental batches. EdiTyper is available in github (https://github.com/LappalainenLab/edityper) under the MIT license.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- HQAlign: Aligning nanopore reads for SV detection using current-level modeling 95%
- Variant Library Annotation Tool (VaLiAnT): an oligonucleotide library design and annotation tool for Saturation Genome Editing and other Deep Mutational Scanning experiments 94%
- Figbird: A probabilistic method for filling gaps in genome assemblies 94%
Similar papers in this journal
- MTG-Link: leveraging barcode information from linked-reads to assemble specific loci 95%
- PtWAVE: A High-Sensitive deconvolution software of sequencing trace for the Detection of Large Indels in Genome Editing 95%
- Identification and Utilization of Copy Number Information for Correcting Hi-C Contact Map of Cancer Cell Line 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.