Advanced pipeline for CRISPR/Cas9 offtargets detection in Guide-seq and related integration-based assays
Corre, G.; Rouillon, M.; Mombled, M.; Amendola, M.
Show abstract
BackgroundThe advent of CRISPR-Cas9 genome editing has brought about a paradigm shift in molecular biology and gene therapy. However, the persistent challenge of off-target effects continues to hinder its therapeutic applications. Unintended genomic alterations can lead to significant genomic damage, thereby compromising the safety and efficacy of CRISPR-based therapies. Although in-silico prediction tools have made substantial progress, they are not sufficient for capturing the complexity of genomic alterations and experimental validation remains crucial for accurate identification and quantification of off-target effects. In this context, Genome-wide Unbiased Identification of Double-strand breaks Enabled by Sequencing (GUIDE-Seq) has emerged as a gold standard method for the experimental detection of off-target sites and assessment of their prevalence by introducing short double-stranded oligonucleotides (dsODNs) at the break sites created by the nuclease. The bioinformatic analysis of GUIDE-Seq data plays a pivotal yet challenging role in accurately mapping and interpreting editing sites and current pipelines suffer limitations we aim to address in this work. ResultsIn this study, we present a rapid and versatile single-command pipeline designed for the comprehensive analysis of GuideSeq and similar techniques of sequencing. Our pipeline is capable of simultaneously processing multiplexed libraries from different organisms, PCR orientations, and Cas with different PAM specificities in a single run, all based on user-specified sample information. To ensure reproducibility, the pipeline operates within a closed environment and incorporates a suite of well-established bioinformatics tools. Key novel features include the ability to manage bulges in gDNA/gRNA interaction and multi-hit reads, and a built-in tool for off-target site prediction. The pipeline generates a detailed report that consolidates quality control metrics and provides a curated list of off-target candidates along with their corresponding gRNA alignments. ConclusionsOur pipeline has been tested and successfully applied to analyze samples under a variety of experimental conditions, including different source organisms, PAM motifs, dsODN sequences and PCR orientations. The robustness and flexibility of our pipeline make it a valuable tool for researchers in the field of genome editing. The source code and comprehensive documentation are freely accessible on our GitHub repository: https://github.com/gcorre/GNT_GuideSeq.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Unlocking the Full Potential of Nanopore Sequencing: Tips, Tricks, and Advanced Data Analysis Techniques 95%
- Dysgu: efficient structural variant calling using short or long reads 95%
- REVERSE: A user-friendly web server for analyzing next-generation sequencing data from in vitro selection/evolution experiments 95%
Similar papers in this journal
- miRge3.0: a comprehensive microRNA and tRF sequencing analysis pipeline 96%
- FLYNC: A Machine Learning-Driven Framework for Discovering Long Non-Coding RNAs in Drosophila melanogaster 95%
- Kmerator Suite: design of specific k-mer signatures andautomatic metadata discovery in large RNA-Seq datasets. 94%
Similar papers in this journal
- wg-blimp: an end-to-end analysis pipeline for whole genome bisulfite sequencing data 96%
- PIPETS: A statistically informed, gene-annotation agnostic analysis method to study bacterial termination using 3'-end sequencing. 95%
- PtWAVE: A High-Sensitive deconvolution software of sequencing trace for the Detection of Large Indels in Genome Editing 95%
Similar papers in this journal
- Variant Library Annotation Tool (VaLiAnT): an oligonucleotide library design and annotation tool for Saturation Genome Editing and other Deep Mutational Scanning experiments 93%
- Genome size estimation from long read overlaps 93%
- AnthOligo: Automating the design of oligonucleotides for capture/enrichment technologies 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.