Back

Python-based automation of INDIGO webserver using Selenium: A high throughput analysis of Sanger sequence data to detect allelic variations created by CRISPR/Cas9-mediated genome editing of crop plants

Suresh, V.; Girish, C.; Tavva, V. S. S.

2025-08-05 bioinformatics
10.1101/2025.08.04.668409 bioRxiv
Show abstract

CRISPR-Cas9 has revolutionized plant genome editing by enabling precise introduction of insertion/deletion (indel) mutations, critical for functional genomics and crop improvement studies. Sanger sequencing, combined with bioinformatics tools like the INDIGO webserver from Gear Genomics, is essential for validating these mutations. However, manual analysis of large numbers of Sanger sequencing (.ab1) files is labor-intensive, particularly when analyzing multiple guide RNA (gRNA) target regions. We developed a Python-based automation pipeline using Selenium with integrated highlighting of protospacer adjacent motif (PAM) regions in the resulting HTML reports. This pipeline enhances scalability of Sanger sequence data analysis and improves result interpretability by automating PAM region identification and supporting multiple gRNA regions. This tool significantly accelerates CRISPR-Cas9-mediated mutation analysis, offering a high throughput, reproducible solution for genome editing research.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.