peakScout - a user-friendly and reversible peak-to-gene translator for genomic peak calling results
Lin, A. L.; Cartailler, L. A.; Cartailler, J.-P.
Show abstract
SummarypeakScout is a command line and web-based bioinformatics tool designed to quickly and easily bridge the gap between genomic peak data and gene annotations, enabling researchers to understand the relationship between measurements of regulatory elements and their target genes. At its core, peakScout processes genomic peak files obtained through various means chromatin profiling and maps them to nearby genes using reference genome annotations. The workflow begins with input processing, where peak files are standardized and reference GTF files are decomposed into chromosome-specific feature collections. The core analysis modules then perform bidirectional mapping: peak-to-gene identifies which genes are potentially regulated by specific genomic regions, while gene-to-peak reveals which regulatory elements might influence particular genes of interest. Throughout this process, nearest-feature detection algorithms handle the complex spatial relationships between genomic elements, considering factors like distance constraints and feature overlaps. Finally, the results are formatted into researcher-friendly CSV and Excel outputs, providing a comprehensive view of the genomic landscape that connects regulatory elements to their potential gene targets. Availability and implementationThe web version of peakScout is available at https://vandydata.github.io/peakScout/. The command line version is available at https://github.com/vandydata/peakScout and archived on Zenodo (URL to be provided upon version 1.0 release) under the GNU Affero General Public License v3.0. Installation instructions, example datasets, and detailed usage examples are provided in the GitHub repository README file. peakScout is implemented in Python and is platform independent, but the web version is implemented in Amazon Web Services and thus uses proprietary infrastructure.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- txtools: an R package facilitating analysis of RNA modifications, structures, and interactions 95%
- edgeR v4: powerful differential analysis of sequencing data with expanded functionality and improved support for small counts and larger datasets 94%
- Echtvar: Compressed variant representation for rapid annotation and filtering of SNPs and indels 93%
Similar papers in this journal
- AnnSQL: A Python SQL-based package for fast large-scale single-cell genomics analysis using minimal computational resources 96%
- RNApysoforms: Fast rendering interactive visualization of RNA isoform structure and expression in Python 95%
- tinyRNA: precision analysis of small RNA-seq data with user-defined hierarchical selection rules 95%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.