SigRescueR: A Pan-System Framework for Noise Correction and Mutational Signature Identification Across Sequencing Platforms
Nguyen, P.; Zhivagui, M.
Show abstract
Mutational signatures serve as molecular fingerprints of the biological processes and exposures that shape cancer genomes. However, accurate signal recovery remains challenging due to pervasive background variants, sequencing artifacts, technical noise, and platform-specific biases that obscure true mutagenic patterns, hampering biomarker discovery and mechanistic interpretation. Here we introduce SigRescueR, a rigorous, pan-system, computational framework designed for noise correction and mutational signature identification. SigRescueR applies statistically robust baseline correction to effectively disentangle true mutational signals from confounding noise and artifacts. When applied to extensive datasets spanning experimental models and human cancers, SigRescueR reliably identified canonical mutational signatures associated with environmental mutagens such as colibactin, benzo[a]pyrene, and UV radiation, and chemotherapeutic agents, namely 5-fluorouracil and cisplatin. SigRescueR effectively operated across diverse mutation classes, including single base substitutions, insertions and deletions, and doublet base substitutions, while also integrating strand bias and duplex sequencing data for toxicology applications. SigRescueR offers a unified, high-precision platform that seamlessly integrates cancer genomics, molecular toxicology, and mechanistic studies. It enables precise mapping of mutagenic processes and identification of robust genomic biomarkers of environmental and therapeutic exposures, providing a transformative framework for translational cancer research. Availability and implementationSigRescueR is implemented in R and provided as open-source software on GitHub at https://github.com/ZhivaguiLab/SigRescueR/
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A benchmark of computational methods for correcting biases of established and unknown origin in CRISPR-Cas9 screening data 95%
- DiMSum: an error model and pipeline for analyzing deep mutational scanning data and diagnosing common experimental pathologies 94%
- BASCULE: Bayesian inference and clustering of mutational signatures leveraging biological priors 94%
Similar papers in this journal
- Highly accurate barcode and UMI error correction using dual nucleotide dimer blocks allows direct single-cell nanopore transcriptome sequencing 93%
- A pan-cancer landscape of somatic substitutions in non-unique regions of the human genome 93%
- Multi-resolution deconvolution of spatial transcriptomics data reveals continuous patterns of inflammation 93%
Similar papers in this journal
- Sequence dependencies and mutation rates of localized mutational processes in cancer 95%
- Pan-cancer detection of driver genes at the single-patient resolution 92%
- Nanopore sequencing with unique molecular identifiers enables accurate mutation analysis and haplotyping in the complex Lipoprotein(a) KIV-2 VNTR 92%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.