SAGA (Simplified Association Genomewide Analyses): a user-friendly Pipeline to Democratize Genome-Wide Association Studies
Cieza, B.; Pandey, N.; Ruhela, V.; Ali, S.; tosto, g.
Show abstract
Genome-wide association studies (GWAS) have enabled clinicians and researchers to identify genetic variants linked to complex traits and diseases(1-3). However, GWAS still face several challenges, particularly regarding accessibility and reproducibility (4-6). Conducting these analyses often requires substantial bioinformatics expertise for data preprocessing, software installation, and scripting(7-10). We then developed SAGA ("Simplified Association Genome-wide Analyses"), a BASH-based, open-source, fully automated pipeline that integrates three widely adopted tools--PLINK(11), GMMAT(12), and SAIGE(13)--for accessible, robust, and reproducible GWAS. After installation, users simply need to provide genotype and phenotype files in standard formats. The pipeline automates preprocessing, association testing, and visualization, outputting summary statistics, Manhattan plots, and quantile-quantile plots. SAGA enables robust GWAS for users without scripting experience, expanding access to complex genetic analyses.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- kTWAS: integrating kernel-machine with transcriptome-wide association studies improves statistical power and reveals novel genes 93%
- xQTLbiolinks: a comprehensive and scalable tool for integrative analysis of molecular QTLs 93%
- BayesKAT: Bayesian Optimal Kernel-based Test for genetic association studies reveals joint genetic effects in complex diseases 92%
Similar papers in this journal
- RegionScan: A comprehensive R package for region-level genome-wide association testing with integration and visualization of multiple-variant and single-variant hypothesis testing 93%
- ntRoot: Computational Inference of Human Ancestry at Scale from Genomic Data 92%
- VAREANT : a bioinformatics application for gene variant reduction and annotation 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.