fastman: A fast algorithm for visualizing GWAS results using Manhattan and Q-Q plots
Paria, S. S.; Rahman, S. R.; Adhikari, K.
Show abstract
Visualization of GWAS summary statistics, specifically P-values, as Manhattan plots is widespread in GWAS publications, and many popular software tools are available, such as the R package qqman. But there is substantial need for further development, such as the handling of non-human data. We provide a new R package, fastman, with major additional capabilities. It handles genomes of non-model organisms, even those at a draft stage, i.e. contigs that havent been compiled to chromosomes. Non-numeric chromosome IDs are supported. It supports plotting of other genetic scores, such as FST, D statistics, selection statistics such as PBS, or other kinds of GWAS statistics such as beta. Importantly, negative or two-tailed values are supported in this package. We implement a heuristic algorithm that drastically reduces plotting time for huge datasets without any loss of visual precision, while allowing for many different data types and missing data. We provide substantial additional flexibility in highlighting and annotation. In summary, we have developed a package fastman in R for fast and efficient visualization of GWAS results and other genomewide scores using Manhattan and Q-Q plots. The package can create plots directly from association outputs by PLINK. Alternatively, it can produce plots from any R data frame with custom columns and is equipped to handle big datasets with fast plot generation. It is available for public use on https://github.com/kaustubhad/fastman.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- CNVpytor: a tool for CNV/CNA detection and analysis from read depth and allele imbalance in whole genome sequencing 95%
- SnpHub: an easy-to-set-up web server framework for exploring large-scale genomic variation data in the post-genomic era with applications in wheat 95%
- DivBrowse - interactive visualization and exploratory data analysis of variant call matrices 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.