BulkLMM: Real-time genome scans for multiple quantitative traits using linear mixed models
Yu, Z.; Farage, G.; Williams, R. W.; Broman, K. W.; Sen, S.
Show abstract
Genetic studies often collect data using high-throughput phenotyping. That has led to the need for fast genomewide scans for large number of traits using linear mixed models (LMMs). Computing the scans one by one on each trait is time consuming. We have developed new algorithms for performing genome scans on a large number of quantitative traits using LMMs, BulkLMM, that speeds up the computation by orders of magnitude compared to one trait at a time scans. On a mouse BXD Liver Proteome data with more than 35,000 traits and 7,000 markers, BulkLMM completed in a few seconds. We use vectorized, multi-threaded operations and regularization to improve optimization, and numerical approximations to speed up the computations. Our soft-ware implementation in the Julia programming language also provides permutation testing for LMMs and is available at https://github.com/senresearch/BulkLMM.jl.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Fast Numerical Optimization for Genome Sequencing Data in Population Biobanks 96%
- Efficient Permutation-based Genome-wide Association Studies for Normal and Skewed Phenotypic Distributions 96%
- Novel Approach for Parallelizing Pairwise Comparison Problems as Applied to Detecting Segments Identical By Decent in Whole-Genome Data 96%
Similar papers in this journal
Similar papers in this journal
- Interpretable Artificial Neural Networks incorporating Bayesian Alphabet Models for Genome-wide Prediction and Association Studies 95%
- A Multiple-trait Bayesian Variable Selection Regression Method for Integrating Phenotypic Causal Networks in Genome-Wide Association Studies 95%
- Comparing Heritability Estimators under Alternative Structures of Linkage Disequilibrium 95%
Similar papers in this journal
- Using encrypted genotypes and phenotypes for collaborative genomic analyses to maintain data confidentiality 96%
- Estimating SNP heritability in presence of population substructure in large biobank-scale data 96%
- Bayesian Hierarchical Hypothesis Testing in Large-Scale Genome-Wide Association Analysis 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.