traceax: a JAX-based framework for stochastic trace estimation
Nahid, A. A.; Serafin, L.; Mancuso, N.
Show abstract
In many applications, from statistical inference to machine learning, calculating the trace of a matrix is a fundamental operation, yet may be infeasible due to memory constraints. Stochastic trace estimation offers a practical solution by using randomized matrix-vector products to obtain accurate, unbiased estimates without constructing the full matrix in memory. Here, we present traceax, a Python framework for scalable trace estimation that leverages efficient linear operator representations of matrices while supporting automatic differentiation and hardware acceleration. traceax supports state-of-the-art trace estimators and through simulations we recapitulate results demonstrating their high accuracy while significantly reducing runtime and memory usage when compared with direct trace computation. As a proof of concept, we implemented a stochastic heritability estimator using traceax requiring only several lines of code. Overall, traceax provides a versatile tool for stochastic trace estimation that can be easily integrated into existing inferential pipelines. traceax is freely available at: https://github.com/mancusolab/traceax
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- K2R: Tinted de Bruijn Graphs implementation for efficient read extraction from sequencing datasets 93%
- HTRX: an R package for learning non-contiguous haplotypes associated with a phenotype 93%
- AnnSQL: A Python SQL-based package for fast large-scale single-cell genomics analysis using minimal computational resources 93%
Similar papers in this journal
- CoGAPS 3: Bayesian non-negative matrix factorization for single-cell analysis with asynchronous updates and sparse data structures 96%
- MGMM: An R Package for fitting Gaussian Mixture Models on Incomplete Data 94%
- HARVESTMAN: A framework for hierarchical featurelearning and selection from whole genome sequencingdata 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.