SimMiL: Simulating Microbiome Longitudinal Data
Weaver, N. E.; Hendricks, A.
Show abstract
0.Structured AbstractO_ST_ABSMotivationC_ST_ABSThe quantity of statistical tools designed for omics data analysis has grown rapidly with the ability to collect large sets of human health data, particularly longitudinal data sets. Most tools are assessed for performance using simulated datasets constructed to mimic a handful of relevant characteristics from real world data sets. Consequently, the simulated data sets, and their respective simulation frameworks, are too narrow in scope to qualify as a standard for assessment in longitudinal omics analyses. ResultsHere we present the flexible and accessible simulation framework and software package called SimMiL (Simulating Microbiome Longitudinal data) capturing three general components of longitudinal microbiome data: (i) absence/presence of microbes, (ii) individual microbe abundance, and (iii) microbiome community composition over time. The framework is assessed by replicating the Type I error and Power analyses of a broad range of statistical tools (MirKAT, repeated measures permANOVA, and a modified kernel association test). Software AvaliabilityThe simulation framework is at https://github.com/nweaver111/SimMiL
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- parafac4microbiome: Exploratory analysis of longitudinal microbiome data using Parallel Factor Analysis 95%
- Directional Gaussian Mixture Models of the gut microbiome elucidate microbial spatial structure 94%
- Simulation-based approaches to characterize the effect of sequencing depth on the quantity and quality of metagenome-assembled genomes 94%
Similar papers in this journal
- Feature selection and causal analysis for microbiome studies in the presence of confounding using standardization 95%
- Functional Analysis of Metagenomes by Likelihood Inference (FAMLI) Successfully Compensates for Multi-Mapping Short Reads from Metagenomic Samples 95%
- A negative binomial latent factor model for paired microbiome sequencing data 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.