Neural posterior estimation for population genetics
Min, J.; Ning, Y.; Pope, N. S.; Baumdicker, F.; Kern, A. D.
Show abstract
Simulation-based inference methods are increasingly being used in population genetics due to their flexibility and ability to be applied in settings where likelihood-based methods are intractable. Perhaps the best known such method is Approximate Bayesian Computation (ABC); however, its popularity is offset by its shortcomings which include computational expense and an unfortunate inability to efficiently fit models to high-dimensional summaries of the data. An alternative approach that solves these issues is supervised machine learning (ML); however, ML methods generally do not yield Bayesian uncertainty estimates of the quantities they predict. Here, we apply a recently introduced method, neural posterior estimation (NPE), that combines the best facets of ABC and supervised ML by training a neural network to estimate the posterior distribution of a population genetics model. We first compare neural posterior estimation with other inference methods for a variety of population genetic tasks, and show that neural posterior estimators yield posterior distributions with high accuracy and efficiency. We compare learned posterior distributions given raw genotypes and various summary statistics as input data. Additionally, we apply neural posterior estimation for demographic inference for simple and more complex models to highlight its application, including an analysis of demographic history in Drosophila melanogaster. Finally, we provide a user friendly workflow that enables others to perform neural posterior estimation on their own genetic data.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Tree sequences as a general-purpose tool for population genetic inference 97%
- Computationally efficient demographic history inference from allele frequencies with supervised machine learning 97%
- Fast and accurate estimation of selection coefficients and allele histories from ancient and modern DNA 97%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.