Local Ancestry Prediction with PyLAE
Smetanin, A.; Moshkov, N.; Tatarinova, T. V.
Show abstract
SummaryWe developed PyLAE - a new tool for determining local ancestry along a genome using whole-genome sequencing data or high-density genotyping experiments. PyLAE can process an arbitrarily large number of ancestral populations (with or without an informative prior). Since PyLAE does not involve estimation of many parameters, it can process thousands of genomes within a day. Computational efficiency, straightforward presentation of results, and an ease of installation makes PyLAE a useful tool to study admixed populations. Availability and implementationThe source code and installation manual are available at https://github.com/smetam/pylae.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Benchmarking phasing software with a whole-genome sequenced cattle pedigree 95%
- A Nextflow pipeline for molecular quantitative trait loci mapping in small sample size datasets with an application in Atlantic salmon 93%
- Fine-Tuning GBS Data with Comparison of Reference and Mock Genome Approaches for Advancing Genomic Selection in Less Studied Farmed Species 93%
Similar papers in this journal
Similar papers in this journal
- VarSCAT: A computational tool for sequence context annotations of genomic variants 95%
- An assembly-free method of phylogeny reconstruction using short-read sequences from pooled samples without barcodes 95%
- Variant calling tool evaluation for variable size indel calling from next generation whole genome and targeted sequencing data 95%
Similar papers in this journal
- Poking COVID-19: insights on genomic constraints among immune-related genes between Qatari and Italian populations 94%
- More rule than exception: Parallel evidence of ancient migrations in grammars and genomes of Finno-Ugric speakers 92%
- The FORCE panel: An all-in-one SNP marker set for confirming investigative genetic genealogy leads and for general forensic applications 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.