Genomic demography of world's ethnic groups and their genomic identity between two individuals
Kim, B.-J.; Choi, J.; Kim, S.-H.
Show abstract
All current categorizations of human population, such as ethnicity, ancestry and race, are based on various selections and combinations of subjectively- and/or qualitatively-defined characteristics, such as ancestral lineage/location, cultural/societal norm, language, skin color and other phenotypes and traits perceived by the members within or from outside of the categorized group. Yet, such categorization has been broadly used also in the fields of human genetics, health sciences and medical practices (e.g., 1,2,3), where the observed health characteristics are objectively and quantitatively definable, but the population categorization is not yet available. Here we show the feasibility of deriving a whole-genome-based categorization that is objectively definable and quantitatively measurable. We observe that: (a) the worlds ethnic populations form about 14 genomic groups (GGs); (b) each GG consists of multiple ethnic groups (EGs); and (c) at an individual level, approximately 99.8%, on average, of the whole genome contents are identical between any two individuals regardless of their GGs or EGs.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Assessing human genome-wide variation in the Massim region of Papua New Guinea and implications for the Kula trading tradition 96%
- Network-based analysis of allele frequency distribution among multiple populations identifies adaptive genomic structural variants 94%
- Mozambican genetic variation provides new insights into the Bantu expansion 94%
Similar papers in this journal
- Joint effects of balancing selection and population bottlenecks on the evolution of a regulatory region of human anti-viral APOBEC3 94%
- The impact of natural selection on the evolution and function of placentally expressed galectins 93%
- Tracing the evolution of human gene regulation and its association with shifts in environment 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.