Evolution profile of 13415 SNVs in 33 language/cognition genes measured by five types of distance calculation
Zhang, Z.; Xu, Y.
Show abstract
This study aims to quantify the genetic similarity of different species (from fish to humans) to the human reference genome (pp6, Homo sapiens.GRCh38) based on the allele presence/absence patterns of 33 language/cognition related gene SNV loci, identify key breakpoints during evolution, and evaluate the enrichment of language and cognition genes at these breakpoints. We designed a similarity calculation method relying on binary features (four columns for A/T/C/G), adopted five difference/distance measures (Sorensen, Rogers, Nei, Reynolds, and Hellinger), and converted them into similarity values (1/(1+distance)). For each method, samples were independently ranked, the first derivative of similarity was computed, and the top 12 peaks were selected as candidate breakpoints. Results show that the similarity curves from the five methods are highly consistent (correlation coefficients >0.9), with major peaks concentrated at positions 355, 363, 381, 382, 390, 400, etc., where the corresponding samples are predominantly ancient hominins and primates. Furthermore, we defined 13 peak groups (starting positions 355-401). For each peak within a group, pairwise SNV differences between the peak apex sample and its immediate left neighbor were compared, and the intersection F_INTERSECTION (shared differential loci) was obtained. For each F_INTERSECTION, we calculated the proportions of language genes and cognition genes. In addition, we computed the differential sets between adjacent groups' F_INTERSECTION to trace the gradual emergence of new loci. In F_INTERSECTION, language genes accounted for an average of 59.5%, and cognition genes for an average of 62.9%. The proportion of language genes reached a peak at position 383 (61.2%), while cognition genes peaked at position 386 (64.9%). High frequency peak samples include c25, c27, and ja2, suggesting that language cognition genes may have undergone independent intensification during Eurasian evolution. Differential analysis between adjacent F_INTERSECTION revealed a stepwise acquisition of new loci from position 355 to 401, with three bursts of newly added loci along the entire evolutionary axis. This study provides a quantitative framework based on similarity curves, offers a novel molecular perspective for understanding the evolution of language and cognitive abilities, and highlights the potential importance of East Asian archaic hominins in the evolution of language cognition genes.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Comparative analysis of corrected tiger genome provides clues to their neuronal evolution 93%
- Recurrent erosion of COA1/MITRAC15 demonstrates gene dispensability in oxidative phosphorylation 91%
- Inner Asian maternal genetic origin of the Avar period nomadic elite in the 7th century AD Carpathian Basin 91%
Similar papers in this journal
- Mitochondrial Genome Diversity in the Central Siberian Plateau with Particular Reference to Prehistory of Northernmost Eurasia 93%
- Investigating demic versus cultural diffusion and sex bias in the spread of Austronesian languages in Vietnam 91%
- Lower promoter activity of the ST8SIA2 gene has been favored in evolving human collective brains 90%
Similar papers in this journal
- Time-series analyses of directional sequence changes in SARS-CoV-2 genomes and an efficient search method for advantageous mutations for growth in human cells 90%
- Codon usage bias in radioresistant bacteria 89%
- A unique neurogenomic state emerges after aggressive confrontations in males of the fish Betta splendens 88%
Similar papers in this journal
- Cross-species alignment along the chronological axis reveals evolutionary effect on structural development of human brain 90%
- The whale shark genome reveals patterns of vertebrate gene family evolution 89%
- Ancient DNA reveals the lost domestication history of South American camelids in Northern Chile and across the Andes 89%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.