Genetic diversity analysis of the D614G mutation in SARS-CoV-2
Felix, P. T.; Silva, E. D. A. B.; Venancio, D. B. R.; Ramos, R. d. S.
Show abstract
In this work, we evaluated the levels of genetic diversity in 18 genomes of SARS-CoV-2 carrying the D614G mutation, coming from Malaysia and Venezuela and publicly available at the National Center of Biotechnology and Information (NCBI). These haplotypes were previously used for phylogenetic analysis, following the LaBECom protocols. All gaps and unconserved sites were extracted for the construction of a phylogenetic tree. As specific methodologies for paired FST estimators, Molecular Variance (AMOVA), Genetic Distance, mismatch, demographic and spatial expansion analyses, molecular diversity and evolutionary divergence time analyses, 20,000 random permutations were always used. The results revealed the presence of only 57 sites of polymorphic and parsimonium-informative among the 29,827bp analyzed and the analyses based on FST values confirmed the presence of two distinct genetic entities with fixation index of 22% and with a higher component of population variation (78.14%). Tau variations revealed a significant time of divergence, supported by mismatch analysis of the observed distribution ({tau} = 42%). It is safe to say that the small number of existing polymorphisms should not reflect major changes in the protein products of viral populations in both countries and this consideration provides the safety that, although there are differences in the haplotypes studied, these differences are minimal for both regions analyzed geographically and, therefore, it seems safe to extrapolate the levels of polymorphism and molecular diversity found in the samples for other mutant genomes of SARS-CoV-2 in other countries. This reduces speculation about the possibility of large differences between mutant strains of SARS-CoV-2 (D614G) and wild strains, at least at the level of their protein products, although the mutant form has higher transmission speed and infection. The analyses suggest that possible variations in protein products, of the wild virus in relation to its mutant form, should be minimal, bringing peace of mind as to the increased risk of death from the new form of the virus, as well as possible problems of gradual adjustments in some molecular targets for vaccines.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genetic analyses reveal population structure and recent decline in leopards (Panthera pardus fusca) across Indian subcontinent 94%
- Application of High Resolution Melt analysis (HRM) for screening haplotype variation in non-model plants: a case study of Honeybush (Cyclopia Vent.) 93%
- DiscoSnp-RAD: de novo detection of small variants for population genomics 91%
Similar papers in this journal
- Sample Size Impact (SaSii): an R script for estimating optimal sample sizes in population genetics and population genomics studies 94%
- metaVaR: introducing metavariant species models for reference-free metagenomic-based population genomics 93%
- Linkage disequilibrium and haplotype block patterns in popcorn populations 93%
Similar papers in this journal
- Inference of the worldwide invasion routes of the pinewood nematode Bursaphelenchus xylophilus using approximate Bayesian computation analysis 93%
- Better confidence intervals in simulation-based inference 92%
- Spontaneous parthenogenesis in the parasitoid wasp Cotesia typhae: low frequency anomaly or evolving process? 92%
Similar papers in this journal
- Integrating population genetics to define conservation units from the core to the edge of Rhinolophus ferrumequinum western range 93%
- Historical demography and species distribution models shed light on past speciation in primates of northeast India 92%
- A highly divergent mitochondrial genome in extant Cape buffalo from Addo Elephant National Park, South Africa 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.