SARS-CoV-2 has observably higher propensity to accept uracil as nucleotide substitution: Prevalence of amino acid substitutions and their predicted functional implications in circulating SARS-CoV-2 in India up to July, 2020
Roy, S.; Nath, H.; Mallick, A.; Biswas, S.
Show abstract
SARS-CoV-2 has emerged as pandemic all over the world since late 2019. In this study, we investigated the diversity of the virus in the context of SARS-CoV-2 spread in India. Full-length SARS-CoV-2 genome sequences of the circulating viruses from all over India were collected from GISAID, an open data repository, until 25thJuly, 2020. We have focused on the non-synonymous changes across the genome that resulted in amino acid substitutions. Analysis of the genomic signatures of the non-synonymous mutations demonstrated a strong association between the time of sample collection and the accumulation of genetic diversity. Most of these isolates from India belonged to the A2a clade (63.4%) which has overcome the selective pressure and is spreading rapidly across several continents. Interestingly a new clade I/A3i has emerged as the second-highest prevalent type among the Indian isolates, comprising 25.5% of the Indian sequences. Emergence of new mutations in the S protein was observed. Major SARS-CoV-2 clades in India have defining mutations in the RdRp. Maximum accumulation of mutations was observed in ORF1a. Other than the clade-defining mutations, few representative non-synonymous mutations were checked against the available crystal structures of the SARS-CoV-2 proteins in the DynaMut server to assess their thermodynamic stability. We have observed that SARS-CoV-2 genomes contain more uracil than any other nucleotide. Furthermore, substitution of nucleotides to uracil was highest among the non-synonymous mutations observed. The A+U content in SARS-CoV-2 genome is much higher compared to other RNA viruses, suggesting that the virus RdRp has a propensity towards uracil incorporation in the genome. This implies that thymidine analogues may have a better chance to competitively inhibit SARS-CoV-2 RNA replication than other nucleotide analogues.
Matching journals
The top 10 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- In silico comparative genomics of SARS-CoV-2 to determine the source and diversity of the pathogen in Bangladesh 98%
- Whole Genome Sequencing Analysis of Spike D614G Mutation Reveals Unique SARS-CoV-2 Lineages of B.1.524 and AU.2 in Malaysia 97%
- Comparative Genomic Study for Revealing the Complete Scenario of COVID-19 Pandemic in Bangladesh 97%
Similar papers in this journal
- Mutational spectra of SARS-CoV-2 orf1ab polyprotein and Signature mutations in the United States of America 97%
- Genomic diversity of SARS-CoV-2 in Pakistan during fourth wave of pandemic 96%
- Dynamic tracking of variant frequencies depicts the evolution of mutation sites amongst SARS-CoV-2 genomes from India 95%
Similar papers in this journal
Similar papers in this journal
- Characterizations of SARS-CoV-2 mutational profile, spike protein stability and viral transmission 97%
- Global variation in the SARS-CoV-2 proteome reveals the mutational hotspots in the drug and vaccine candidates 95%
- Genome based Evolutionary study of SARS-CoV-2 towards the Prediction of Epitope Based Chimeric Vaccine 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.