Back

Analyzing hCov genome sequences: Applying Machine Intelligence and beyond

Sawmya, S.; Saha, A.; Tasnim, S.; Anjum, N.; Toufikuzzaman, M.; Rafid, A. H. M.; Rahman, M. S.; Rahman, M. S.

2020-06-03 bioinformatics
10.1101/2020.06.03.131987 bioRxiv
Show abstract

BackgroundCovid-19 pandemic, caused by the SARS-CoV-2 genome sequence of coronavirus, has affected millions of people all over the world and taken thousands of lives. It is of utmost importance that the character of this deadly virus be studied and its nature be analyzed. MethodsWe present here an analysis pipeline comprising a classification exercise to identify the virulence of the genome sequences and extraction of important features from its genetic material that are used subsequently to predict mutation at those interesting sites using deep learning techniques. ResultsWe have classified the SARS-CoV-2 genome sequences with high accuracy and predicted the mutations in the sites of Interest. ConclusionsIn a nutshell, we have prepared an analysis pipeline for hCov genome sequences leveraging the power of machine intelligence and uncovered what remained apparently shrouded by raw data.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.