Back

Regression Analysis of COVID-19 Spread in India and its Different States

Chauhan, P.; Kumar, A.; Jamdagni, P.

2020-05-29 epidemiology
10.1101/2020.05.29.20117069 medRxiv
Show abstract

Linear and polynomial regression model has been used to investigate the COVID-19 outbreak in India and its different states using time series epidemiological data up to 26th May 2020. The data driven analysis shows that the case fatality rate (CFR) for India (3.14% with 95% confidence interval of 3.12% to 3.16%) is half of the global fatality rate, while higher than the CFR of the immediate neighbors i.e. Bangladesh, Pakistan and Sri Lanka. Among Indian states, CFR of West Bengal (8.70%, CI: 8.21-9.18%) and Gujrat (6.05%, CI: 4.90-7.19%) is estimated to be higher than national rate, whereas CFR of Bihar, Odisha and Tamil Nadu is less than 1%. The polynomial regression model for India and its different states is trained with data from 21st March 2020 to 19th May 2020 (60 days). The performance of the model is estimated using test data of 7 days from 20th May 2020 to 26th May 2020 by calculating RMSE and % error. The model is then used to predict number of patients in India and its different states up to 16th June 2020 (21 days). Based on the polynomial regression analysis, Maharashtra, Gujrat, Delhi and Tamil Nadu are continue to remain most affected states in India.

Matching journals

The top 9 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.