Back

Analysis of Feature Influence on Covid-19 Death Rate Per Country Using a Novel Orthogonalization Technique

Gonnet, G.; Stewart, J.; Lafleur, J.; Keith, S.; McLellan, M.; Jiang-Gorsline, D.; Snider, T.

2021-07-05 health informatics
10.1101/2021.07.02.21259929 medRxiv
Show abstract

We have developed a new technique of Feature Importance, a topic of machine learning, to analyze the possible causes of the Covid-19 pandemic based on country data. This new approach works well even when there are many more features than countries and is not affected by high correlation of features. It is inspired by the Gram-Schmidt orthogonalization procedure from linear algebra. We study the number of deaths, which is more reliable than the number of cases at the onset of the pandemic, during Apr/May 2020. This is while countries started taking measures, so more light will be shed on the root causes of the pandemic rather than on its handling. The analysis is done against a comprehensive list of roughly 3,200 features. We find that globalization is the main contributing cause, followed by calcium intake, economic factors, environmental factors, preventative measures, and others. This analysis was done for 20 different dates and shows that some factors, like calcium, phase in or out over time. We also compute row explainability, i.e. for every country, how much each feature explains the death rate. Finally we also study a series of conditions, e.g. comorbidities, immunization, etc. which have been proposed to explain the pandemic and place them in their proper context. While there are many caveats to this analysis, we believe it sheds light on the possible causes of the Covid-19 pandemic. One-Sentence SummaryWe use a novel feature importance technique to find that globalization, followed by calcium intake, economic factors, environmental factors, and some aspects of societal quality are the main country-level data that explain early Covid-19 death rates.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

1
Scientific Reports
3102 papers in training set
Top 2%
14.5%
2
Frontiers in Artificial Intelligence
18 papers in training set
Top 0.1%
10.2%
3
PLOS Computational Biology
1633 papers in training set
Top 4%
7.3%
4
PLOS ONE
4510 papers in training set
Top 31%
4.9%
5
Mathematics
11 papers in training set
Top 0.1%
3.6%
6
Computers in Biology and Medicine
120 papers in training set
Top 0.8%
3.6%
7
Artificial Intelligence in Medicine
15 papers in training set
Top 0.1%
3.3%
8
Journal of Medical Internet Research
85 papers in training set
Top 2%
2.8%
50% of probability mass above
9
BMC Medical Informatics and Decision Making
39 papers in training set
Top 1%
2.5%
10
BMC Bioinformatics
383 papers in training set
Top 4%
2.1%
11
Cureus
67 papers in training set
Top 2%
1.9%
12
Physica A: Statistical Mechanics and its Applications
10 papers in training set
Top 0.1%
1.8%
13
Heliyon
146 papers in training set
Top 2%
1.7%
14
Patterns
70 papers in training set
Top 1%
1.5%
15
International Journal of Medical Informatics
25 papers in training set
Top 1.0%
1.5%
16
Expert Systems with Applications
11 papers in training set
Top 0.1%
1.5%
17
Communications Biology
886 papers in training set
Top 11%
1.5%
18
Bioinformatics
1061 papers in training set
Top 8%
1.5%
19
JAMIA Open
37 papers in training set
Top 1.0%
1.3%
20
Biology Methods and Protocols
53 papers in training set
Top 1%
1.3%
21
Chaos, Solitons & Fractals
32 papers in training set
Top 1%
1.3%
22
Vaccines
196 papers in training set
Top 2%
1.2%
23
Frontiers in Bioinformatics
45 papers in training set
Top 0.5%
1.1%
24
Frontiers in Applied Mathematics and Statistics
10 papers in training set
Top 0.3%
0.9%
25
Nature Communications
4913 papers in training set
Top 59%
0.9%
26
Computer Methods and Programs in Biomedicine
27 papers in training set
Top 0.8%
0.8%
27
Frontiers in Microbiology
375 papers in training set
Top 9%
0.8%
28
Physical Biology
43 papers in training set
Top 2%
0.8%
29
JMIR Public Health and Surveillance
45 papers in training set
Top 4%
0.8%
30
Informatics in Medicine Unlocked
21 papers in training set
Top 1%
0.8%