Back

Leveraging machine learning and self-administered tests to predict COVID-19:An olfactory and gustatory dysfunction assessment through crowd-sourced data in India

Kumar, R.; Singh, M.; Singh, P.; Parma, V.; Ohla, K.; Olsson, S. B.; Saini, V.; Rani, J.; Kishore, K.; Kumari, P.; Ichhpujani, P.; Sharma, A.; Kumar, S.; Sharma, M.; Bhondekar, A. P.; Kothari, A.; Sardana, V.; Iyengar, S.; Dash, D.; Kaur, R.

2021-10-26 infectious diseases
10.1101/2021.10.20.21265247 medRxiv
Show abstract

It has been established that smell and taste loss are frequent symptoms during COVID-19 onset. Most evidence stems from medical exams or self-reports. The latter is particularly confounded by the common confusion of smell and taste. Here, we tested whether practical smelling and tasting with household items can be used to assess smell and taste loss. We conducted an online survey and asked participants to use common household items to perform a smell and taste test. We also acquired generic information on demographics, health issues including COVID-19 diagnosis, and current symptoms. We developed several machine learning models to predict COVID-19 diagnosis. We found that the random forest classifier consistently performed better than other models like support vector machines or logistic regression. The smell and taste perception of self-administered household items were statistically different for COVID-19 positive and negative participants. The most frequently selected items that also discriminated between COVID-19 positive and negative participants were clove, coriander seeds, and coffee for smell and salt, lemon juice, and chillies for taste. Our study shows that the results of smelling and tasting household items can be used to predict COVID-19 illness and highlight the potential of a simple home-test to help identify the infection and prevent the spread.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.