Challenges in Estimating Time-Varying Epidemic Severity Rates from Aggregate Data
Goldwasser, J.; Hu, A.; Bilinski, A.; McDonald, D. J.; Tibshirani, R.
Show abstract
Severity rates like the case-fatality rate and infection-fatality rate are key metrics in public health. To guide decision-making in response to changes like new variants or vaccines, it is imperative to understand how these rates shift in real time. In practice, time-varying severity rates are typically estimated using a ratio of aggregate counts. We demonstrate that these estimators are capable of exhibiting large statistical biases, with concerning implications for public health practice, as they may fail to detect heightened risks or falsely signal nonexistent surges. We supplement our mathematical analyses with experimental results on real and simulated COVID-19 data. Finally, we briefly discuss strategies to mitigate this bias, drawing connections with effective reproduction number (Rt) estimation.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Probabilistic Cause-of-disease Assignment using Case-control Diagnostic Tests: A Latent Variable Regression Approach 96%
- A Double Machine Learning Approach for the Evaluation of COVID-19 Vaccine Effectiveness under the Test-Negative Design: Analysis of Québec Administrative Data 96%
- Estimation of Vaccine Efficacy for Variants that Emerge After the Placebo Group Is Vaccinated 94%
Similar papers in this journal
Similar papers in this journal
- Adjusting for time of infection or positive test when estimating the risk of a post-infection outcome in an epidemic 95%
- Tight Fit of the SIR Dynamic Epidemic Model to Daily Cases of COVID-19 Reported During the 2021-2022 Omicron Surge in New York City: A Novel Approach 95%
- A bivariate zero-inflated negative binomial model and its applications to biomedical settings 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.