Rethinking respiratory disease forecasting: temporal heterogeneity between surveillance predictors and outcomes drives forecast instability
Topazian, H. M.; Sheets, T. R.; Gruninger, R. J.; Kelley, J.; LaCross, N.; Samore, M. H.; Lofgren, E.; Keegan, L. T.
Show abstract
Since the COVID-19 pandemic, forecasting hubs and non-traditional respiratory disease surveillance streams have become increasingly common. However, many forecasting approaches assume that relationships between surveillance predictors and disease outcomes remain stable over time and that incorporating additional historical data will improve forecast performance. To evaluate these assumptions in a real-world setting, we developed and evaluated forecasts of SARS-CoV-2 and influenza hospitalizations in Utah using syndromic surveillance, test positivity, and wastewater data. Rather than identifying a single, best-performing model, we examined whether relationships between surveillance predictors and hospitalization outcomes remained stable across seasons and whether longer historical training periods consistently improved forecast accuracy. Relationships between surveillance predictors and hospitalizations varied substantially by pathogen and season. Analyses using pooled data across multiple years suggested strong positive correlations between predictors and outcomes, but these aggregated patterns often obscured weak or negative correlations observed during SARS-CoV-2 variant waves and influenza seasons. Forecast performance similarly varied over time. Models that performed well during some seasons, transmission phases, or under certain training strategies frequently performed worse than benchmark models in others. Training on additional historical data generally reduced forecast accuracy, though this varied by disease and transmission phase. Forecasting groups should prioritize continual evaluation of surveillance predictors, adaptive strategies, and diverse ensembles, rather than relying on a single model, data stream, or historical training framework each year.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Machine learning-based short-term forecasting of COVID-19 hospital admissions using routine hospital patient data 94%
- Assessing the utility of COVID-19 case reports as a leading indicator for hospitalization forecasting in the United States 92%
- A prospective real-time transfer learning approach to estimate Influenza hospitalizations with limited data 92%
Similar papers in this journal
- Infectious disease modeling for public health practice: projections, scenarios, and uncertainty in three phases of outbreak response 91%
- A quantitative framework to define the end of an outbreak: application to Ebola Virus Disease 89%
- Performance of Existing and Novel Symptom- and Antigen Testing-Based COVID-19 Case Definitions in a Community Setting 89%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.