Forecastability of infectious disease time series: are some seasons and pathogens intrinsically more difficult to forecast?
White, L. A.; Leon, T. M.
Show abstract
For infectious disease forecasting challenges, individual model performance typically varies across space and time. This phenomenon raises the question: are there properties of the target time series that contribute to a particular season, location, or disease being more difficult to forecast? Here we characterize a time series future predictability using a forecastability metric that calculates the spectral density of the time series. Forecastability of syndromic influenza hospital admissions for the state of California varied widely across seasons and was positively correlated with peak burden. Next, using archived U.S. state and national forecasts targeting laboratory-confirmed COVID-19 and influenza hospital admissions, we investigated the relationship between forecastability and: (i) population size of the forecasting target, and (ii) forecast performance as measured by mean absolute error, weighted interval score (WIS), and scaled relative WIS. Forecastability increased with increasing population size of the forecasting target, and forecasting performance generally improved with higher forecastability when controlling for population size across scales. These preliminary results support the idea that some targets and respiratory virus seasons may be inherently more difficult to forecast and could help explain inter-seasonal variation in model performance. Author summaryCould intrinsic properties of an epidemiological time series help explain why a particular season, location, or disease is more difficult to predict in the future? To answer this question, this analysis uses a measure of a time series future predictability called "forecastability," which describes the inherent uncertainty or surprise in the signal. Influenza and COVID-19 hospital admissions had higher forecastability scores for locations with larger population sizes, possibly due to larger counts leading to smoother time series. At the same time, forecasting performance generally improved for time series with higher forecastability scores when controlling for population size, suggesting that this metric is helpful for understanding ease of forecasting. These preliminary results support the idea that some epidemiological targets and respiratory virus seasons may be inherently more difficult to forecast and could help explain why forecasting model performance changes across different respiratory virus seasons.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Stacked ensemble method for forecasting influenza-like illness visit volumes at emergency departments 93%
- Early Detection of COVID-19 Outbreaks Using Human Mobility Data 92%
- Inclusion of environmentally themed search terms improved Elastic Net regression nowcasts of regional Lyme disease rates 92%
Similar papers in this journal
- Performance of Existing and Novel Symptom- and Antigen Testing-Based COVID-19 Case Definitions in a Community Setting 92%
- Adjusting COVID-19 seroprevalence survey results to account for test sensitivity and specificity 91%
- A quantitative framework to define the end of an outbreak: application to Ebola Virus Disease 90%
Similar papers in this journal
Similar papers in this journal
- Accuracy of US CDC COVID-19 Forecasting Models 94%
- AI-Driven Early Detection of Severe Influenza in Jiangsu, China: A Deep Learning Model Validated Through The Design of Multi-Center Clinical Trials and Prospective Real-World Deployment 91%
- Assessing COVID-19 vaccination strategies in varied demographics using an individual-based model 90%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.