The tale of two assumptions: incorporating healthcare-seeking behaviour in epidemic forecasting
Bibliographic record
Abstract
BACKGROUND: Modelling efforts during the COVID-19 pandemic highlighted the importance of incorporating human behaviour into mathematical models and the challenges of making accurate forecasts. Case detection is affected by different healthcare-seeking behaviors, including visiting physicians to seek help, which can impact the number of laboratory tests performed and the number of cases identified by surveillance systems throughout an epidemic. Mathematical models for forecasting epidemics of respiratory viruses such as influenza and COVID-19 generally assume a constant rate of case detection, and only a few studies have previously used time-dependent rates. PURPOSE: This study aims to compare constant and time-dependent case detection rate approaches for the forecast and retrospective fitting of seasonal influenza data. METHODS: An age-stratified Susceptible-Infected-Removed (SIR) model that incorporates case detection for influenza is formulated. Influenza case data and case detection rates for the 2016–2019 seasons in Alberta, Canada, are used for model training. The model fitting results are compared for the constant and time-dependent case detection assumptions. The model forecasting results using partial-season data are compared to the data for the remainder of the season for validation. RESULTS: While both constant and time-dependent case detection rate assumptions allowed an accurate retrospective fitting to the case data of an entire season, the forecasting performance showed a significant difference between the two assumptions. Models with a time-dependent case detection rate accurately predicted the influenza peak time four weeks before the actual peak occurred. The average total infections per case detected, an estimate that includes both under-ascertainment and underreporting, also showed a significant difference between the two assumptions. CONCLUSION: The incorporation of healthcare-seeking behaviour in mathematical modelling helps quantify the dynamic process of how infections are detected by surveillance systems. This is an important consideration for influenza forecasting. Since not all individuals engage in healthcare-seeking behaviour and only a fraction of those who do seek help may get tested, a proportion of infections remain undetected by the surveillance system. The retrospective forecasting results highlight that a time-dependent case detection rate is more representative of changes in healthcare-seeking behaviour during the influenza seasons than a constant case detection rate. This approach provides a reliable solution for improving forecasts of seasonal influenza and may be adaptable to other respiratory viral infections.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.013 | 0.052 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.003 | 0.005 |
| Open science | 0.002 | 0.002 |
| Research integrity | 0.002 | 0.003 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".