MétaCan
Menu
Back to cohort
Record W4396685046 · doi:10.1371/journal.pcbi.1011200

Challenges of COVID-19 Case Forecasting in the US, 2020–2021

2024· article· en· W4396685046 on OpenAlexaff
Velma K. Lopez, Estee Y. Cramer, Roberto Pagano, John M. Drake, Eamon B. O’Dea, Madeline Adee, Turgay Ayer, Jagpreet Chhatwal, Özden O. Dalgıç, Mary A. Ladd, Benjamin P. Linas, Peter P. Mueller, Jade Xiao, Johannes Bracher, Alvaro J. Castro Rivadeneira, Aaron Gerding, Tilmann Gneiting, Yuxin Huang, Abdul Hannan Kanji, Khoa Le, Anja Mühlemann, Jarad Niemi, Evan L Ray, Ariane Stark, Yijin Wang, Nutcha Wattanachit, Martha Zorn, Sen Pei, Jeffrey Shaman, Teresa K. Yamana, Samuel R. Tarasewicz, Daniel J. Wilson, Sid Baccam, Heidi Gurung, Steve Stage, Brad Suchoski, Lei Gao, Zhiling Gu, Myungjin Kim, Xinyi Li, Guannan Wang, Li Wang, Yueying Wang, Shan Yu, Lauren Gardner, Sonia Jindal, Maximilian Marshall, Kristen Nixon, Juan Dent, Alison L. Hill, Joshua Kaminsky, Elizabeth C. Lee, Joseph C. Lemaitre, Justin Lessler, Claire P. Smith, Shaun Truelove, Matt Kinsey, Luke C. Mullany, Kaitlin Rainwater‐Lovett, Lauren Shin, Katharine Tallaksen, Shelby Wilson, D. Karlen, Lauren Castro, Geoffrey Fairchild, Isaac Michaud, Dave Osthus, Jiang Bian, Wei Cao, Zhifeng Gao, Juan Lavista Ferres, Chaozhuo Li, Tie‐Yan Liu, Xing Xie, Shun Zhang, Shun Zheng, Matteo Chinazzi, Jessica T. Davis, Kunpeng Mu, Ana Pastore y Piontti, Alessandro Vespignani, Xinyue Xiong, Robert Walraven, Jinghui Chen, Quanquan Gu, Lingxiao Wang, Pan Xu, Weitong Zhang, Difan Zou, Graham Gibson, Daniel Sheldon, Ajitesh Srivastava, Aniruddha Adiga, Benjamin Hurt, Gursharn Kaur, Bryan Lewis, Madhav Marathe, Akhil Sai Peddireddy, Przemyslaw Porebski, Srinivasan Venkatramanan, Lijing Wang, Pragati Prasad, Jo Walker, Alexander E. Webber, Rachel B. Slayton, Matthew Biggerstaff, Nicholas G Reich, Michael A. Johansson

Bibliographic record

VenuePLoS Computational Biology · 2024
Typearticle
Languageen
FieldMathematics
TopicCOVID-19 epidemiological studies
Canadian institutionsUniversity of VictoriaTRIUMF
FundersState of CaliforniaNational Institute of General Medical SciencesLaboratory Directed Research and DevelopmentJohns Hopkins Bloomberg School of Public HealthDefense Threat Reduction AgencyCenters for Disease Control and PreventionAndrew and Corey Morris-Singer FoundationLos Angeles County Department of Public HealthNational Science FoundationGoogleCouncil of State and Territorial EpidemiologistsJohns Hopkins UniversityDivision of Intramural Research, National Institute of Allergy and Infectious DiseasesVirginia Department of HealthU.S. Department of Homeland SecurityLos Alamos National LaboratoryNational Institute of Allergy and Infectious DiseasesIowa State UniversityAmazon Web ServicesNational Institutes of HealthSchweizerischer Nationalfonds zur Förderung der Wissenschaftlichen ForschungU.S. Department of Health and Human ServicesU.S. Department of Defense
KeywordsCoronavirus disease 2019 (COVID-19)Baseline (sea)EconometricsPandemicConfidence intervalStatisticsForecast skillGovernment (linguistics)Severe acute respiratory syndrome coronavirus 2 (SARS-CoV-2)MathematicsMedicineDiseaseBiology

Abstract

fetched live from OpenAlex

During the COVID-19 pandemic, forecasting COVID-19 trends to support planning and response was a priority for scientists and decision makers alike. In the United States, COVID-19 forecasting was coordinated by a large group of universities, companies, and government entities led by the Centers for Disease Control and Prevention and the US COVID-19 Forecast Hub (https://covid19forecasthub.org). We evaluated approximately 9.7 million forecasts of weekly state-level COVID-19 cases for predictions 1-4 weeks into the future submitted by 24 teams from August 2020 to December 2021. We assessed coverage of central prediction intervals and weighted interval scores (WIS), adjusting for missing forecasts relative to a baseline forecast, and used a Gaussian generalized estimating equation (GEE) model to evaluate differences in skill across epidemic phases that were defined by the effective reproduction number. Overall, we found high variation in skill across individual models, with ensemble-based forecasts outperforming other approaches. Forecast skill relative to the baseline was generally higher for larger jurisdictions (e.g., states compared to counties). Over time, forecasts generally performed worst in periods of rapid changes in reported cases (either in increasing or decreasing epidemic phases) with 95% prediction interval coverage dropping below 50% during the growth phases of the winter 2020, Delta, and Omicron waves. Ideally, case forecasts could serve as a leading indicator of changes in transmission dynamics. However, while most COVID-19 case forecasts outperformed a naïve baseline model, even the most accurate case forecasts were unreliable in key phases. Further research could improve forecasts of leading indicators, like COVID-19 cases, by leveraging additional real-time data, addressing performance across phases, improving the characterization of forecast confidence, and ensuring that forecasts were coherent across spatial scales. In the meantime, it is critical for forecast users to appreciate current limitations and use a broad set of indicators to inform pandemic-related decision making.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.009
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMetaresearch
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Theoretical or conceptual · Consensus signal: Theoretical or conceptual
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.111
Threshold uncertainty score0.999

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0010.009
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0000.000
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.483
GPT teacher head0.460
Teacher spread0.022 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designTheoretical or conceptual
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations32
Published2024
Admission routes1
Has abstractyes

Explore more

Same venuePLoS Computational BiologySame topicCOVID-19 epidemiological studiesFrench-language works237,207