Advancing Sub-Seasonal to Seasonal Streamflow Forecasting in Canada: A Review of Conventional and Emerging Approaches for Operational Applications
Bibliographic record
Abstract
• A structured review of the state-of-the-art S2S streamflow forecasting, focusing on Canada • An overview of machine learning applications in S2S streamflow forecasting • A multi-model framework for developing operational tools for multi-sectoral S2S forecasting Sub-seasonal to seasonal (S2S) streamflow forecasts play a critical role in the planning and management of water resources for various purposes, such as optimization of hydropower production, ensuring sufficient water supplies for various usages, mitigating flood and drought risks, and management of nutrients from industrial and agricultural sources. Contrary to day-to-day operational activities, such forecasts can provide an extended operational window to various levels of the government for taking appropriate actions and issuing timely directives. Compared to the vast amount of hydrologic literature on short-term streamflow forecasting, S2S forecasting area is still not well-developed. This paper reviews state-of-the-art in S2S streamflow forecasting, considering conventional process-based and statistical modeling approaches, emerging machine learning (ML) techniques, and hybrid options. The generated knowledge and insights are intended to guide the development of operational tools for S2S forecasting for Alberta, Saskatchewan, and Manitoba provinces of Canada, and can also be used for developing similar tools for other regions of the world. Apart from discussing various modeling challenges, data availability constraints, and quantification of uncertainties, the paper also presents a systematic framework for developing ML-based S2S streamflow forecasting tools. Various limitations of the reviewed approaches and potential avenues of future research are also discussed to advance research and applications in S2S forecasting area. It is found that the potential of ML in addressing scaling issues in hydrology, through S2S forecasting, and investigating relevant hydrologic mechanisms at coarse spatial and temporal resolutions are not adequately explored. This is a significant path forward for ML in hydrology.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.005 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.005 | 0.012 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.003 | 0.002 |
| Open science | 0.002 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".