Forecasting streamflow using Artificial Neural Network (ANN) with different spatial discretizations of the watershed : use case on the Au Saumon watershed in Quebec (Canada).
Notice bibliographique
Résumé
Improving streamflow forecasts helps in reducing socio-economical impacts of hydrological-related damages. Among them, improving hydropower production is a challenge, even more so in a context of climate change. Deep learning models drew the attention of scientists working on forecasting models based on physical laws, since they got recognition in other domains. Artificial Neural Network (ANN) offer promising performance for streamflow forecasts, including good accuracy and lesser time to run compared to traditional physically-based models. The objective of this study is to compare different spatial discretization schemes of inputs in an ANN model for streamflow forecast. The study focuses on the “Au Saumon” watershed in Southern Quebec (Canada) during summer periods, with a forecast window of 7 days at a daily timestep. Parameterization of the ANN was a key preliminary step: the number of neurons in the hidden layer was first optimized, leading to 6 neurons. The model was trained on a 11-year dataset (2000-2005 and 2007-2011) followed by model validation on one dry (2012) and one wet (2006) year to take into account extreme hydrologic regimes. To lead this study, the physically-based hydrological ‘Hydrotel’ model is the reference to compare our results. The model defines watershed heterogeneity using hydrological units based on land uses, soil types, and topography, called Relative Homogeneous Hydrological Units (RHHU). The Nash-Sutcliffe Efficiency score (NSE) is the main evaluation criteria calculated. In a preliminary step, we have to ensure the ANN model can satisfactorily mimic Hydrotel. With the same model inputs, that is same variables and same spatial discretizations of variables (total precipitation, daily maximum and minimum temperatures, and soil surface humidity), the ANN forecasts were found to be better than those of Hydrotel for one to 7-day forecasts. Three different watershed spatial discretizations were tested: global, fully distributed, and semi-distributed. For the global model, hydrometeorological data used as inputs to the ANN model were averaged across all RHHUs. The complexity is reduced with loss of spatial details and heterogeneity. For the fully distributed model, a regular grid was defined with six cells of 28x28km2 covering all the watershed. For the semi-distributed model, spatial distribution of the input data was that of the RHHUs. For this discretization, the state variables (soil moisture and outflow) were updated at each forecast timestep, whether on all RHHUs, or only on the RHHU of the outlet. Depending on the spatial discretization of inputs used, the accuracy differed. The fully distributed model offered the least performance, with NSE values of 0.85 ,while the global model surprisingly performed better with a 0.93 NSE. Moreover, updating soil moisture on all the RHHUs of the semi-distributed model improved the NSE across the entire window of forecast. This research will assess the ANN model performance developed using ERA5-land precipitation and temperature reanalysis and ground observations of soil moisture. Given the promising results obtained with the fully and semi distributed models, our ANN model will be tested with state variables retrieved from satellite data, such as surface soil moisture from SMAP and SMOS missions.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,000 | 0,001 |
| Méta-épidémiologie (sens strict) | 0,001 | 0,000 |
| Méta-épidémiologie (sens large) | 0,000 | 0,000 |
| Bibliométrie | 0,000 | 0,001 |
| Études des sciences et des technologies | 0,001 | 0,000 |
| Communication savante | 0,001 | 0,000 |
| Science ouverte | 0,001 | 0,000 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,002 | 0,000 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».