MétaCan
Menu
Retour à la cohorte
Enregistrement W2965835723

Group-based estimation of missing hydrological data

2001· article· en· W2965835723 sur OpenAlexvenueno aff
Amin Elshorbagy

Notice bibliographique

RevueLibrary and Archives Canada (Government of Canada) · 2001
Typearticle
Langueen
DomaineEnvironmental Science
ThématiqueHydrology and Watershed Management Studies
Établissements canadiensnon disponible
Organismes subventionnairesnon disponible
Mots-clésMissing dataEstimationComputer scienceStatisticsMathematicsEngineering
DOInon disponible

Résumé

récupéré en direct d'OpenAlex

Water resources planning and management require complete data sets of many variables, such as rainfall, streamflow, and temperature. Unfortunately, records of hydrologic processes are usually short and often have missing observations. Attracted by the importance of estimating missing data, hydrologic researchers have adopted and developed various models and techniques to in-fill missing data. The diversity of the adopted techniques does not necessarily indicate diversity in the approach. A major commonality exists in most of the applications of these techniques; that is, any hydrologic time series record is perceived as a sequence of single-valued observations irrespective of the time scale of the data or their underlying structure. In this research, the group approach, different from the traditional single-valued approach, is proposed. The approach perceives the periodic hydrologic data as sequence of groups rather than single-valued observations. The techniques suggested to handle the group approach, after modification, are regression, time series analysis, partitioning modeling, and artificial neural networks. Various models representing these four techniques are briefly presented and applied to single series and bi-series cases, respectively. Also group time series models are developed in this thesis for the same purpose. It turns out that the group approach is highly useful for estimating consecutive missing values, and possibly other applications, such as long-term forecast. On the other hand, in non-periodic data (e.g., daily flows) where seasonality does not play a major role and a definite number of repetitive low dimensional groups of observations cannot be found in the geophysical year, another approach of identifying and modeling groups is sought. The nonlinearity and dynamic behavior of non-periodic hydrologic data sets have been indicated in water resources literature as issues that influence the performance of modeling tools that ignore nonl nearity and dynamics inherent in the data structure. Consecutive missing streamflows are estimated, using the principles of chaos theory, in two steps. First, the existence of chaotic behavior in daily flows of the river is investigated. Second, the analysis of chaos is used to configure two models employed to estimate missing data: artificial neural networks and K-nearest neighbors. Also, another local linear model is applied for comparison purposes. The results highlight the utility of using the analysis of chaos for configuring the models. In an unprecedented trial, in the chaos literature in water resources, the effect of the chaotic behavior on the analysis of two cross-correlated time series is investigated. The effect of both nonlinearity and dynamics is shown through application to daily streamflows. Other issues such as noise reduction and the reliability of its application to hydrologic time series are discussed. It is recommended that current noise reduction algorithms should be applied with caution and used for better estimation of chaotic invariants. The raw data should always be the basis for any further hydrologic analysis. After decades of adopting stochastic hydrology, chaos analysis, which has been recently introduced to hydrology, provides challenges and opportunities in hydrologic research. It has the potential to change the way in which hydrologic, and other real, processes are perceived, analyzed, and interpreted. The phenomenon that used to be treated as random may turn out to be nonlinear deterministic (chaotic) process.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,005
score de la tête « metaresearch » (Gemma)0,016
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Simulation ou modélisation · Signal consensuel: Simulation ou modélisation
GenreSignal candidat: Empirique · Signal consensuel: aucune
Score de désaccord entre enseignants0,005
Score d'incertitude au seuil0,025

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,0050,016
Méta-épidémiologie (sens strict)0,0010,001
Méta-épidémiologie (sens large)0,0020,001
Bibliométrie0,0020,002
Études des sciences et des technologies0,0000,001
Communication savante0,0010,002
Science ouverte0,0020,002
Intégrité de la recherche0,0010,001
Charge utile insuffisante (le modèle a refusé de juger)0,0020,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,008
Tête enseignante GPT0,163
Écart entre enseignants0,155 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeSimulation ou modélisation
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations8
Publié2001
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueLibrary and Archives Canada (Government of Canada)Même sujetHydrology and Watershed Management StudiesTravaux en français237 207