MétaCan
Menu
Retour à la cohorte
Enregistrement W4391110503 · doi:10.18822/edgcc625761

Multi-model ensemble successfully predicted atmospheric methane consumption in soils across the complex landscape

2024· article· en· W4391110503 sur OpenAlexaff
М. В. Глаголев, D. V. Ilyasov, А. Ф. Сабреков, Irina Terentieva, Д. В. Карелин

Notice bibliographique

RevueEnvironmental Dynamics and Global Climate Change · 2024
Typearticle
Langueen
DomaineEnvironmental Science
ThématiqueAtmospheric and Environmental Gas Dynamics
Établissements canadiensUniversity of Calgary
Organismes subventionnairesnon disponible
Mots-clésMethaneEnvironmental scienceSoil waterAtmospheric methaneConsumption (sociology)Atmospheric sciencesEarth sciencePhysical geographySoil scienceGeographyGeologyEcologySociology

Résumé

récupéré en direct d'OpenAlex

Methane consumption by soils is a crucial component of the CH4 and carbon cycle. It is essential to thoroughly investigate CH4 uptake by soils, particularly considering its anticipated increase by the end of the century [Zhuang et al., 2013]. Numerous mathematical models, both empirical and detailed biogeochemical [Glagolev et al., 2023], have been developed to quantify methane consumption by soils from the atmosphere. These models are instrumental in handling spatio-temporal variability and can offer reliable estimates of regional and global methane consumption by soils. Furthermore, they enhance our comprehension of the physical and biological processes that influence methanotrophy intensity. Consequently, we can forecast the response of CH4 consumption by soil to global climate shifts [Murguia-Flores et al., 2018], especially since many models consider the effects of atmospheric CH4 concentration changes on methanotrophy and ecosystem type [Zhuang et al., 2013]. In addition to the utilization of individual models, such as those cited by [Hagedorn et al., 2005; Glagolev et al., 2014; Ito et al., 2016; Silva et al., 2016], there has been extensive advancement in employing multiple models in an ensemble format. This approach aims to integrate as much a priori information as feasible [Lapko, 2002]. Throughout the 20th century, the concept of ensemble modeling evolved from merely drawing conclusions based on multiple independent experts (F. Sanders, 1963) to structured ensemble mathematical modeling [Hagedorn et al., 2005]. In this context, the term "ensemble" consistently refers to a collection containing more than one model. Complexities in describing the physiology and biochemistry of methanotrophic bacteria in natural environments [Bedard, Knowles, 1989; Hanson, Hanson, 1996; Belova et al., 2013; Oshkin et al., 2014] make it difficult to develop accurate biological models and determine their specific biokinetic parameters [Curry, 2007]. At the same time, broader and often empirical models, such as those by [Potter et al., 1996; Ridgwell et al., 1999; Curry, 2007; Murguia-Flores et al., 2018], demonstrate reasonable estimates of global methane consumption. Employing model ensembles could enhance accuracy, not just in global and large-scale modeling, but also at the granular level of local study sites. Nonetheless, ensemble modeling doesn't always ensure optimal outcomes, as all models within an ensemble might overlook a biological process or effect that significantly influences the dynamics of a real ecosystem [Ito et al., 2016]. For instance, no model considered anaerobic methane oxidation until this process was empirically identified [Xu et al., 2015]. Therefore, it's crucial to validate the realism of an ensemble against specific in situ data for every application. This study aimed to develop an ensemble model describing methane consumption by soils and to test its efficacy on a randomly selected study site. In our research, we closely examined and replicated the algorithms of four soil methane consumption models: the modification by Glagolev, Filippov [2011] of Dörr et al. [1993], Curry's model [2007], the CH4 consumption block from the DLEM model [Tian et al., 2010], and the MeMo model excluding autochthonous CH4 sources [Murguia-Flores et al., 2018]. Using these, we developed an ensemble of four models. For experimental in situ data, we utilized field measurements from the Kursk region in Russia. Additionally, we introduced a method to average the ensemble model's prediction by assigning weight coefficients to each model. This approach acknowledges the idea that the total available information doubles every few years. Thus, newer models were given higher weights, while older ones received lower weights. The model ensemble effectively predicted CH4 consumption based on in situ measurements, albeit with a notably broad confidence interval for the predictions. Notably, there was minimal variance between the standard averaging of model predictions and weighted averaging. As anticipated, individual models underperformed compared to the ensemble. We computed the Theil inconsistency coefficient for various types of means, such as quadratic mean, cubic mean, and biquadratic mean, among others [Gini, Barbensi, 1958], both for ensemble modeling results and individual models. The ensemble predictions, when averaged using diverse methods, yielded Theil inconsistency coefficients ranging from 0.156 to 0.267. The most favorable outcome (0.156) was derived from the power mean with a power index of 0.7. However, the power mean presents a challenge as its power index isn't predetermined but chosen to best fit the experimental data. A similar limitation exists for the exponential mean. While the experimental data allows for the selection of a parameter yielding a Theil coefficient of 0.157, pre-determining this optimal value (1.3) is not feasible. Regarding other estimations that don't necessitate selecting optimal parameters, it was surprising to find that one of the best results (Theil's coefficient = 0.166) came from the half-sum of extreme terms. Surprisingly, the median provided a less satisfactory result, with a Theil's coefficient of 0.222. The merit of the ensemble approach stems from P.D. Thompson's 1977 observation, which he stated assertively: "It is an indisputable fact that two or more inaccurate, but independent predictions of the same event can be combined in such a way that their "combined" forecast, on average, will be more accurate than any of these individual forecasts" [Hagedorn et al., 2005]. Examining our ensemble of models through this lens reveals a limitation, as the condition of independence isn't fully satisfied. The models by Dörr et al. [1993], Curry [2007], and MeMo [Murguia-Flores et al., 2018] share underlying similarities and can be seen as part of a cohesive cluster. Only DLEM, crafted on entirely distinct principles, stands apart from these models. To enhance the ensemble's robustness in future iterations, the inclusion of genuinely independent models, such as a modified version of MDM [Zhuang et al., 2013] and the model by Ridgwell et al. [1999], is recommended. The ensemble, comprising four models and implemented without specific parameter adjustments, effectively captured methane consumption across diverse sites in the Kursk region, such as fields and forests. On average, the relative simulation error for all these sites was 36%, with the experimental data displaying a variation of 26%. Notably, while the variation is modest for this dataset, methane absorption measurements generally tend to fluctuate by several tens of percent [Crill, 1991, Fig. 1; Ambus, Robertson, 2006, Fig. 3; Kleptsova et al., 2010; Glagolev et al., 2012]. Considering this broader perspective, the simulation error achieved is indeed favorable. Upon evaluating different methods for combining individual model results within the ensemble (specifically those methods that can be applied without prior parameter adjustments based on experimental data), it was found that the most straightforward operators yielded the best outcomes. This assessment was based on Theil's inequality coefficient criterion. Both the semi-sum of extreme terms and the arithmetic mean stood out in their performance. However, a significant drawback of the constructed ensemble is the extensive confidence interval for its predictions, averaging ±78% at a 90% probability level. We hypothesize that expanding the number of independent models within the ensemble could potentially narrow this interval.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,000
score de la tête « metaresearch » (Gemma)0,001
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Simulation ou modélisation · Signal consensuel: Simulation ou modélisation
GenreSignal candidat: Empirique · Signal consensuel: Empirique
Score de désaccord entre enseignants0,031
Score d'incertitude au seuil0,061

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,0000,001
Méta-épidémiologie (sens strict)0,0010,000
Méta-épidémiologie (sens large)0,0010,002
Bibliométrie0,0010,001
Études des sciences et des technologies0,0000,000
Communication savante0,0010,001
Science ouverte0,0010,001
Intégrité de la recherche0,0010,001
Charge utile insuffisante (le modèle a refusé de juger)0,0020,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,023
Tête enseignante GPT0,266
Écart entre enseignants0,243 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeSimulation ou modélisation
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations1
Publié2024
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueEnvironmental Dynamics and Global Climate ChangeMême sujetAtmospheric and Environmental Gas DynamicsTravaux en français237 207