Quantifying the impact of model errors on top-down estimates of carbon monoxide emissions using satellite observations
Bibliographic record
Abstract
[1] We conduct inverse analyses of atmospheric CO, using the GEOS-Chem model and observations from the Measurement of Pollution in the Troposphere satellite instrument, to quantify the potential contribution of systematic model errors on top-down source estimates of CO. We assess how the specification of the source of CO from the oxidation of biogenic nonmethane volatile organic compounds (NMVOCs) in the inversion impacts the top-down estimates. Our results show that when the NMVOC source of CO is comparable to or larger than the combustion source, optimizing the CO from NMVOC emissions on larger spatial scales than the combustion emissions could result in significant overadjustment for the a posteriori CO emissions and could lead to negative sources of CO, such as we found for the top-down South American emissions in June. We quantify the impact of aggregation errors on the source estimates, associated with conducting the inversion at a lower resolution than the atmospheric model. We find that aggregating the emissions across spatial scales in which the a priori error in the emissions changes sign could introduce biases exceeding 20% in the flux estimates since the inversion cannot correct the a priori error by uniformly scaling the emissions across the region. We also use the GEOS-3 and GEOS-4 meteorological fields in GEOS-Chem to examine the impact of discrepancies in atmospheric transport and in the atmospheric OH distribution on the source estimates. We find that the differences in the OH distribution and transport fields associated with the GEOS-3 and GEOS-4 products introduce comparably large differences of as much as 20% in the source estimates. Our results indicate that mitigating systematic model error is critical for improving the accuracy of the inferred source estimates.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".