Using transport diagnostics to understand chemistry climate model ozone simulations
Bibliographic record
Abstract
[1] We use observations of N2O and mean age to identify realistic transport in models in order to explain their ozone predictions. The results are applied to 15 chemistry climate models (CCMs) participating in the 2010 World Meteorological Organization ozone assessment. Comparison of the observed and simulated N2O, mean age and their compact correlation identifies models with fast or slow circulations and reveals details of model ascent and tropical isolation. This process-oriented diagnostic is more useful than mean age alone because it identifies models with compensating transport deficiencies that produce fortuitous agreement with mean age. The diagnosed model transport behavior is related to a model's ability to produce realistic lower stratosphere (LS) O3 profiles. Models with the greatest tropical transport problems compare poorly with O3 observations. Models with the most realistic LS transport agree more closely with LS observations and each other. We incorporate the results of the chemistry evaluations in the Stratospheric Processes and their Role in Climate (SPARC) CCMVal Report to explain the range of CCM predictions for the return-to-1980 dates for global (60°S–60°N) and Antarctic column ozone. Antarctic O3 return dates are generally correlated with vortex Cly levels, and vortex Cly is generally correlated with the model's circulation, although model Cl chemistry and conservation problems also have a significant effect on return date. In both regions, models with good LS transport and chemistry produce a smaller range of predictions for the return-to-1980 ozone values. This study suggests that the current range of predicted return dates is unnecessarily broad due to identifiable model deficiencies.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.003 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".