Agreement in late twentieth century Southern Hemisphere stratospheric temperature trends in observations and CCMVal‐2, CMIP3, and CMIP5 models
Bibliographic record
Abstract
We present a comparison of temperature trends using different satellite and radiosonde observations and climate (GCM) and chemistry‐climate model (CCM) outputs, focusing on the role of photochemical ozone depletion in the Antarctic lower stratosphere during the second half of the twentieth century. Ozone‐induced stratospheric cooling peaks during November at an altitude of approximately 100 hPa in radiosonde observations, with 1969 to 1998 trends in the range of −3.8 to −4.7 K/dec. This stratospheric cooling trend is more than 50% greater than the previously estimated value of −2.4 K/dec, which suggested that the CCMs were overestimating the stratospheric cooling, and that the less complex GCMs forced by prescribed ozone were matching observations better. Corresponding ensemble mean model trends are −3.8 K/dec for the CCMs, −3.5 K/dec for the CMIP5 GCMs, and −2.7 K/dec for the CMIP3 GCMs. Accounting for various sources of uncertainty—including sampling uncertainty, measurement error, model spread, and trend confidence intervals—observations and CCM and GCM ensembles are consistent in this new analysis. This consistency does not apply to each individual that makes up the GCM and CCM ensembles, and some do not show significant ozone‐induced cooling. Nonetheless, analysis of the joint ozone and temperature trends in the CCMs suggests that the modeled cooling/ozone‐depletion relationship is within the range of observations. Overall, this study emphasizes the need to use a wide range of observations for model validation as well as sufficient accounting of uncertainty in both models and measurements.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".