Numerical impacts on tracer transport: A proposed intercomparison test of Atmospheric General Circulation Models
Bibliographic record
Abstract
Abstract The transport of trace gases by the atmospheric circulation plays an important role in the climate system and its response to external forcing. Transport presents a challenge for Atmospheric General Circulation Models (AGCMs), as errors in both the resolved circulation and the numerical representation of transport processes can bias their abundance. In this study, two tests are proposed to assess transport by the dynamical core of an AGCM. To separate transport from chemistry, the tests focus on the age‐of‐air, an estimate of the mean transport time by the circulation. The tests assess the coupled stratosphere–troposphere system, focusing on transport by the overturning circulation and isentropic mixing in the stratosphere, or Brewer–Dobson Circulation, where transport time‐scales on the order of months to years provide a challenging test of model numerics. Four dynamical cores employing different numerical schemes (finite‐volume, pseudo‐spectral, and spectral‐element) and discretizations (cubed sphere versus latitude–longitude) are compared across a range of resolutions. The subtle momentum balance of the tropical stratosphere is sensitive to model numerics, and the first intercomparison reveals stark differences in tropical stratospheric winds, particularly at high vertical resolution: some cores develop westerly jets and others easterly jets. This leads to substantial spread in transport, biasing the age‐of‐air by up to 25% relative to its climatological mean, making it difficult to assess the impact of the numerical representation of transport processes. This uncertainty is removed by constraining the tropical winds in the second intercomparison test, in a manner akin to specifying the Quasi‐Biennial Oscillation in an AGCM. The dynamical cores exhibit qualitative agreement on the structure of atmospheric transport in the second test, with evidence of convergence as the horizontal and vertical resolution is increased in a given model. Significant quantitative differences remain, however, particularly between models employing spectral versus finite‐volume numerics, even in state‐of‐the‐art cores.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.029 | 0.049 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.002 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.003 | 0.003 |
| Open science | 0.003 | 0.003 |
| Research integrity | 0.003 | 0.002 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".