Assessment of Odin-OSIRIS ozone measurements from 2001 to the present using MLS, GOMOS, and ozonesondes
Bibliographic record
Abstract
Abstract. The Optical Spectrograph and InfraRed Imaging System (OSIRIS) was launched aboard the Odin satellite in 2001 and is continuing to take limb-scattered sunlight measurements of the atmosphere. This work aims to characterize and assess the stability of the OSIRIS 11 yr v5.0x ozone data set. Three validation data sets were used: the v2.2 Microwave Limb Sounder (MLS) and v6 Global Ozone Monitoring by Occultation of Stars (GOMOS) satellite data records, and ozonesonde measurements. Global mean percent differences between coincident OSIRIS and validation measurements are within 5% at all altitudes above 18.5 km for MLS, above 21.5 km for GOMOS, and above 17.5 km for ozonesondes. Below 17.5 km, OSIRIS measurements agree with ozonesondes within 5% and are well-correlated (R > 0.75) with them. For low OSIRIS optics temperatures (< 16 °C), OSIRIS ozone measurements have a negative bias of 1–6% compared with the validation data sets for 25.5–40.5 km. Biases between OSIRIS ascending and descending node measurements were investigated and found to be related to aerosol retrievals below 27.5 km. Above 30 km, agreement between OSIRIS and the validation data sets was related to the OSIRIS retrieved albedo, which measures apparent upwelling, with a positive bias in OSIRIS data with large albedos. In order to assess the long-term stability of OSIRIS measurements, global average drifts relative to the validation data sets were calculated and were found to be < 3% per decade for comparisons with MLS for 19.5–36.5 km, GOMOS for 18.5–54.5 km, and ozonesondes for 12.5–22.5 km. Above 36.5 km, the relative drift for OSIRIS versus MLS ranged from ~ 0 to 6% per decade, depending on the data set used to convert MLS data to the OSIRIS altitude versus number density grid. Overall, this work demonstrates that the OSIRIS 11 yr ozone data set from 2001 to the present is suitable for trend studies.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".