A sulfur dioxide Covariance-Based Retrieval Algorithm (COBRA): application to TROPOMI reveals new emission sources
Bibliographic record
Abstract
Sensitive and accurate detection of sulfur dioxide (SO 2 ) from space is important for monitoring and estimating global sulfur emissions. Inspired by detection methods applied in the thermal infrared, we present here a new scheme to retrieve SO 2 columns from satellite observations of ultraviolet back-scattered radiances. The retrieval is based on a measurement error covariance matrix to fully represent the SO 2 -free radiance variability, so that the SO 2 slant column density is the only retrieved parameter of the algorithm. We demonstrate this approach, named COBRA, on measurements from the TROPOspheric Monitoring Instrument (TROPOMI) aboard the Sentinel-5 Precursor (S-5P) satellite. We show that the method reduces significantly both the noise and biases present in the current TROPOMI operational DOAS SO 2 retrievals. The performance of this technique is also benchmarked against that of the principal component algorithm (PCA) approach. We find that the quality of the data is similar and even slightly better with the proposed COBRA approach. The ability of the algorithm to retrieve SO 2 accurately is further supported by comparison with ground-based observations. We illustrate the great sensitivity of the method with a high-resolution global SO 2 map, considering 2.5 years of TROPOMI data. In addition to the known sources, we detect many new SO 2 emission hotspots worldwide. For the largest sources, we use the COBRA data to estimate SO 2 emission rates. Results are comparable to other recently published TROPOMI-based SO 2 emissions estimates, but the associated uncertainties are significantly lower than with the operational data. Next, for a limited number of weak sources, we demonstrate the potential of our data for quantifying SO 2 emissions with a detection limit of about 8 kt yr −1 , a factor of 4 better than the emissions derived from the Ozone Monitoring Instrument (OMI). We anticipate that the systematic use of our TROPOMI COBRA SO 2 column data set at a global scale will allow missing sources to be identified and quantified and help improve SO 2 emission inventories.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".