Comment on acp-2022-502
Bibliographic record
Abstract
Abstract. In this study, three methods, i.e., the random forest (RF) algorithm, boosted regression trees (BRTs) and the improved complete ensemble empirical-mode decomposition with adaptive noise (ICEEMDAN), were adopted for investigating emission-driven interannual variations in concentrations of air pollutants including PM2.5, PM10, O3, NO2, CO, SO2 and NO2 + O3 monitored in six cities in South China from May 2014 to April 2021. The first two methods were used to calculate the deweathered hourly concentrations, and the third one was used to calculate decomposed hourly residuals. To constrain the uncertainties in the calculated deweathered or decomposed hourly values, a self-developed method was applied to calculate the range of the deweathered percentage changes (DePCs) of air pollutant concentrations on an annual scale (each year covers May to the next April). These four methods were combined together to generate emission-driven trends and percentage changes (PCs) during the 7-year period. Consistent trends between the RF-deweathered and BRT-deweathered concentrations and the ICEEMDAN-decomposed residuals of an air pollutant in a city were obtained in approximately 70 % of a total of 42 cases (for seven pollutants in six cities), but consistent PCs calculated from the three methods, defined as the standard deviation being smaller than 10 % of the corresponding mean absolute value, were obtained in only approximately 30 % of all the cases. The remaining cases with inconsistent trends and/or PCs indicated large uncertainties produced by one or more of the three methods. The calculated PCs from the deweathered concentrations and decomposed residuals were thus combined with the corresponding range of DePCs calculated from the self-developed method to gain the robust range of DePCs where applicable. Based on the robust range of DePCs, we identified significant decreasing trends in PM2.5 concentration from 2014 to 2020 in Guangzhou and Shenzhen, which were mainly caused by the reduced air pollutant emissions and to a much lesser extent by weather perturbations. A decreasing or probably decreasing emission-driven trend was identified in Haikou and Sanya with inconsistent PCs, and a stable or no trend was identified in Zhanjiang with positive PCs. For O3, a significant increasing trend from 2014 to 2020 was identified in Zhanjiang, Shenzhen, Guangzhou and Haikou. An increasing trend in NO2 + O3 was also identified in Zhanjiang and Guangzhou and an increasing or probably increasing trend in Haikou, suggesting the contributions from enhanced formation of O3. The calculated PCs from using different methods implied that the emission changes in O3 precursors and the associated atmospheric chemistry likely played a dominant role than did the perturbations from varying weather conditions. Results from this study also demonstrated the necessity of combining multiple decoupling methods in generating emission-driven trends in atmospheric pollutants.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.032 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.002 |
| Bibliometrics | 0.002 | 0.003 |
| Science and technology studies | 0.004 | 0.002 |
| Scholarly communication | 0.005 | 0.003 |
| Open science | 0.003 | 0.002 |
| Research integrity | 0.025 | 0.015 |
| Insufficient payload (model declined to judge) | 0.243 | 0.223 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".