Proposed standardized definitions for vertical resolution and uncertainty inthe NDACC lidar ozone and temperature algorithms – Part 1: Verticalresolution
Bibliographic record
Abstract
Abstract. A standardized approach for the definition and reporting of vertical resolution of the ozone and temperature lidar profiles contributing to the Network for the Detection for Atmospheric Composition Change (NDACC) database is proposed. Two standardized definitions homogeneously and unequivocally describing the impact of vertical filtering are recommended. The first proposed definition is based on the width of the response to a finite-impulse-type perturbation. The response is computed by convolving the filter coefficients with an impulse function, namely, a Kronecker delta function for smoothing filters, and a Heaviside step function for derivative filters. Once the response has been computed, the proposed standardized definition of vertical resolution is given by Δz = δz × HFWHM, where δz is the lidar's sampling resolution and HFWHM is the full width at half maximum (FWHM) of the response, measured in sampling intervals. The second proposed definition relates to digital filtering theory. After applying a Laplace transform to a set of filter coefficients, the filter's gain characterizing the effect of the filter on the signal in the frequency domain is computed, from which the cut-off frequency fC, defined as the frequency at which the gain equals 0.5, is computed. Vertical resolution is then defined by Δz = δz∕(2fC). Unlike common practice in the field of spectral analysis, a factor 2fC instead of fC is used here to yield vertical resolution values nearly equal to the values obtained with the impulse response definition using the same filter coefficients. When using either of the proposed definitions, unsmoothed signals yield the best possible vertical resolution Δz = δz (one sampling bin). Numerical tools were developed to support the implementation of these definitions across all NDACC lidar groups. The tools consist of ready-to-use “plug-in” routines written in several programming languages that can be inserted into any lidar data processing software and called each time a filtering operation occurs in the data processing chain. When data processing implies multiple smoothing operations, the filtering information is analytically propagated through the multiple calls to the routines in order for the standardized values of vertical resolution to remain theoretically and numerically exact at the very end of data processing.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.026 | 0.056 |
| Meta-epidemiology (narrow) | 0.003 | 0.001 |
| Meta-epidemiology (broad) | 0.002 | 0.002 |
| Bibliometrics | 0.009 | 0.007 |
| Science and technology studies | 0.001 | 0.003 |
| Scholarly communication | 0.008 | 0.005 |
| Open science | 0.005 | 0.005 |
| Research integrity | 0.003 | 0.003 |
| Insufficient payload (model declined to judge) | 0.002 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".