Short communication: a new dataset for estimating organic carbon storage to 3 m depth in soils of the northern circumpolar permafrost region
Bibliographic record
Abstract
Abstract. High latitude terrestrial ecosystems are key components in the global carbon (C) cycle. The Northern Circumpolar Soil Carbon Database (NCSCD) was developed to quantify stocks of soil organic carbon (SOC) in the northern circumpolar permafrost region (18.7 × 106 km2). The NCSCD is a digital Geographical Information systems (GIS) database compiled from harmonized regional soil classification maps, in which data on soil coverage has been linked to pedon data from the northern permafrost regions. Previously, the NCSCD has been used to calculate SOC content (SOCC) and mass (SOCM) to the reference depths 0–30 cm and 0–100 cm (based on 1778 pedons). It has been shown that soils of the northern circumpolar permafrost region also contain significant quantities of SOC in the 100–300 cm depth range, but there has been no circumpolar compilation of pedon data to quantify this SOC pool and there are no spatially distributed estimates of SOC storage below 100 cm depth in this region. Here we describe the synthesis of an updated pedon dataset for SOCC in deep soils of the northern circumpolar permafrost regions, with separate datasets for the 100–200 cm (524 pedons) and 200–300 cm (356 pedons) depth ranges. These pedons have been grouped into the American and Eurasian sectors and the mean SOCC for different soil taxa (subdivided into Histels, Turbels, Orthels, Histosols, and permafrost-free mineral soil taxa) has been added to the updated NCSCDv2. The updated version of the database is freely available online in several different file formats and spatial resolutions that enable spatially explicit usage in e.g. GIS and/or terrestrial ecosystem models. The potential applications and limitations of the NCSCDv2 in spatial analyses are briefly discussed. An open access data-portal for all the described GIS-datasets is available online at: http://dev1.geo.su.se/bbcc/dev/v3/ncscd/download.php. The NCSCDv2 database has the doi:10.5879/ECDS/00000002.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.003 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.005 | 0.007 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.008 | 0.007 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".