The Unified North American Soil Map and its implication on the soil organic carbon stock in North America
Bibliographic record
Abstract
Abstract. The Unified North American Soil Map (UNASM) was developed to provide more accurate regional soil information for terrestrial biosphere modeling. The UNASM combines information from state-of-the-art US STATSGO2 and Soil Landscape of Canada (SLCs) databases. The area not covered by these datasets is filled with the Harmonized World Soil Database version 1.1 (HWSD1.1). The UNASM contains maximum soil depth derived from the data source as well as seven soil attributes (including sand, silt, and clay content, gravel content, organic carbon content, pH, and bulk density) for the top soil layer (0–30 cm) and the sub soil layer (30–100 cm) respectively, of the spatial resolution of 0.25° in latitude and longitude. There are pronounced differences in the spatial distributions of soil properties and soil organic carbon between UNASM and HWSD, but the UNASM overall provides more detailed and higher-quality information particularly in Alaska and Central Canada. To provide more accurate and up-to-date estimate of soil organic carbon stock in North America, we incorporated Northern Circumpolar Soil Carbon Database (NCSCD) into the UNASM. The estimate of total soil organic carbon mass in the upper 100 cm soil profile based on the improved UNASM is 347.70 Pg, of which 24.7% is under trees, 14.2% is under shrubs, and 1.3% is under grasses and 3.8% under crops. This UNASM data will provide a resource for use in land surface and terrestrial biogeochemistry modeling both for input of soil characteristics and for benchmarking model output.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".