The ERA5-Land soil temperature bias in permafrost regions
Bibliographic record
Abstract
Abstract. ERA5-Land (ERA5L) is a reanalysis product derived by running the land component of ERA5 at increased resolution. This study evaluates ERA5L soil temperature in permafrost regions based on observations and published permafrost products. We find that ERA5L overestimates soil temperature in northern Canada and Alaska but underestimates it in mid–low latitudes, leading to an average bias of −0.08 ∘C. The warm bias of ERA5L soil is stronger in winter than in other seasons. As calculated from its soil temperature, ERA5L overestimates active-layer thickness and underestimates near-surface (<1.89 m) permafrost area. This is thought to be due in part to the shallow soil column and coarse vertical discretization of the land surface model and to warmer simulated soil. The soil temperature bias in permafrost regions correlates well with the bias in air temperature and with maximum snow height. A review of the ERA5L snow parameterization and a simulation example both point to a low bias in ERA5L snow density as a possible cause for the warm bias in soil temperature. The apparent disagreement of station-based and areal evaluation techniques highlights challenges in our ability to test permafrost simulation models. While global reanalyses are important drivers for permafrost simulation, we conclude that ERA5L soil data are not well suited for informing permafrost research and decision making directly. To address this, future soil temperature products in reanalyses will require permafrost-specific alterations to their land surface models.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".