Exploring the decision-making process in model development: focus on the Arctic snowpack
Bibliographic record
Abstract
Abstract. The Arctic poses many challenges for Earth system and snow physics models, which are commonly unable to simulate crucial Arctic snowpack processes,such as vapour gradients and rain-on-snow-induced ice layers. These limitations raise concerns about the current understanding of Arctic warming and its impact on biodiversity, livelihoods, permafrost, and the global carbon budget. Recognizing that models are shaped by human choices, 18 Arctic researchers were interviewed to delve into the decision-making process behind model construction. Although data availability, issues of scale, internal model consistency, and historical and numerical model legacies were cited as obstacles to developing an Arctic snowpack model, no opinion was unanimous. Divergences were not merely scientific disagreements about the Arctic snowpack but reflected the broader research context. Inadequate and insufficient resources, partly driven by short-term priorities dominating research landscapes, impeded progress. Nevertheless, modellers were found to be both adaptable to shifting strategic research priorities – an adaptability demonstrated by the fact that interdisciplinary collaborations were the key motivation for model development – and anchored in the past. This anchoring and non-epistemic values led to diverging opinions about whether existing models were “good enough” and whether investing time and effort to build a new model was a useful strategy when addressing pressing research challenges. Moving forward, we recommend that both stakeholders and modellers be involved in future snow model intercomparison projects in order to drive developments that address snow model limitations currently impeding progress in various disciplines. We also argue for more transparency about the contextual factors that shape research decisions. Otherwise, the reality of our scientific process will remain hidden, limiting the changes necessary to our research practice.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".