Examining the operational use of avalanche problems with decision trees and model-generated weather and snowpack variables
Bibliographic record
Abstract
Abstract. Avalanche problems are used in avalanche forecasting to describe snowpack, weather, and terrain factors that require distinct risk management techniques. Although they have become an effective tool for assessing and communicating avalanche hazard, their definitions leave room for interpretation and inconsistencies. This study uses conditional inference trees to explore the application of avalanche problems over eight winters in Glacier National Park, Canada. The influences of weather and snowpack variables on each avalanche problem type were explored by analysing a continuous set of weather and snowpack variables produced with a numerical weather prediction model and a physical snow cover model. The decision trees suggest forecasters' assessments are based on not only a physical analysis of weather and snowpack conditions but also contextual information about the time of season, the location, and interactions with other avalanche problems. The decision trees showed clearer patterns when new avalanche problems were added to hazard assessments compared to when problems were removed. Despite discrepancies between modelled variables and field observations, the model-generated variables produced intuitive explanations for conditions influencing most avalanche problem types. For example, snowfall in the past 72 h was the most significant variable for storm slab avalanche problems, skier penetration depth was the most significant variable for dry loose avalanche problems, and slab density was the most significant variable for persistent-slab avalanche problems. The explanations for wind slab and cornice avalanche problems were less intuitive, suggesting potential inconsistencies in their application as well as shortcomings of the model-generated data. The decision trees illustrate how forecasters apply avalanche problems and can inform discussions about improved operational practices and the development of data-driven decision aids.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".