MétaCan
Menu
Back to cohort
Record W4407798055 · doi:10.3389/fspor.2025.1570591

Corrigendum: Addressing grading bias in rock climbing: machine and deep learning approaches

2025· erratum· en· W4407798055 on OpenAlexaboutno aff
B. J. O’Mara, M. A. Mahmud

Bibliographic record

VenueFrontiers in Sports and Active Living · 2025
Typeerratum
Languageen
FieldEngineering
TopicMedical Imaging and Analysis
Canadian institutionsnot available
Fundersnot available
KeywordsClimbingGrading (engineering)Artificial intelligenceComputer scienceMachine learningEngineeringCivil engineeringStructural engineering

Abstract

fetched live from OpenAlex

1 IntroductionRock climbing’s popularity as a recreational sport is growing dramatically. It is a unique activity for both the body and the mind. In a puzzle-solving manner, climbers strategically scale vertical natural or artificial rock routes–a series of rock features–using their hands and feet. People are drawn to rock climbing because it is an activity in which one can improve physical fitness, problem-solving skills, and self-confidence (1, 2). It is estimated that the rock climbing gym market size was valued at 3 billion USD in 2023 (1), and this projected to double by 2032 (1). In the last five years, the establishment of rock climbing gyms in the US has grown by 6.46% per year (3). Within this time frame, three of the disciplines of rock climbing: sport climbing, bouldering, and speed climbing made their debut appearance at the 2020 Tokyo Olympics. More recently, it reached global audiences again at the 2024 Paris Olympics. Of the three disciplines, bouldering has the largest share of the rock climbing gym market (1, 2). This is because bouldering is the most accessible discipline of climbing, as it requires little equipment and technical knowledge. Recent trends in gym establishment highlight the increasing accessibility to bouldering. In the past decade, 50% of the gyms established in the United States and Canada are only bouldering gyms (3). To capitalize on this popularity, accessibility is crucial for climbing gym success.Climbing accessibility is highly dependent on route setters. Route setters produce climbing routes, the central service of a climbing gym. They are responsible for producing routes that are varied yet consistent in difficulty. Gyms vary their route difficulties to capture the largest audience possible [Market Research Future], catering to a range of climber experience levels from novice to advanced. However, the grading scales used to rate climbing route difficulty are often subjective according to the region, the gym, and the setter of the route (4). General factors considered when determining route difficulty are rock hold types, the number of rock holds on a route, the distance between the rock holds, and the angle of ascent (5). Therefore, it seems that the positioning and sequencing of holds are critical to route difficulty. But holds may be positioned and sequenced in an almost infinite number of ways. Setting a route is like composing a song (6, 7); there are constraints that govern its composition, but the liberty to operate within those constraints is quite large. When operating within these constraints, a route can be developed in a multitude of ways. This wide variance of route generation is a challenge for generalizing route difficulty. without a large sample size, route setters introduce their own biases when determining route difficulty, which then inadvertently affects the climber (i.e., the customer).2 MotivationRoute setters are in an awkward position. The act of setting routes is inherently subjective, but the success of a climb depends on the ability of the setter to objectively set routes. This is the Grading Bias Problem: the setter of a route introduces their biases when declaring a route’s difficulty.Reporting the objective grade of a climbing route is critical in the climbing community and can be aided by machine and deep learning technology (5). Increasingly, machine learning and deep learning techniques are being used to objectively classify the route difficulty. The objectives of this review article are to (1) understand how today’s route setters maintain objectivity in their setting, (2) to review the state-of-the-art approaches in determining climbing route difficulty with machine learning and deep learning, and (3) to suggest new areas for research. Together, these objectives are intended to address how climbing gyms can integrate machine and deep learning systems to streamline route setting and eliminate route difficulty bias for greater consistency and accessibility.3 Document layoutThe Grading Bias Problem will be thoroughly explored in subsequent sections of this paper. Section 4 provides context on rock climbing grade scales and how route setters currently set routes with the goals of objectivity and accessibility. Section 5 details the survey methodology and inclusion criteria. Section 6 identifies the approaches and methods of various deep and machine learning techniques to determine the climbing route difficulty and their success rates. Section 7 discusses the trends, performance, and shortcomings of current machine learning and deep learning techniques. In addition, it is argued that a route-centric approach with a natural language-like model is most optimal. Section 7 continues by proposing future areas for research, in which some proposals are based on works auxiliary to the survey.4 Background: route grading systems and settingClimbing route difficulty can be graded on a variety of scales (Figure 1). The grade scale depends greatly on the discipline, subdisciplines, and climbing systems. In free climbing, the climber ascends a route without any artificial aid. The climber ascends a route by only the natural or artificial features of the rock. But a free climber can still use safety equipment (e.g., rope) in the event that they fall. In the subdisciplines of traditional (trad), sport, and ice climbing, of which the climber ascends a vertical face that is typically greater than 4 meters, the climber’s main tool for protection is a belay system consisting of rope, harness, and either temporary or permanent anchor points. The grading scales of these three “roped” disciplines account for risk to the climber in addition to the technical difficulty of climbing movement. Risk to the climber is most apparent in trad and ice climbing. In both trad and ice climbing, the climber sets and removes protection in rock crevices as they ascend. This protective gear is more prone to fail because they are not intended to be permanent fixtures in the rock or ice face. For ice climbing, it is particularly critical to gauge risk to the climber because the conditions of ice is greatly dependent of weather factors such as temperature, humidity, and precipitation. Although risk to the climber is still a concern in sport climbing, it is greatly reduced because

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.000
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMeta-epidemiology (narrow)
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Simulation or modeling · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.908
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0000.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0010.000
Bibliometrics0.0010.000
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0000.000
Research integrity0.0000.002
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.031
GPT teacher head0.228
Teacher spread0.197 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designSimulation or modeling
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2025
Admission routes1
Has abstractyes

Explore more

Same venueFrontiers in Sports and Active LivingSame topicMedical Imaging and AnalysisFrench-language works237,207