Developing priority criteria for hip and knee replacement: results from the Western Canada Waiting List Project.
Bibliographic record
Abstract
INTRODUCTION: The Western Canada Waiting List Project (WCWL), a federally funded partnership of 19 organizations, was created to develop tools for managing waiting lists. The WCWL panel on hip and knee replacement surgery was 1 of 5 panels constituted under this project. METHODS: The panel developed and tested a collection of standardized clinical criteria for setting priorities among patients awaiting hip and knee replacement. The criteria were applied to 405 patients in 4 provinces. Regression analysis was used to determine the set of criteria weights that collectively best predicted clinicians' overall urgency ratings. Inter-rater and test-retest reliability was assessed from 6 videotaped patient interviews, scored by orthopedic surgeons, related professionals and general practitioners. RESULTS: The priority criteria accounted for over two-thirds of the observed variance in overall urgency ratings (adjusted R2 = 0.676). The panel modified the criteria and weights based on the empirical findings and on clinical judgement. The reliability of the priority criteria for the hip and knee replacement tool was among the strongest of the 5 instruments developed in the WCWL project. CONCLUSIONS: The panel considered the criteria easy to use and reasonably reflective of expert surgical judgement regarding clinical urgency for hip and knee replacement. Further development and testing of the tool appears warranted.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.015 | 0.046 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.003 |
| Science and technology studies | 0.002 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.002 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".