Development and Initial Validation of a Radiographic Scoring System for the Hip in Juvenile Idiopathic Arthritis
Bibliographic record
Abstract
OBJECTIVE: To develop and validate a radiographic scoring system for the assessment of radiographic damage in the hip joint in patients with juvenile idiopathic arthritis (JIA). METHODS: The Childhood Arthritis Radiographic Score of the Hip (CARSH) assesses and scores these radiographic abnormalities: joint space narrowing (JSN), erosion, growth abnormalities, subchondral cysts, malalignment, sclerosis of the acetabulum, and avascular necrosis of the femoral head. Score validation was accomplished by evaluating reliability and correlational, construct, and predictive validity in 148 JIA patients with hip disease who had a total of 381 hip radiographs available for study. RESULTS: JSN was the most frequently observed radiographic abnormality, followed by erosion and sclerosis of the acetabulum. The least common abnormalities were avascular necrosis, growth abnormalities, and malalignment. Interobserver and intraobserver reliability on baseline and longitudinal score values and on score changes was good, with intraclass correlation coefficients ranging from 0.76 to 0.98. Early score changes, but not absolute baseline score values, were moderately correlated (r(s) > 0.4) with clinical indicators of disease damage at last followup observation, thereby demonstrating that the CARSH has good construct and predictive validity. The amount of structural damage in the hip radiograph at last followup observation was predicted better by baseline to 1-year score change (r(s) = 0.66; p < 0.0001) than by absolute baseline score values (r(s) = 0.40; p = 0.002). CONCLUSION: Our results show that the CARSH is reliable and valid for the assessment of radiographic hip damage and its progression in patients with JIA.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".