Spinal mobility in radiographic axial spondyloarthritis: criterion concurrent validity of classic and novel measurements
Bibliographic record
Abstract
BACKGROUND: Limitations in spinal mobility are a characteristic feature of Axial Spondyloarthritis. Current clinical measurements of spinal mobility have shown low criterion-concurrent validity. This study sought to evaluate criterion-concurrent validity for a clinically feasible measurement method of measuring spine mobility using tri-axial accelerometers. METHODS: Fifteen radiographic-Spondyloarthritis patients were recruited for this study. Two postural reference radiographs, followed by three trials in forward, left and right lateral bending were taken. For all trials, three measurements were collected: tape (Original Schober's, Modified Schober's, Modified-Modified Schober's, Lateral Spinal Flexion Test and Domjan Test), followed immediately by synchronized radiograph and accelerometer measurements at end range of forward and bilateral lateral flexion. The criterion-concurrent validity of all measurement methods was compared to the radiographic measures using Pearson's correlation coefficients. A Bland-Altman analysis was conducted to assess agreement. RESULTS: In forward bending, the accelerometer method (r = 0.590, p = 0.010) had a stronger correlation to the radiographic measures than all tape measures. In lateral bending, the Lateral Spinal Flexion tape measure (r = 0.743, p = 0.001) correlated stronger than the accelerometer method (r = 0.556, p = 0.016). The Domjan test of bilateral bending (r = 0.708, p = 0.002) had a stronger correlation to the radiographic measure than the accelerometer method. CONCLUSIONS: Accelerometer measures demonstrated superior criterion-concurrent validity compared to current tape measures of spinal mobility in forward bending. While a moderate correlation exists between accelerometer and radiographs in lateral bending, the Lateral Spinal Flexion Test and Domjan Test were found to have the best criterion-concurrent validity of all tests examined in this study.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".