Psychometric properties of the global rating of change scales in patients with low back pain, upper and lower extremity disorders. A systematic review with meta-analysis
Bibliographic record
Abstract
OBJECTIVE: The purpose of this systematic review was to critically appraise and synthesize the psychometric properties of the Global Rating of Change (GRoC) scales on the assessment of patients with low back pain (LBP), upper extremity and lower extremity disorders. METHODS: A search was performed in 4 databases (MEDLINE, EMBASE, CINAHL, SCOPUS) until February 2019. Eligible articles were appraised using Consensus-based Standards for the selection of health Measurement Instruments (COSMIN) checklist and the Quality Appraisal for Clinical Measurement Research Reports Evaluation Form. RESULTS: The 8 eligible studies included participants with orthopedic lumbar spine impairments (n = 52,767), patients with work-related musculoskeletal disorders (n = 1944), patients with low back pain (n = 183) and individuals with upper extremity disorders (n = 151). Risk of bias was ranging from "adequate" to "very good" and quality was found excellent for all studies. Based on pooled data, test-retest reliability of 11-item GRoC for patients with low back pain was found excellent ICC = 0.84, 95% CI: 0.65 to 0.94. Test-retest reliability in patients with shoulder pain was found fair to good ICC of 0.62 in a 15-point GRoC scale. Seven studies (n = 7) examined the convergent validity between GRoC and another outcome measure. Minimum important change on the Portuguese version of Global Perceived Effect (GPE) for patients with LBP was 2.5 points out of 11 points. CONCLUSIONS: The current pool of clinical measurement studies indicates that the GRoC has excellent test-retest reliability for patients with low back pain, shoulder pain and with lumbar spine disorders. However, the validity of it as a reference standard in responsiveness studies or as an accurate overall assessment of change has been questioned. While future studies might provide more insight into its measurement properties, this limitation is unlikely to change. Therefore, we suggest that future responsiveness in the studies that want a global indicator measure need to use an additional measure to mitigate recall bias. PROSPERO REGISTRATION NUMBER: CRD 42020149122.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.020 | 0.050 |
| Meta-epidemiology (narrow) | 0.002 | 0.001 |
| Meta-epidemiology (broad) | 0.013 | 0.030 |
| Bibliometrics | 0.004 | 0.004 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.003 | 0.001 |
| Open science | 0.002 | 0.001 |
| Research integrity | 0.002 | 0.002 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".