Completeness of Outcomes Description Reported in Low Back Pain Rehabilitation Interventions: A Survey of 185 Randomized Trials
Bibliographic record
Abstract
Purpose: To assess reporting completeness of the most frequent outcome measures used in randomized controlled trials (RCTs) of rehabilitation interventions for mechanical low back pain. Methods: We performed a cross-sectional study of RCTs included in all Cochrane systematic reviews (SRs) published up to May 2013. Two authors independently evaluated the type and frequency of each outcome measure reported, the methods used to measure outcomes, the completeness of outcome reporting using a eight-item checklist, and the proportion of outcomes fully replicable by an independent assessor. Results: Our literature search identified 11 SRs, including 185 RCTs. Thirty-six different outcomes were investigated across all RCTs. The 2 most commonly reported outcomes were pain (n=165 RCTs; 89.2%) and disability (n=118 RCTs; 63.8%), which were assessed by 66 and 44 measurement tools, respectively. Pain and disability outcomes were found replicable in only 10.3% (n=17) and 10.2% (n=12) of the RCTs, respectively. Only 40 RCTs (21.6%) distinguished between primary and secondary outcomes. Conclusions: A large number of outcome measures and a myriad of measurement instruments were used across all RCTs. The reporting was largely incomplete, suggesting an opportunity for a standardized approach to reporting in rehabilitation science.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.339 | 0.244 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.008 | 0.002 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.005 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".