Core outcome sets for spinal and associated limb, trunk, abdomen or pelvic pain: A systematic review
Bibliographic record
Abstract
BACKGROUND: Spinal pain is a significant global health issue, affecting millions and ranking as one of the leading causes of disability worldwide. Despite the wide scope of research conducted on spinal and associated pain, the lack of standardised core outcome measures poses challenges for comparing and synthesising research data. Core Outcome Sets (COSs) are intended to harmonise assessment and facilitate comparison across studies. This review aimed to identify, map, and examine published core outcome sets (COSs) designed for the assessment of spinal pain-including cervical, thoracic, lumbar-and spinal-related limb, trunk, abdomen, or pelvic pain. It also sought to synthesise consistent outcome domains across these COSs, categorising them by anatomical region and measurement type, including patient-reported, physical, biological, psychological, social, and environmental measures. METHODS: This systematic review followed PRISMA guidelines and was registered with PROSPERO. A comprehensive literature search of 13 electronic databases and grey literature sources was conducted from 2000 to April 2025. Two independent reviewers assessed study eligibility and quality using predefined criteria. Data extraction was performed to identify core outcome domains, and a thematic analysis was conducted to categorise domains based on anatomical regions, patient-reported outcomes, performance measures, and biopsychosocial factors. RESULTS: Thirteen studies met inclusion criteria, addressing core outcome sets for cervical (n = 4), thoracolumbar (n = 1), and lumbar (n = 8) spinal regions. Patient-reported outcome measures were the most frequently recommended outcome type. The most commonly endorsed domains were physical function n = 9 (100%), pain intensity n = 8 (88.9%), participation in work or daily activities n = 7 (77.8%), and disability n = 6 (66.7%). However, few studies incorporated psychological, social, environmental, or physiological domains, highlighting critical gaps in the multidimensional assessment of spinal pain. CONCLUSION: This systematic review identified key domains in current use and significant gaps in biopsychosocial and biological measurement. Findings will support researchers, clinicians, and policymakers in selecting appropriate outcomes for spinal pain research and practice. A Delphi study to develop an internationally agreed "Essential Universal Set" for spinal pain, inclusive of multidimensional biopsychosocial domains, is a sound next step.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.034 | 0.131 |
| Meta-epidemiology (narrow) | 0.002 | 0.001 |
| Meta-epidemiology (broad) | 0.012 | 0.013 |
| Bibliometrics | 0.018 | 0.016 |
| Science and technology studies | 0.002 | 0.002 |
| Scholarly communication | 0.004 | 0.005 |
| Open science | 0.003 | 0.003 |
| Research integrity | 0.002 | 0.002 |
| Insufficient payload (model declined to judge) | 0.005 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".