Assessment of Patient-Reported Outcome Instruments to Assess Chronic Low Back Pain
Bibliographic record
Abstract
Objective: To identify patient-reported outcome (PRO) instruments that assess chronic low back pain (cLBP) symptoms (specifically pain qualities) and/or impacts for potential use in cLBP clinical trials to demonstrate treatment benefit and support labeling claims. Design: Literature review of existing PRO measures. Methods: Publications detailing existing PRO measures for cLBP were identified, reviewed, and summarized. As recommended by the US Food & Drug Administration (FDA) PRO development guidance, standard measurement characteristics were reviewed, including development history, psychometric properties (validity and reliability), ability to detect change, and interpretation of observed changes. Results: Thirteen instruments were selected and reviewed: Low Back Pain Bothersomeness Scale, Neuropathic Pain Symptom Inventory, PainDETECT, Pain Quality Assessment Scale Revised, Revised Short Form McGill Pain Questionnaire, Low Back Pain Impact Questionnaire, Oswestry Disability Index, Pain Disability Index, Roland-Morris Disability Questionnaire, Brief Pain Inventory and Brief Pain Inventory Short Form, Musculoskeletal Outcomes Data Evaluation and Management System Spine Module, Orebro Musculoskeletal Pain Questionnaire, and the West Haven-Yale Multidimensional Pain Inventory Interference Scale. The instruments varied in the aspects of pain and/or impacts that they assessed, and none of the instruments fulfilled all criteria for use in clinical trials to support labeling claims based on recommendations outlined in the FDA PRO guidance. Conclusions: There is an unmet need for a validated PRO instrument to evaluate cLBP-related symptoms and impacts for use in clinical trials.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.011 | 0.009 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.004 | 0.001 |
| Bibliometrics | 0.001 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".