Patient-reported outcome measures for hip-related pain: a review of the available evidence and a consensus statement from the International Hip-related Pain Research Network, Zurich 2018
Bibliographic record
Abstract
Hip-related pain is a well-recognised complaint among active young and middle-aged active adults. People experiencing hip-related disorders commonly report pain and reduced functional capacity, including difficulties in executing activities of daily living. Patient-reported outcome measures (PROMs) are essential to accurately examine and compare the effects of different treatments on disability in those with hip pain. In November 2018, 38 researchers and clinicians working in the field of hip-related pain met in Zurich, Switzerland for the first International Hip-related Pain Research Network meeting. Prior to the meeting, evidence summaries were developed relating to four prioritised themes. This paper discusses the available evidence and consensus process from which recommendations were made regarding the appropriate use of PROMs to assess disability in young and middle-aged active adults with hip-related pain. Our process to gain consensus had five steps: (1) systematic review of systematic reviews; (2) preliminary discussion within the working group; (3) update of the more recent high-quality systematic review and examination of the psychometric properties of PROMs according to established guidelines; (4) formulation of the recommendations considering the limitations of the PROMs derived from the examination of their quality; and (5) voting and consensus. Out of 102 articles retrieved, 6 systematic reviews were selected and assessed for quality according to AMSTAR 2 (A MeaSurement Tool to Assess systematic Reviews). Two showed moderate quality. We then updated the most recent review. The updated literature search resulted in 10 additional studies that were included in the qualitative synthesis. The recommendations based on evidence summary and PROMs limitations were presented at the consensus meeting. The group makes the following recommendations: (1) the Hip and Groin Outcome Score (HAGOS) and the International Hip Outcome Tool (iHOT) instruments (long and reduced versions) are the most appropriate PROMs to use in young and middle-aged active adults with hip-related pain; (2) more research is needed into the utility of the HAGOS and the iHOT instruments in a non-surgical treatment context; and (3) generic quality of life measures such as the EuroQoL-5 Dimension Questionnaire and the Short Form Health Survey-36 may add value for researchers and clinicians in this field. We conclude that as none of the instruments shows acceptable quality across various psychometric properties, more methods studies are needed to further evaluate the validity of these PROMS-the HAGOS and iHOT-as well as the other (currently not recommended) PROMS.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.021 | 0.016 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.003 | 0.001 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".