A pilot study on the validity and psychometric properties of the electronic EQ-5D-5L in routine clinical practice
Bibliographic record
Abstract
BACKGROUND: Electronic measurement of health-related quality of life (HRQOL) may facilitate timely and regular assessments in routine clinical practice. This study evaluated the validity and psychometric properties of an electronic version of the EQ-5D-5L (e-EQ-5D-5L) in Chinese patients with chronic knee and/or back problems. METHODS: 151 Chinese subjects completed an electronic version of the Chinese (Hong Kong) EQ-5D-5L when they attended a primary care or orthopedics specialist out-patient clinic in Hong Kong. They also completed the Chinese Western Ontario and McMaster University Osteoarthritis Index (WOMAC), a Pain Rating Scale, and a structured questionnaire on socio-demographics, co-morbidities and health service utilization. 32 subjects repeated the e-EQ-5D-5L two weeks after the baseline. 102 subjects completed e-EQ-5D-5L and 99 completed the Global Rating on Change Scale at three-month clinic follow up. Construct validity was assessed by the association of EQ-5D-5L scores with external criterion of WOMAC scores. We tested mean differences of WOMAC scores between adjacent response levels of the EQ-5D-5L dimensions by one-way ANOVA, test-retest reliability by intra-class correlation, sensitivity by known group comparisons and responsiveness by changes in EQ-5D-5L scores over 3 months. RESULTS: There was an association between EQ-5D-5L and WOMAC scores. Mean WOMAC scores increased with the increase in adjacent response levels of EQ-5D-5L dimensions. Test-retest intraclass correlation coefficient (ICC) of EQ-5D-5L utility and EQ-VAS scores were 0.76 and 0.83, respectively, indicating good reliability. There were significant differences in the proportions reporting limitations in the EQ-5D-5L dimensions, the utility and VAS scores between the mild and severe pain groups (utility = 0.28, p = 0.001; VAS = 11.46, p < 0.001), and between primary care and specialist out-patient clinic patients (utility = 0.15, p = 0.001; VAS = 10.21, p < 0.001), supporting sensitivity. Among those reporting 'better' global health at three-months, their EQ-5D-5L utility and EQ-VAS scores were significantly increased from baseline (utility = 0.18, p < 0.001; VAS = 10.75, p = 0.005). CONCLUSIONS: The electronic version of the EQ-5D-5L is valid, reliable, sensitive and responsive in the measurement of HRQOL in Chinese patients with chronic knee or back pain in routine clinical practice.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.021 | 0.042 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.001 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.001 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.001 |
| Insufficient payload (model declined to judge) | 0.002 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".