Can the standard gamble and rating scale be used to measure quality of life in rhinoconjunctivitis? Comparison with the RQLQ and SF‐36
Bibliographic record
Abstract
BACKGROUND: With interest in health economics growing, it is important to know whether utilities may be used to measure health-related quality of life in patients with rhinoconjunctivitis. The objective was to compare the validity and measurement properties of disease-specific versions of the standard gamble and rating scale with those of the Rhinoconjunctivitis Quality of Life Questionnaire (RQLQ) and the Short-Form 36 (SF-36). METHODS: One hundred adults with symptomatic rhinoconjunctivitis participated in a 5 week observational study, completing the standard gamble, rating scale, RQLQ and SF-36 at baseline and after 1 and 5 weeks. Symptom diaries were completed for 1 week before each follow-up visit. RESULTS: Reliability was highest for the RQLQ (intraclass correlation coefficient = 0.97), followed by the rating scale (0.75), the SF-36 physical (0.75), the SF-36 mental (0.74) and the standard gamble (0.12). The responsiveness index was highest for the RQLQ (0.76), followed by the rating scale (0.56) and the SF-36 mental (0.28). Both cross-sectional and longitudinal validity were strongest for the RQLQ and the rating scale. CONCLUSIONS: Both the rating scale and the RQLQ have strong evaluative and discriminative properties. The SF-36 has acceptable discriminative properties but its evaluative properties are poor. All measurement properties for the standard gamble are inadequate. Poor correlation between the standard gamble and the rating scale indicates that utilities cannot be derived from rating scale data.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".