Feasibility and Reliability of a Mailed Questionnaire to Obtain Visual Analogue Scale Valuations for Health States Defined by the Health Utilities Index Mark 3
Bibliographic record
Abstract
To establish the generalizability (external validity) of the Health Utilities Index Mark 3 (HUI3) as a single-summary score generic outcome measure in numerous countries/subgroups (including children), repeated studies of community preferences should be performed in various settings. In performing multiple HUI3 studies, a mailed questionnaire approach, if feasible and reliable, might be substituted for oral interviews. In the present study, we assessed the feasibility and reliability of a mailed questionnaire approach originally developed for the EQ-5D, for the purpose of collecting Visual Analogue Scale (VAS) valuations from parents as surrogate responders for 65 pediatric HUI3 health states and for the state of being dead. Untransformed mean VAS scores of the health states and scores converted into preliminary Standard Gamble (SG)-utilities were compared with Canadian and French multiattribute utility estimates. A random sample of 1920 parents of schoolchildren (aged 4 to 13) received a mailed questionnaire. Each parent was asked to rate 6 HUI3 health states on a 0 to 100 VAS. Response was 70%. Mean completion time was 20 minutes (SD 9). The questionnaire was rated difficult by only 9%. The current format was, however, inappropriate for valuing the state of being dead. Interrater reliability of health state valuations was.87. Spearman's rank correlations, Pearson-R correlations and intra class correlation coefficients (ICCs) between untransformed VAS valuations and Canadian/French utility estimates were > or =.87. However, preliminary SG-utilities showed diminished ICCs (.71 to.72). The data support the feasibility and reliability of mailed HUI3 valuation questionnaires to a considerable extent, but further methodological studies regarding other formats and different populations are recommended.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.031 | 0.021 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".