Minimally Important Differences in Patient or Proxy-Reported Outcome Studies Relevant to Children: A Systematic Review
Bibliographic record
Abstract
CONTEXT: No study has characterized and appraised all anchor-based minimally important differences (MIDs) associated with patient-reported outcome (PRO) instruments in pediatric studies. OBJECTIVE: To complete a comprehensive systematic survey and appraisal of published anchor-based MIDs associated with PRO instruments used in children. DATA SOURCES: Medline, Embase, and PsycINFO (1989 to February 11, 2015). STUDY SELECTION: Studies reporting empirical ascertainment of anchor-based MIDs among PROs used in pediatric care. DATA EXTRACTION: All pertinent data items related to the characteristics of PRO instruments, anchors, and MIDs. RESULTS: Of 4179 unique citations, 30 studies (including 32 cohorts) proved eligible and reported on 28 unique PROs (8 generic, 13 disease-specific, 5 symptoms-specific, 2 function-specific), with 9 (32%) classified as patient-reported, 11 (39%) proxy-reported, and 8 (29%) both patient- and proxy-reported. Of the 30 studies, we rated 14 (44%) as providing highly credible estimates of the MID. Most cohorts (n = 20, 62%) recorded patients’ direct response to the target PRO and the use of an independent standard of comparison (n = 25, 78%). Most, however, failed to effectively report measurement properties of the anchor (n = 24, 75%). LIMITATIONS: We have not yet addressed the measurement properties of instrument to measure credibility; our search was restricted to 3 electronic sources, and we used a single data abstractor. CONCLUSIONS: Our study found 28 PROs that have been developed for children, with fewer than half providing credible estimates. Clinicians, clinical trialists, systematic reviewers, and guideline developers seeking to effectively summarize and interpret results of studies addressing PROs in child health are likely to find our comprehensive compendium of MIDs of use, both in providing best estimates of MIDs and identifying credible estimates.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.049 | 0.272 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.011 | 0.009 |
| Bibliometrics | 0.014 | 0.014 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.004 | 0.004 |
| Open science | 0.003 | 0.003 |
| Research integrity | 0.003 | 0.002 |
| Insufficient payload (model declined to judge) | 0.004 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".