International Consensus on Standard Outcome Measures for Neurodevelopmental Disorders
Bibliographic record
Abstract
Importance: The use of evidence-based standardized outcome measures is increasingly recognized as key to guiding clinical decision-making in mental health. Implementation of these measures into clinical practice has been hampered by lack of clarity on what to measure and how to do this in a reliable and standardized way. Objective: To develop a core set of outcome measures for specific neurodevelopmental disorders (NDDs), such as attention-deficit/hyperactivity disorder (ADHD), communication disorders, specific learning disorders, and motor disorders, that may be used across a range of geographic and cultural settings. Evidence Review: An international working group composed of clinical and research experts and service users (n = 27) was convened to develop a standard core set of accessible, valid, and reliable outcome measures for children and adolescents with NDDs. The working group participated in 9 video conference calls and 8 surveys between March 1, 2021, and June 30, 2022. A modified Delphi approach defined the scope, outcomes, included measures, case-mix variables, and measurement time points. After development, the NDD set was distributed to professionals and service users for open review, feedback, and external validation. Findings: The final set recommends measuring 12 outcomes across 3 key domains: (1) core symptoms related to the diagnosis; (2) impact, functioning, and quality of life; and (3) common coexisting problems. The following 14 measures should be administered at least every 6 months to monitor these outcomes: ADHD Rating Scale 5, Vanderbilt ADHD Diagnostic Rating Scale, or Swanson, Nolan, and Pelham Rating Scale IV; Affective Reactivity Index; Children's Communication Checklist 2; Colorado Learning Disabilities Questionnaire; Children's Sleep Habits Questionnaire; Developmental-Disability Children's Global Assessment Scale; Developmental Coordination Disorder Questionnaire; Family Strain Index; Intelligibility in Context Scale; Vineland Adaptive Behavior Scale or Repetitive Behavior Scale-Revised and Social Responsiveness Scale; Revised Child Anxiety and Depression Scales; and Yale Global Tic Severity Scale. The external review survey was completed by 32 professionals and 40 service users. The NDD set items were endorsed by more than 70% of professionals and service users in the open review survey. Conclusions and Relevance: The NDD set covers outcomes of most concern to patients and caregivers. Use of the NDD set has the potential to improve clinical practice and research.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".