Concurrent and construct validation of the short form of the Bruininks‐Oseretsky Test of Motor Proficiency and the Movement‐ABC when administered under field conditions: implications for screening
Bibliographic record
Abstract
RATIONALE: Among the most widely used instruments to assess developmental co-ordination disorder (DCD) in children are the Bruininks-Oseretsky Test of Motor Proficiency (BOTMP) and the Movement Assessment Battery for Children (M-ABC). However, there is little research on agreement between these tests, when administered to children in field-based settings by trained non-clinicians. METHOD: Ten of 75 schools participating in a larger study were randomly selected. All children in grade 4 (n= 340) in each of these schools were assessed at the same time using both the BOTMP-SF and the M-ABC in May of 2005. The order of tests was balanced, with an average gap in time between tests of 10-15 min. All tests were administered by trained research assistants. RESULTS: The correlation between tests was moderate (r= 0.50, P < 0.01). Kappas were low at the fifth (k= 0.19) and 15th (k= 0.29) percentile cut-points, which are generally used to identify cases of DCD. Re-analysis using the relative improvement over chance (RIOC) statistic, however, revealed slightly better agreement at both cut-points (fifth percentile, RIOC = 0.29; 15th percentile, RIOC = 0.47). Children who scored as probable for DCD on both motor tests, as well as on only the BOTMP-SF, had higher body mass index, poorer physical fitness and lower levels of teacher-reported physical ability than those positive for DCD on the M-ABC only or those who scored negatively on both tests. DISCUSSION: In general, the agreement between tests, even after adjustment for RIOC, was poor. Children identified with poor motor competence by both tests or by the BOTMP-SF only are at particular risk for poor physical fitness, overweight/obesity and physical inactivity. It appears that each assessment measures different dimensions of motor ability but that under field-based conditions the M-ABC may be less useful when applied by non-clinicians.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".