Measuring the complex syntax of school‐aged children in language sample analysis: A known‐groups validation study
Bibliographic record
Abstract
BACKGROUND: Complex syntax is affected by developmental language disorder (DLD) during the school years. Targeting areas of syntactic difficulty for children with DLD may yield useful assessment techniques. AIMS: To determine whether wh-movement can be measured in language samples from typically developing mono- and bilingual school-aged children, and, if so, to provide preliminary evidence of validity by comparison with traditional measures of syntax in a cross-sectional, known-groups design. METHODS & PROCEDURES: Participants were 48 typically developing children recruited from the Canadian province of Nova Scotia in four groups: monolingual English and bilingual French-English children in early (7-8 years of age) and late (11-12 years of age) elementary school. Language samples were collected and analysed with mean use of wh-movement, mean length of utterance and clausal density. These measures were compared for effects of age, bilingual development and elicitation task. OUTCOMES & RESULTS: The results from all measures closely paralleled each other, providing preliminary evidence of validity. Wh-movement-based and traditional measures demonstrated similar age-related and discourse genre effects. Neither demonstrated an effect of mono- versus bilingual development. CONCLUSIONS & IMPLICATIONS: The results confirm research interest in syntactic movement as an area of language assessment. Further research is required to understand its application to clinical populations. What this paper adds What is already known on the subject Complex syntax is known to be an area of difficulty for children with DLD. Certain syntactic constructions appear to be particularly difficult for these children. Assessments targeting these areas of difficulty are emerging. What this paper adds to existing knowledge The paper compares traditional measures of syntax with measures based on wh-movement. It shows similar results for both types of measures, suggesting construct and convergent validity. Results suggest that syntactic movement is an age-appropriate area of assessment for elementary school-aged children's language. What are the potential or actual clinical implications of this work? Language sample assessment measures based on wh-movement appear promising. The impact of task effects of the discourse genre on assessing syntax must be carefully considered in research and clinical practice.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.002 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.002 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".