Building partnerships, capacity, and knowledge through a use of newly linked child development and education datasets in Ontario, Canada.
Notice bibliographique
Résumé
ObjectivesThe objective of this study was to establish a partnership between a university and a jurisdictional education body (Education Quality and Assessment Organization, EQAO) which would allow creation of a linked dataset from kindergarten to later grades in order to examine educational trajectory in mathematics in Ontario. ApproachBuilding on mutual goals of improving the understanding of children’s learning trajectories, we developed a project with an investigator team that included university researchers and representatives of the provincial educational assessment body, to link a database of child development status in kindergarten (Early Development Instrument/EDI data, including neighbourhood socioeconomic/SES index) with academic assessment EQAO data, and received research funding. A deterministic matching process was employed to match the datasets. We examined differences between the unmatched and fully matched cases and constructed a growth mixture model of math scores in grades 3, 6 and 9, with key EDI/SES variables as covariates. ResultsDespite lacking a common identifier, we successfully matched approximately 50% of the EDI cases from 2002-2014 (n=183,771). Effect sizes indicated negligible differences between matched and unmatched, except for SES and child development status, which were poorer for unmatched group. A 3-class solution was the best fit for a 20,000-person subsample of math trajectories based on AIC, BIC, ICL, and entropy values as well as sufficiently high proportions of posterior probabilities, which indicate confidence in class membership. 61% of sample showed steady moderate-high achievement; 9% started high, but declined, and 30% deteriorated then improved. Males, children in low SES, and those with adequate kindergarten EDI outcomes had better math achievement trajectories than females, children in high SES, and those with poor kindergarten outcomes. ConclusionGiven the two datasets were collected without explicit linkage plan, the matching was only 50%, nevertheless resulting in a large database that allows study of early development antecedents of students’ educational trajectories. The partnership between university and EQAO ensures a wide dissemination of results in both academia and policy worlds.
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,014 | 0,044 |
| Méta-épidémiologie (sens strict) | 0,000 | 0,001 |
| Méta-épidémiologie (sens large) | 0,001 | 0,001 |
| Bibliométrie | 0,007 | 0,015 |
| Études des sciences et des technologies | 0,004 | 0,001 |
| Communication savante | 0,004 | 0,002 |
| Science ouverte | 0,002 | 0,006 |
| Intégrité de la recherche | 0,001 | 0,001 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,004 | 0,001 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».