MétaCan
Menu
Retour à la cohorte
Enregistrement W4317207440 · doi:10.1136/bmjgh-2022-009827

The creation of the Global Scales for Early Development (GSED) for children aged 0–3 years: combining subject matter expert judgements with big data

2023· article· en· W4317207440 sur OpenAlexaff
Gareth McCray, Dana Charles McCoy, Patricia Kariger, Magdalena Janus, Maureen M. Black, Susan M. Chang, Fahmida Tofail, Iris Eekhout, Marcus Waldman, Stef van Buuren, Rasheda Khanam, Sunil Sazawal, Ambreen Nizar, Yvonne Schönbeck, Arsène Zongo, Alexandra Brentani, Yunting Zhang, Tarun Dua, Vanessa Cavallera, Abbie Raikes, Ann M. Weber, Kieran Bromley, Abdullah H Baqui, Arunangshu Dutta, Muhammad Imran Nisar, Symone Detmar, Romuald Anago, Pacifico Mercadante, Fan Jiang, Raghbir Kaur, Katelyn Hepworth, Marta Rubio‐Codina, Samuel Kembou Nzalé, Salahuddin Ahmed, Gill A. Lancaster, Melissa Gladstone

Notice bibliographique

RevueBMJ Global Health · 2023
Typearticle
Langueen
DomaineSocial Sciences
ThématiqueEarly Childhood Education and Development
Établissements canadiensMcMaster University
Organismes subventionnairesJacobs FoundationKing Baudouin Foundation United StatesChildren's Investment Fund FoundationBernard van Leer FoundationWorld Health OrganizationBill and Melinda Gates Foundation
Mots-clésRasch modelSubject-matter expertPolytomous Rasch modelPsychologyPopulationApplied psychologyItem response theorySet (abstract data type)PsychometricsScale (ratio)Item bankDevelopmental psychologyArtificial intelligenceComputer scienceMedicineExpert systemGeography

Résumé

récupéré en direct d'OpenAlex

INTRODUCTION: With the ratification of the Sustainable Development Goals, there is an increased emphasis on early childhood development (ECD) and well-being. The WHO led Global Scales for Early Development (GSED) project aims to provide population and programmatic level measures of ECD for 0-3 years that are valid, reliable and have psychometrically stable performance across geographical, cultural and language contexts. This paper reports on the creation of two measures: (1) the GSED Short Form (GSED-SF)-a caregiver reported measure for population-evaluation-self-administered with no training required and (2) the GSED Long Form (GSED-LF)-a directly administered/observed measure for programmatic evaluation-administered by a trained professional. METHODS: We selected 807 psychometrically best-performing items using a Rasch measurement model from an ECD measurement databank which comprised 66 075 children assessed on 2211 items from 18 ECD measures in 32 countries. From 766 of these items, in-depth subject matter expert judgements were gathered to inform final item selection. Specifically collected were data on (1) conceptual matches between pairs of items originating from different measures, (2) developmental domain(s) measured by each item and (3) perceptions of feasibility of administration of each item in diverse contexts. Prototypes were finalised through a combination of psychometric performance evaluation and expert consensus to optimally identify items. RESULTS: We created the GSED-SF (139 items) and GSED-LF (157 items) for tablet-based and paper-based assessments, with an optimal set of items that fit the Rasch model, met subject matter expert criteria, avoided conceptual overlap, covered multiple domains of child development and were feasible to implement across diverse settings. CONCLUSIONS: State-of-the-art quantitative and qualitative procedures were used to select of theoretically relevant and globally feasible items representing child development for children aged 0-3 years. GSED-SF and GSED-LF will be piloted and validated in children across diverse cultural, demographic, social and language contexts for global use.

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction distillée sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Apprise à partir de 10 348 étiquettes directes de Codex et de 10 348 étiquettes directes de Gemma. Le mode candidate est l'union des têtes enseignantes seuillées; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont ni des étiquettes humaines ni des étiquettes directes de modèles de pointe.

score de la tête « metaresearch » (Codex)0,002
score de la tête « metaresearch » (Gemma)0,000
Version: codex-gemma-dda1882f352aStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Observationnel · Signal consensuel: Observationnel
GenreSignal candidat: Empirique · Signal consensuel: Empirique
Score de désaccord entre enseignants0,238
Score d'incertitude au seuil0,900

Scores Codex et Gemma par catégorie

CatégorieCodexGemma
Métarecherche0,0020,000
Méta-épidémiologie (sens strict)0,0000,000
Méta-épidémiologie (sens large)0,0000,000
Bibliométrie0,0000,001
Études des sciences et des technologies0,0010,000
Communication savante0,0000,000
Science ouverte0,0010,000
Intégrité de la recherche0,0000,000
Charge utile insuffisante (le modèle a refusé de juger)0,0000,000

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,069
Tête enseignante GPT0,409
Écart entre enseignants0,340 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule tête enseignante, pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeObservationnel
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations48
Publié2023
Routes d'admission1
Résumé présentoui

Explorer davantage

Même revueBMJ Global HealthMême sujetEarly Childhood Education and DevelopmentTravaux en français237 207