Measuring early childhood development in Brazil: validation of the Caregiver Reported Early Development Instruments (CREDI)
Bibliographic record
Abstract
The present study aims to analyze the psychometric properties and general validity of the Caregiver Reported Early Development Instruments (CREDI) short form for the population-level assessment of early childhood development for Brazilian children under age 3. The study analyzed the acceptability, test-retest reliability, internal consistency and discriminant validity of the CREDI short-form tool. The study also analyzed the concurrent validity of the CREDI with a direct observational measure (Inter-American Development Bank's Regional Project on Child Development Indicators; PRIDI). The full sample includes 1,265 Brazilian caregivers of children from 0 to 35 months (678 of which comprising an in-person sample and 587 an online sample). Results from qualitative interviews suggest overall high rates of acceptability. Most of the items showed adequate test-retest reliability, with an average agreement of 84%. Cronbach's alpha suggested adequate internal consistency/inter-item reliability (α > 0.80) for the CREDI within each of the six age groups (0–5, 6–11, 12–17, 18–23, 24–29 and 30–35 months of age). Multivariate analyses of construct validity showed that a significant proportion of the variance in CREDI scores could be explained by child gender and family characteristics, most importantly caregiver-reported cognitive stimulation in the home (p < 0.0001). Regarding concurrent validity, scores on the CREDI were significantly correlated with overall PRIDI scores within the in-person sample at r = 0.46 (p < 0.001). The results suggested that the CREDI short form is a valid, reliable, and acceptable measure of early childhood development for children under the age of 3 years in Brazil. O presente estudo visa analisar as propriedades psicométricas e a validade geral do formulário curto dos Instrumentos sobre o Desenvolvimento na Primeira Infância Relatado por Cuidados (CREDI) para avaliação em nível populacional do desenvolvimento na primeira infância de crianças brasileiras com menos de três anos. O estudo analisou a aceitabilidade, a confiabilidade teste-reteste, a consistência interna e a validade discriminante da ferramenta CREDI. O estudo também analisou a validade concorrente do CREDI com uma medida observacional direta (Projeto Regional sobre os Indicadores de Desenvolvimento na Infância do Banco Interamericano de Desenvolvimento; PRIDI). A amostra total inclui 1.265 cuidadores brasileiros de crianças de 0 a 35 meses (678 em uma amostra presencial e 587 em uma amostra on-line). Os resultados das entrevistas qualitativas sugerem altas taxas gerais de aceitabilidade. A maior parte dos itens mostrou confiabilidade teste-reteste adequada, com concordância média de 84%. O coeficiente alfa de Cronbach sugeriu consistência interna/confiabilidade entre itens (α > 0,80) para o CREDI em cada uma das seis faixas etárias (0-5 α = 6-11, 12-17, 18-23, 24-29 e 30-35 meses de idade). As análises multivariadas da validade do constructo mostraram que uma proporção significativa da variação nas pontuações do CREDI pode ser explicada pelo sexo da criança e pelas características familiares, mais importante o estímulo cognitivo em casa relatado pelo cuidador (p < 0,0001). Com relação à validade concorrente, as pontuações do CREDI foram significativamente correlacionadas às pontuações gerais do PRIDI na amostra presencial em r = 0,46 (p < 0,001). Os resultados sugerem que o formulário curto CREDI é uma medida válida, confiável e aceitável de desenvolvimento na primeira infância para crianças com menos de três anos no Brasil.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".