Replication Package for: Concurrent Validity of Short Tests for Early Childhood Development
Metadata & use
| Identifier | https://doi.org/10.60966/n2y42g0v |
|---|---|
| License | Creative Commons Attribution–NonCommercial–NoDerivs 3.0 IGO |
| Related Knowledge Product | |
| Citation |
Rubio-Codina, Marta (2016). Replication Package for: Concurrent Validity of Short Tests for Early Childhood Development. IDB Open Data. https://doi.org/10.60966/n2y42g0v |
| Published date | 2016-10-08 |
| Modified date | 2026-08-26 |
| Tags/Keywords | Child Development |
| Language |
|
| Temporal coverage | 2011-2011 |
| Country |
Colombia
|
| Region | Latin America and the Caribbean |
| Publisher |
Inter-American Development Bank
|
| Author |
Rubio-Codina, Marta
|
| Data collection type | Survey Data |
| Statistical type | Cross-sectional Data |
| Data structure | Structured Data |
| Data notes |
What is the Bayley test used in this dataset?The Bayley Scales of Infant and Toddler Development (Bayley-III) is a standardized assessment of early development. It measures Cognitive, Language, and Motor domains and reports composite scores that are normed with a mean of 100 and SD of 15. In this dataset, Bayley-III was administered alongside short tests (ASQ-3, Denver, Battelle, MacArthur) to examine concurrent validity in a Bogotá sample. Who was assessed and when?
What are the average development scores (Colombia sample) on the Bayley test?Across the full sample (n=1,311), Bayley-III composite score means were:
- Cognitive: 98.4 (SD 8.9) Reminder: Bayley composite scores are scaled to mean=100, SD=15 in the reference norming sample. This Bogotá distribution reflects this specific study population and age mix. Does the dataset include tests beyond the Bayley test?Yes. Variables and scripts in the package indicate concurrent measures from:
- ASQ-3 (Ages & Stages Questionnaires) How should I interpret the Bayley composite scores?
Are there citywide or national benchmarks to compare with these Bogotá results?The dataset is designed for concurrent validity across instruments rather than for official city- or national-level benchmarks. For benchmarking, you would either:
- Compare against Bayley-III norms (mean 100, SD 15), and/or What else is included to help researchers?
What are the dataset’s limitations?
What kinds of policy decisions can be informed by this dataset?
What is the Bayley Scales of Infant and Toddler Development?The Bayley Scales are considered the “gold standard” diagnostic test for measuring developmental levels of infants and toddlers. They assess multiple domains including cognition, language, and motor skills, and are widely used in controlled research and clinical settings. In the Bogota study, the Bayley-III was administered by trained psychologists in standardized environments. How long does a Bayley assessment take and what happens during it?The Bayley-III was administered in public libraries or childcare centers, typically requiring a dedicated session with trained psychologists. Each child was tested individually in quiet, well-lit environments, with caregivers present. The dataset notes that Bayley-III administration is time-consuming and requires specialized training. Bayley Scales vs ASQ-3: which developmental screening is better?The dataset finds that ASQ-3 performs poorly under 31 months, while Denver-II shows stronger validity. ASQ-3’s cognitive, language, and fine motor scales had low concurrent validity with Bayley-III below 19 months. How reliable are the Bayley Scales for developmental delay, and what is their validity?The dataset confirms Bayley-III as the gold standard with sensitivity to differences in outcomes due to interventions. It was used as the benchmark for concurrent validity comparisons. |