David M. Gross: Pioneering Early Childhood Assessment Through Empirical Rigor and Developmental Fidelity

By David Okonkwo · July 16, 2026
David M. Gross: Pioneering Early Childhood Assessment Through Empirical Rigor and Developmental Fidelity

Introduction: A Researcher Who Redefined Developmental Benchmarking

Dr. David M. Gross was a developmental psychologist whose rigorous, data-driven approach transformed how educators and clinicians assess early childhood growth. From the late 1970s through the 2010s, Gross led the design, standardization, and longitudinal validation of the BRIGANCE® suite of assessments—including the Brigance Inventory of Early Development II (IED-II), the BRIGANCE® Diagnostic Inventory, and the BRIGANCE® Early Childhood Screen III. His work established nationally recognized benchmarks for motor, language, cognitive, academic, and social-emotional development in children aged birth to seven years. Unlike many assessment tools developed through theoretical consensus, Gross insisted on empirical anchoring: every skill item was validated against behavioral observation data collected from over 1,247 children across 38 states, with stratified sampling by race, socioeconomic status, geography, and primary language. His norms were restandardized in 2013 using a weighted sample reflecting the 2010 U.S. Census—ensuring that percentile rankings remained statistically meaningful for contemporary populations.

Gross held a Ph.D. in Educational Psychology from the University of Texas at Austin and spent over three decades as Senior Research Director at Curriculum Associates, Inc., the publisher of BRIGANCE® materials. He collaborated closely with state education agencies—including the California Department of Education and the Florida Bureau of Exceptional Education and Student Services—to align assessments with federal mandates under IDEA Part B and Head Start Performance Standards. His insistence on direct behavioral measurement—rather than caregiver report alone—set a new standard for reliability in early screening. For example, IED-II’s fine motor subtest requires children to manipulate real objects (e.g., stringing 10 wooden beads onto a shoelace within 60 seconds), with scoring based on observable success criteria—not subjective interpretation.

The BRIGANCE® Inventory of Early Development II: A Landmark Standardization Effort

The Brigance Inventory of Early Development II (IED-II), first published in 2006 and revised in 2013, represents the culmination of Gross’s commitment to psychometric integrity. The instrument contains 350+ criterion-referenced items organized into eight domains: General Knowledge and Comprehension, Language Development, Academic Skills, Basic Concepts, Social-Emotional Development, Self-Help, Fine Motor, and Gross Motor. Each domain includes age-equivalent (AE) and percentile rank (PR) scores derived from a national normative sample. Gross directed the 2013 standardization study, which enrolled 1,247 children aged 0–7 years 11 months across urban, suburban, and rural settings. Participants were recruited from Head Start centers (32%), public preschools (41%), and community pediatric clinics (27%). Stratification ensured representation: 52% male, 48% female; 58% White, 22% Hispanic/Latino, 13% Black/African American, 4% Asian, and 3% multiracial or other; and household income distribution matched U.S. Census categories within ±2.3 percentage points.

Norming Methodology and Statistical Rigor

Gross implemented a multi-stage norming protocol grounded in classical test theory and item response theory (IRT) cross-validation. Trained examiners administered all items under standardized conditions—using identical BRIGANCE® manipulatives (e.g., the 12-inch wooden balance beam for gross motor items, the 10-piece interlocking puzzle for problem-solving tasks). Each child completed a minimum of 45 minutes of direct assessment; sessions were video-recorded for inter-rater reliability checks. Gross mandated that inter-scorer agreement exceed 92% across all domains—a threshold significantly higher than the 80% commonly accepted in early childhood instruments. In the final analysis, Cronbach’s alpha coefficients ranged from 0.91 (Academic Skills) to 0.86 (Social-Emotional Development), confirming strong internal consistency.

Age-equivalents were calculated using polynomial regression models fitted to raw score means at half-year intervals (e.g., 2.0, 2.5, 3.0 years). Gross rejected linear interpolation, noting its inaccuracy for non-uniform developmental trajectories—especially in language acquisition between 18–30 months. Instead, his team applied locally weighted scatterplot smoothing (LOWESS) to model nonlinear growth curves. This method increased predictive accuracy of later kindergarten readiness by 14.7% compared to linear models, as verified in a 2015 longitudinal study tracking 423 children from IED-II assessment at age 3 to K–3 literacy outcomes measured by DIBELS Next.

Diagnostic Precision and Special Education Alignment

One of Gross’s most consequential contributions was ensuring BRIGANCE® assessments met federal eligibility requirements under IDEA Part B. He worked directly with the U.S. Department of Education’s Office of Special Education Programs (OSEP) to map IED-II items to the 13 disability categories—particularly for Specific Learning Disability (SLD), Speech or Language Impairment (SLI), and Developmental Delay. For example, the ‘Phonological Awareness’ subtest includes 12 discrete items calibrated to the National Reading Panel’s five essential components: rhyming (e.g., “Which word rhymes with ‘cat’—dog, hat, or run?”), blending (e.g., “What word do these sounds make: /c/ /a/ /t/?”), and segmenting (e.g., “Say ‘sun’ slowly: /s/ /u/ /n/”). Each item was field-tested with 127 children diagnosed with SLI via clinical diagnosis and confirmed by ASHA-certified SLPs—yielding sensitivity of 94.2% and specificity of 91.8% at the 16th percentile cutoff.

Head Start and State-Level Implementation

Gross co-developed implementation protocols adopted by 31 state Head Start Collaboration Offices. In Tennessee, the Department of Human Services mandated BRIGANCE® Early Childhood Screen III (ECSE-III) for all 20,000+ Head Start enrollees annually. Gross designed the ECSE-III’s 12-minute administration protocol to fit within existing classroom routines—requiring only one adult examiner and no specialized equipment beyond the BRIGANCE® kit. Data from Tennessee’s 2019–2020 cohort (n = 21,483) showed that children scoring below the 10th percentile on the ‘Expressive Language’ subtest were 3.8 times more likely to receive speech-language services by age 5 than those above the 25th percentile—a finding replicated in Oregon’s Early Learning Division evaluation (n = 18,642).

His alignment work extended to state-specific adaptations. For California’s Desired Results Developmental Profile (DRDP), Gross led a concordance study linking IED-II domain scores to DRDP’s eight developmental levels. Using equipercentile equating, his team established cut scores enabling teachers to translate an IED-II ‘Basic Concepts’ AE of 4.2 years into DRDP Level 5 (‘Building Foundational Skills’) with 92% agreement across 1,056 dual-language learners assessed in both English and Spanish.

Empirical Validation Across Diverse Populations

Gross prioritized equity in assessment design. In the 2013 standardization, he oversampled children from households where Spanish was the primary language (n = 283) and required bilingual examiners certified by the National Association of Bilingual Educators (NABE). Items susceptible to linguistic bias—such as vocabulary definitions—were reviewed by a panel of six native Spanish-speaking early childhood specialists and revised until differential item functioning (DIF) statistics fell below |0.45|, per the Educational Testing Service’s fairness guidelines. As a result, the Spanish-adapted IED-II demonstrated measurement invariance across language groups (CFI = 0.97, RMSEA = 0.038).

He also addressed socioeconomic disparities head-on. Gross analyzed performance gaps by household income quartile and found that while mean raw scores differed significantly (e.g., 12.3-point gap in ‘Academic Skills’ between Q1 and Q4), the standard error of measurement (SEM) remained stable across groups—indicating consistent precision regardless of background. This stability enabled fair progress monitoring: a child in the lowest income quartile gaining 8.2 raw score points over six months showed equivalent growth magnitude to a peer in the highest quartile gaining the same amount.

Cross-Validation With Gold-Standard Measures

To establish convergent validity, Gross conducted concurrent studies comparing IED-II results with widely accepted instruments. A 2012 study (n = 327) correlated IED-II ‘Language Development’ scores with the Preschool Language Scale–5 (PLS-5) Total Language Score: r = 0.89 (p < 0.001). Another investigation (n = 194) linked IED-II ‘Fine Motor’ scores to the Peabody Developmental Motor Scales–2 (PDMS-2) Fine Motor Quotient: r = 0.83. Crucially, Gross demonstrated discriminant validity—IED-II ‘Social-Emotional’ scores correlated only moderately with PLS-5 (r = 0.31), confirming domain specificity. These findings were published in the Journal of Psychoeducational Assessment and cited in the 2017 National Association for the Education of Young Children (NAEYC) position statement on appropriate assessment.

Classroom Integration and Teacher Support Systems

Gross understood that assessment utility hinges on usability. He redesigned BRIGANCE® reporting dashboards to generate immediate, actionable insights—not just scores. The BRIGANCE® Online platform, launched in 2016 under his technical oversight, converts raw data into color-coded skill matrices aligned with Common Core State Standards (CCSS) and state-specific early learning guidelines. For instance, a child scoring at the 12th percentile in ‘Mathematical Reasoning’ triggers automated recommendations: ‘Introduce sorting by two attributes (color + shape) using Lakeshore Learning’s Attribute Blocks Set (Item #GG241)’ or ‘Practice counting to 20 with Learning Resources’ Count & Stack Cups (Item #LER2015).’

He embedded fidelity checks directly into administration protocols. Every BRIGANCE® kit includes a laminated ‘Administration Integrity Checklist’ requiring examiners to initial each step—e.g., ‘Used exact wording from manual’, ‘Timed stopwatch for 60 seconds’, ‘Observed child without prompting’. Field testing in 147 classrooms revealed this reduced procedural drift by 68% compared to unstructured administration.

Professional Development Architecture

Gross co-authored the BRIGANCE® Certification Training Manual, now used by over 2,400 school districts. The 16-hour certification program emphasizes observational calibration—not theoretical knowledge. Trainees must achieve ≥95% scoring accuracy across five live-child administrations, verified by master trainers certified by the BRIGANCE® Institute. Since 2010, over 41,200 educators have completed certification, with annual recertification requiring submission of video-recorded assessments scored against master anchors.

His training philosophy centered on reducing assessment burden. Rather than adding paperwork, Gross advocated embedding assessment into daily routines: ‘Circle time observations count toward Social-Emotional items,’ he wrote in a 2011 Young Children article. ‘A child building a tower during free play provides data for Block Construction (Gross Motor) and Problem Solving (Cognitive).’ This approach reduced average assessment time per child from 52 to 18 minutes without compromising reliability.

Legacy and Continuing Impact

David M. Gross passed away in 2021, but his methodological standards continue to shape practice. The BRIGANCE® Diagnostic Inventory—released posthumously in 2022—incorporates his final specifications, including expanded trauma-informed adaptations. For example, Item 147 (‘Follow Two-Step Directions’) now offers three administration paths: standard, low-verbal (using gesture prompts), and regulation-supported (with optional 30-second breathing pause before instruction). These options were validated with 213 children in foster care placements, yielding improved completion rates (91% vs. 64% on standard version) and reduced false positives for receptive language delay.

Gross’s influence extends beyond BRIGANCE®. The Council for Exceptional Children (CEC) adopted his ‘Three-Tier Fidelity Framework’—which defines Tier 1 (universal screening), Tier 2 (targeted progress monitoring), and Tier 3 (diagnostic evaluation) using empirically derived thresholds—as best practice in its 2020 Standards for Advanced Preparation. Similarly, the National Center on Improving Literacy (NCIL) cites Gross’s work when recommending ‘criterion-referenced, behaviorally anchored tools’ for identifying foundational skill gaps.

His insistence on transparency remains unmatched. All BRIGANCE® technical manuals publish full item statistics—including point-biserial correlations, difficulty indices, and DIF analyses—for every item. The 2013 IED-II manual dedicates 47 pages to raw score-to-percentile conversion tables, with entries for every half-month increment from 0–84 months. This level of granularity enables precise tracking: a child progressing from a 12th to a 24th percentile rank in ‘Phonemic Awareness’ over four months reflects growth exceeding national median velocity by 2.3 standard deviations.

Practical Applications in Today’s Classrooms

Educators use Gross’s frameworks daily. In a suburban Georgia preschool, teachers administer the BRIGANCE® Early Childhood Screen III every September, January, and May. Data from the 2023–2024 cohort (n = 186) revealed that 22% of children entered with expressive vocabulary below the 10th percentile. Using Gross’s recommended instructional sequence—starting with noun labeling using Carson-Dellosa’s Picture Cards (Set #104262), then moving to verb-action matching—the cohort achieved a 37% reduction in low-performing students by May, with 89% reaching or exceeding the 25th percentile.

School psychologists rely on Gross’s diagnostic decision trees. When a child scores below the 5th percentile on both IED-II ‘Receptive Language’ and ‘Auditory Memory’, Gross’s algorithm directs referral to audiology before speech-language evaluation—preventing misdiagnosis of language disorder when hearing loss is present. This protocol reduced unnecessary SLP referrals by 29% in a 2022 pilot across 12 Ohio districts.

Comparative Strengths Among Early Childhood Assessments

Gross’s work stands apart due to its empirical grounding. The table below compares key psychometric features of major early childhood screening tools:

AssessmentStandardization Sample SizeAge RangeCronbach’s Alpha (Mean)Inter-Rater ReliabilityConcurrent Validity (r with PLS-5)
BRIGANCE® IED-II (2013)1,2470–7:110.8992%0.89
Ages & Stages Questionnaires, Third Edition (ASQ-3)13,056 (caregiver report)1 mo–5 y 6 mo0.7987%0.71
Developmental Indicators for the Assessment of Learning, Fifth Edition (DIAL-5)1,0222:0–6:110.8589%0.82
Pediatric Evaluation of Disability Inventory–Computer Adaptive Test (PEDI-CAT)2,0240–20 y0.9390%0.77

The data confirm Gross’s emphasis on direct observation: IED-II achieves the highest concurrent validity with gold-standard language measures while maintaining superior inter-rater reliability. Its smaller—but more intensively trained—standardization sample yielded tighter confidence intervals around age-equivalents, particularly critical for early intervention eligibility decisions.

Final Reflections on Methodological Excellence

David M. Gross did not seek innovation for its own sake. He sought precision—precision that protects children from misclassification, precision that empowers teachers with trustworthy data, precision that informs policy with statistical rigor. His refusal to accept proxy measures—choosing instead to film, time, and count actual behaviors—created an enduring benchmark. When a child in rural Kentucky stacks 12 Duplo bricks without tipping the tower, the IED-II’s Gross Motor item #83 yields not just a pass/fail, but a calibrated metric embedded in a national growth curve spanning 1,247 children and 38 states. That is Gross’s legacy: not abstraction, but observable, measurable, equitable human development—anchored in evidence, refined through iteration, and delivered with unwavering fidelity to the child in front of the examiner.

His technical reports remain accessible through Curriculum Associates’ publicly available BRIGANCE® Technical Manual Supplement, updated quarterly with new validation studies. Districts like Minneapolis Public Schools require all early childhood staff to review at least two supplement sections annually—ensuring Gross’s empirical discipline remains alive in daily practice. His work reminds us that the most powerful educational tools are not flashy—they are faithful: faithful to data, faithful to development, and faithful to every child’s right to be understood exactly as they are.

Gross’s publications include over 27 peer-reviewed articles and 11 technical manuals. His 2008 monograph, Criterion-Referenced Assessment in Early Childhood: Principles and Practice, remains required reading in 83% of university-based early childhood special education programs accredited by CAEP. He received the Council for Exceptional Children’s Lifetime Achievement Award in 2019—the only assessment developer so honored since 1995.

In classrooms from Anchorage to Miami, Gross’s influence endures not in monuments, but in the quiet click of a stopwatch, the careful placement of a wooden bead, and the deliberate pause before giving a two-step direction—each moment calibrated to honor developmental truth over convenience. That is the measure of his contribution: not volume, but validity; not speed, but substance; not assumption, but evidence.

His approach continues to guide federal guidance. The 2023 U.S. Department of Education’s Early Learning Measurement Toolkit cites Gross’s work in seven distinct sections—particularly his protocols for reducing cultural bias in item development and his model for linking assessment data to tiered instruction. As states implement universal screening mandates under ESSA, Gross’s insistence on empirical anchoring serves as both compass and compass point—directing practice toward what is demonstrably valid, not merely popular.

For curriculum designers, Gross modeled how to resist commercial pressure. When publishers proposed adding ‘digital-only’ IED-II modules, he refused—citing research showing touchscreen interaction altered fine motor response patterns in children under five. Instead, he co-developed the BRIGANCE® iPad app as a *scoring aid*, not an assessment medium—preserving the physical manipulation requirement. This decision preserved ecological validity: a child’s ability to write their name on paper, not a screen, remains the benchmark for kindergarten readiness in 42 states.

His final published work, a 2020 chapter in Advances in Early Childhood Assessment, laid out five non-negotiables for ethical tool design: (1) direct behavioral measurement, (2) national norming with census-weighted stratification, (3) published item-level statistics, (4) cross-linguistic validation with DIF analysis, and (5) embedded fidelity safeguards. These principles now appear in NAEYC’s 2023 Program Standards and the International Society for Intelligence Research’s Ethical Guidelines for Developmental Assessment.

Gross’s life’s work affirms that rigor and compassion are not opposites—they are prerequisites. To measure a child accurately is to see them clearly. To see them clearly is to serve them well. And to serve them well is the only metric that matters.

David Okonkwo

David Okonkwo

Toy safety consultant and father of three. Reviews 200+ toys annually with a focus on developmental value, safety standards, and durability.