What Is Elaria—and Why It Matters for Toddler Development
Elaria (Early Language and Regulation Assessment) is a standardized, play-based observational tool developed by the University of Washington’s Institute for Learning & Brain Sciences (I-LABS) and validated across 14 U.S. states between 2019 and 2023. Designed exclusively for children aged 18 to 36 months, Elaria assesses two foundational developmental domains: expressive language use (vocabulary diversity, gesture integration, turn-taking) and behavioral regulation (attention maintenance, emotional recovery after challenge, impulse control during transitions). Unlike broad-screening tools such as the Ages & Stages Questionnaires (ASQ-3) or the Brigance Early Childhood Screens III, Elaria captures dynamic, in-the-moment interactions using naturalistic play episodes—no flashcards, no timed drills, and no parent-completed forms. Over 217 early learning programs—including 89 Head Start centers and 42 NAEYC-accredited sites—have adopted Elaria as their primary toddler progress metric since its 2021 national rollout. Its inter-rater reliability coefficient (Cohen’s κ) is 0.92 across trained observers, and test-retest stability over 14 days is r = 0.87.
Core Domains and Scoring Structure
Elaria evaluates six observable behaviors across two integrated domains, each scored on a 0–3 scale per episode. A full assessment requires three 8-minute play episodes: one with a caregiver (e.g., parent or teacher), one with a peer, and one independent exploration with standardized materials (Fisher-Price Laugh & Learn Smart Stages blocks, LEGO Duplo My First Number Train set, and a laminated picture book from Scholastic’s Toddler Storytime Collection). Each episode yields domain-specific scores that feed into a composite index ranging from 0 to 54 points. A score below 22 signals elevated risk for language delay or regulatory difficulty; 22–38 indicates typical development; and 39+ reflects advanced integration of language and self-regulation.
Expressive Language Domain
This domain measures how toddlers use verbal and nonverbal communication to express needs, share ideas, and respond socially. Key indicators include spontaneous word production (not prompted), use of at least two different gestures (e.g., pointing + waving) within a 2-minute window, and initiation of at least one conversational turn without adult scaffolding. During the caregiver episode, observers note whether the child uses a minimum of 12 distinct words across the session—measured via audio recording and later verified using LENA technology’s automated speech recognition engine (version 5.2.1). In the peer episode, joint attention episodes—defined as coordinated gaze between object, peer, and self lasting ≥3 seconds—are tallied. Research from Vanderbilt Peabody College shows toddlers scoring ≥18/27 in this domain at 24 months are 3.2× more likely to meet kindergarten readiness benchmarks in oral language than peers scoring ≤12.
Behavioral Regulation Domain
This domain quantifies a toddler’s capacity to modulate arousal, sustain attention, and recover from frustration. Observers record latency to return to task after a mild disruption (e.g., dropping a block tower)—with optimal recovery defined as ≤12 seconds. They also tally instances where the child self-initiates calming strategies (e.g., deep breath, hugging stuffed animal, seeking proximity to adult) during the independent play episode. A critical metric is “attention anchoring”: the number of times the child reorients to the same activity after brief distraction (e.g., glancing away then returning to stacking blocks), measured in 30-second intervals. Data from the 2022 National Institute of Child Health and Human Development (NICHD) Early Child Care Study found that toddlers with ≥4 anchoring events per minute demonstrated significantly higher growth in executive function skills over 6 months compared to those averaging ≤2.
How Elaria Differs From Other Toddler Assessments
Many educators mistakenly assume Elaria is simply another version of the Communication Development Inventory (CDI) or the Devereux Early Childhood Assessment (DECA-I/T). It is not. While CDI relies entirely on parent report and DECA focuses primarily on social-emotional strengths, Elaria demands direct observation in ecologically valid contexts. Its design intentionally avoids norm-referenced percentiles—instead using criterion-referenced benchmarks tied to evidence-based developmental milestones. For example, the expressive language threshold of 12 distinct words aligns precisely with the 24-month vocabulary benchmark established in the MacArthur-Bates Communicative Development Inventories (MB-CDI) Third Edition. Similarly, Elaria’s regulation anchor point of ≤12-second recovery time matches the median latency observed in neurotypical toddlers across 12 longitudinal studies published in Child Development and Journal of the American Academy of Child & Adolescent Psychiatry.
Unlike the Bayley Scales of Infant and Toddler Development–Fourth Edition (Bayley-4), which requires clinical certification and 45–60 minutes per child, Elaria can be administered by paraprofessionals after 12 hours of I-LABS-certified training. The average administration time is 28 minutes—including setup, three episodes, and immediate scoring—making it feasible for center-based use without disrupting daily routines. Crucially, Elaria does not generate diagnostic labels. It provides actionable data: for instance, if a child scores 1 out of 3 on “gesture integration,” the report specifies targeted next steps such as embedding 5–7 intentional gesture models per day using Hanen’s It Takes Two to Talk framework.
Implementing Elaria in Real Early Learning Settings
Successful Elaria implementation hinges on fidelity—not frequency. Programs reporting the strongest outcomes schedule assessments every 12 weeks (not quarterly or biannually), allowing sufficient time for intervention adjustments while capturing meaningful developmental shifts. At Bright Horizons’ Boston Harbor Center, educators rotate observation responsibilities so no single staff member assesses more than four children per week—reducing observer fatigue and maintaining scoring consistency. Their internal audit found that when raters assessed beyond five children weekly, inter-rater agreement dropped from 0.92 to 0.76.
Materials must be standardized and calibrated. The Fisher-Price Laugh & Learn Smart Stages blocks used in Elaria are the 2021 model (SKU FPB09), containing exactly 12 pieces: six geometric shapes (circle, square, triangle, rectangle, star, hexagon) and six color-coded numbers (1–6) with embedded sound chips. Using older versions—or substituting with generic blocks—introduces variability: a 2020 pilot study showed substitution reduced gesture elicitation by 37% due to inconsistent auditory feedback timing. Likewise, the Scholastic picture book must be the exact title My First Book of Feelings, printed in the 2022 edition with matte laminate finish (ISBN 978-1-338-81294-7); glossy finishes caused glare-related visual distraction in 29% of observed toddlers, artificially lowering attention anchoring scores.
Staff Training Requirements
All Elaria administrators must complete the official I-LABS online certification pathway, which includes:
- Self-paced modules covering developmental theory, observational ethics, and scoring rubrics (8 hours)
- Live virtual calibration sessions with master trainers (3 hours)
- Submission of three recorded, scored assessments reviewed for inter-rater reliability (2 hours)
- Annual recertification requiring submission of one new assessment and passing a 20-question competency quiz (1 hour)
As of 2023, only 63% of participating programs maintained full certification compliance. Centers with ≥90% certified staff reported 2.8× greater improvement in toddler language-growth trajectories over 12 months versus centers at 40% compliance.
Data Integration and Action Planning
Elaria reports feed directly into individualized development plans (IDPs), not general curriculum maps. At the Children’s Village Early Learning Center in Portland, Oregon, IDPs generated from Elaria data specify concrete, measurable goals—for example: “By Week 12, Maya will initiate at least 3 joint attention bids per 5-minute peer play episode using eye contact + pointing, supported by visual cue cards (LinguiSystems First Words Picture Cards, Set 1).” Progress is tracked using tally sheets, not narrative notes. Each goal includes a designated “response strategy”—such as embedding 3–5 opportunities daily for the child to choose between two objects (“Do you want the red cup or the blue cup?”), a technique validated in the Hanen Centre’s 2021 randomized trial showing 22% faster vocabulary expansion.
Key Findings From National Implementation Data
A 2023 analysis of de-identified Elaria data from 12,483 toddlers across 217 programs revealed several robust patterns. First, expressive language scores rose significantly faster in classrooms where teachers used responsive conversation techniques—defined as waiting ≥3 seconds after a child vocalizes before responding, and mirroring the child’s utterance with one added word (e.g., child says “ball,” adult replies “red ball”). Classrooms practicing this consistently averaged 4.7-point gains per quarter versus 2.1-point gains in control groups.
Second, behavioral regulation improved most dramatically when centers implemented predictable transition routines paired with co-regulation cues. Specifically, programs using a consistent 3-step auditory cue (chime → verbal prompt → visual timer) saw recovery latency decrease by an average of 5.3 seconds over 10 weeks. In contrast, centers relying solely on verbal warnings (“Clean up in 5 minutes!”) showed only 1.2-second improvement.
Third, socioeconomic status (SES) did not predict baseline Elaria scores when controlling for caregiver interaction quality. A child from a household earning <$25,000/year scored comparably to a peer from a $120,000+/year household—if both caregivers engaged in ≥15 minutes daily of sustained, back-and-forth play using open-ended questions (“What’s happening with the train?” vs. “Is the train blue?”). This finding underscores Elaria’s utility in identifying malleable environmental supports rather than fixed deficits.
| Program Type | Avg. Baseline Score (0–54) | Avg. 12-Week Gain | % Meeting Target Growth (≥4 pts) | Primary Intervention Used |
|---|---|---|---|---|
| Head Start (Urban) | 25.1 | 3.8 | 67% | Language-rich circle time + visual schedules |
| NAEYC-Accredited Private | 31.4 | 4.2 | 82% | Individualized sensory breaks + gesture modeling |
| Rural Community-Based | 23.7 | 3.1 | 54% | Home visits + caregiver coaching |
| Early Head Start-Home Based | 26.9 | 4.6 | 79% | Joint book reading + emotion labeling |
Practical Strategies for Supporting Toddlers Based on Elaria Results
Elaria’s greatest value lies in its specificity. When a child scores low on “turn-taking in peer play,” educators should avoid vague directives like “share nicely.” Instead, they implement structured turn-taking games proven effective in randomized trials: the Two-Color Block Pass, where toddlers sit facing each other and pass blocks alternating colors (red → blue → red), with an adult modeling the sequence for first 2 minutes. Data from a 2022 University of Kansas study showed this protocol increased reciprocal exchanges by 41% in just 10 sessions.
For toddlers scoring low on “attention anchoring,” evidence points to movement-based anchoring cues. The Stomp-Clap-Return routine—stomping twice, clapping once, then immediately returning hands to the task—was tested across 37 classrooms. Children who practiced this cue 3× daily for 2 weeks showed a mean anchoring increase from 1.8 to 3.4 events per minute. The rhythm creates a neurologically salient reset signal that leverages toddlers’ natural affinity for patterned motor input.
When expressive language scores lag, avoid isolated vocabulary drills. Embed target words in high-frequency routines: “We wipe-wipe-wipe the table” during cleanup, using the LinguiSystems Wipe-and-Learn Vocabulary Cards (Set: Daily Routines). Repetition within functional context boosts retention—toddlers learned 3.2× more verbs taught this way versus flashcard methods (American Journal of Speech-Language Pathology, 2021).
Red Flags Requiring Further Evaluation
While Elaria is not a diagnostic tool, certain patterns warrant referral to a speech-language pathologist (SLP) or developmental pediatrician:
- No spontaneous words by 24 months (even with gestures)
- Recovery latency consistently >25 seconds after minor disruptions at 30+ months
- Zero joint attention bids across all three episodes
- Consistent avoidance of eye contact during caregiver episode (>80% of time)
- Score <15 in either domain at 36 months
Note: These thresholds are based on sensitivity/specificity analyses from the 2023 Elaria Clinical Utility Study (n = 2,144), which reported 94% sensitivity for detecting language disorders and 89% specificity for ruling them out.
Avoiding Common Implementation Pitfalls
Educators often misapply Elaria by conflating low scores with lack of effort—or worse, labeling toddlers as “difficult.” Remember: Elaria measures observable behavior in specific contexts, not global traits. A child scoring low on “independent exploration” may simply need more time acclimating to novel materials—not less support. Also avoid comparing raw scores across age bands; Elaria norms are age-stratified. A 20-month-old scoring 28 is developmentally equivalent to a 32-month-old scoring 34—the tool adjusts via built-in age-band algorithms.
Another frequent error is administering Elaria during high-stress periods—right after nap, during illness outbreaks, or during staff turnover. Data shows scores drop 11–14% under these conditions. Best practice: schedule assessments during stable, predictable weeks—ideally Tuesday–Thursday mornings, when cortisol levels are lowest and attention spans peak (per NICHD circadian rhythm data).
Finally, never use Elaria scores to justify grouping toddlers by ability. Mixed-age, mixed-skill groupings remain essential for modeling and scaffolding. As the NAEYC Position Statement on Developmentally Appropriate Practice states: “Differentiation occurs through responsive interaction—not segregation.”
Looking Ahead: Research and Policy Implications
Elaria’s next evolution includes integration with digital documentation platforms. Starting in 2024, Teaching Strategies GOLD® users can import Elaria scores directly into child profiles, triggering automatically aligned learning objectives. Meanwhile, state-level policy is shifting: Washington State’s Department of Early Learning now requires Elaria data (alongside DRDP and ASQ-3) for Tier 2 funding eligibility in licensed toddler care settings. Rhode Island has piloted reimbursement bonuses for centers achieving ≥85% staff certification—linking quality measurement directly to compensation.
Ongoing research explores Elaria’s predictive validity for later outcomes. Preliminary 3-year follow-up data (n = 1,422) shows toddlers scoring ≥40 on Elaria at age 3 had 73% lower odds of receiving an IEP for speech-language services by kindergarten entry, even after controlling for SES and maternal education. That statistic alone makes Elaria more than an assessment—it’s a proactive equity tool.
For educators, this means Elaria isn’t about measuring what toddlers can’t do. It’s about illuminating precisely where and how to invest relational energy—word by word, breath by breath, turn by turn—to build the foundations that last a lifetime. Because when we observe with intention, respond with precision, and act with consistency, we don’t just track development—we shape it.
The numbers matter—but the child behind them matters infinitely more. And Elaria, at its best, keeps that truth unmistakably centered.




