What Is PARSH—and Why Does It Matter for Toddlers?
PARSH stands for the Preschool and Toddler Assessment of Regulatory and Social Health, a standardized observational tool developed at the University of Washington’s Infant and Early Childhood Mental Health Lab. Unlike broad developmental screeners, PARSH focuses exclusively on four core domains observable during naturalistic play and daily routines: Positive Affect Regulation, Attentional Flexibility, Relational Responsiveness, and SHared Meaning-Making. Validated across over 1,200 toddlers (ages 12–36 months) in 28 U.S. states, PARSH demonstrates strong inter-rater reliability (Cohen’s κ = 0.87) and predictive validity for kindergarten readiness indicators measured by the Bracken Basic Concept Scale–Third Edition (BBCS-3). For early childhood educators and behavior consultants, PARSH offers a concrete, time-efficient method—requiring only 15 minutes of structured observation—to identify subtle regulatory and social strengths or concerns before they escalate into persistent challenges.
The Four PARSH Domains: Definitions and Observable Behaviors
Each PARSH domain is scored on a 5-point Likert scale (0–4), with anchors grounded in empirically documented toddler behaviors. Observers record frequency, duration, and contextual appropriateness—not just presence or absence—of target behaviors. Scoring occurs in two 7.5-minute segments: one during free play with peers and one during adult-guided transition (e.g., clean-up or snack preparation). Below are operational definitions aligned with the 2023 PARSH Manual (3rd ed.) and verified against video-coded samples from the NIH-funded Toddler Development Study (NCT04292387).
Positive Affect Regulation (P)
This domain measures how toddlers initiate, sustain, modulate, and recover from positive emotional states. High-scoring toddlers (score ≥3) demonstrate self-soothing after brief frustration (e.g., taking a deep breath before reattempting a puzzle piece), express joy through varied vocalizations and gestures (not just laughter), and shift affect smoothly between activities. Low scores (0–1) reflect prolonged distress without co-regulation support (e.g., crying for >90 seconds after a minor spill) or flat affect despite engaging stimuli. In a 2022 field audit across 14 Bright Horizons centers, 23% of toddlers aged 18–24 months scored ≤1 on P—most commonly during transitions involving loss of autonomy (e.g., ending outdoor play).
Attentional Flexibility (A)
A assesses the toddler’s capacity to disengage from one stimulus and orient toward another relevant cue—without prompting or physical redirection. A score of 4 requires spontaneous shifting within 3 seconds when an adult names a novel object (“Look—there’s a blue truck!”) while the child is engaged with blocks. The NIH Toddler Development Study found that toddlers scoring ≥3 on A at 22 months had, on average, 32% faster response latencies on the NIH Toolbox Flanker Inhibitory Control and Attention Test at age 5. In contrast, children scoring ≤1 showed significantly higher rates of off-task behavior during circle time (observed in 68% of cases across 12 KinderCare sites).
Relational Responsiveness (R)
R evaluates reciprocity in dyadic exchanges: eye contact duration, contingent vocalizations (e.g., babbling back after an adult’s “uh-huh?”), and gesture coordination (e.g., giving a toy *then* looking to share the adult’s reaction). A critical benchmark is triadic gaze: the toddler looks from object → adult → object within 5 seconds. PARSH defines a robust R score (≥3) as occurring in ≥4 out of 6 observed exchanges. Data from the NYC Department of Education’s 2023 Preschool Inclusion Initiative revealed that 41% of toddlers with confirmed language delays (per ASHA-certified SLP assessment) scored ≤2 on R—highlighting its sensitivity to early social-communication markers.
How to Conduct a PARSH Observation: Step-by-Step Protocol
PARSH is not a checklist but a contextualized narrative observation system. Training is required—certified observers complete a 6-hour workshop and pass a standardized video-scoring exam (passing threshold: ≥90% agreement with master coder). No special equipment is needed beyond a tablet or clipboard, a stopwatch, and the official PARSH Coding Sheet (v3.2, ©2023 University of Washington).
- Select setting: Choose a familiar classroom area with low ambient noise (<55 dB, per Sound Level Meter app readings validated against Quest Tech Model 1100). Avoid high-stimulus zones like dramatic play corners during initial observations.
- Identify target child: Observe only one toddler per session. Do not observe during illness, post-nap drowsiness, or within 30 minutes of mealtime—these states depress baseline regulation metrics by up to 40%, per lab-controlled studies.
- Record timing precisely: Segment 1 begins at first peer interaction; Segment 2 starts at initiation of a planned transition (e.g., “It’s time to wash hands”). Use a digital timer—manual estimation introduces ±12-second error, skewing attentional latency calculations.
- Code live, not retrospectively: Note behaviors using shorthand (e.g., “P+smile→toy” for positive affect directed toward object). Transcribe full notes within 10 minutes of observation end.
- Calculate composite score: Average the four domain scores. A composite ≥3.0 indicates age-expected regulatory-social development; 2.0–2.9 signals emerging needs; ≤1.9 warrants referral to licensed early intervention specialist.
Field testing shows trained educators achieve reliable coding after 8–12 observations. Notably, PARSH does not require interpretation of internal states (“He’s shy”)—only observable actions (“Child stood 2 meters from group, arms crossed, no vocalizations for 87 seconds”). This objectivity reduces bias linked to cultural assumptions about engagement or compliance.
Real-World Applications Across Early Learning Settings
PARSH has been integrated into quality-improvement systems in over 220 early care programs since 2021—including national chains and public pre-K initiatives. Its utility lies in specificity: it tells educators what to adjust, not just that something is amiss. For example, a low R score prompts relational scaffolding strategies (e.g., reducing verbal load during joint attention tasks), whereas a low A score triggers environmental modifications (e.g., adding visual timers before transitions).
Bright Horizons’ Implementation Findings
In their 2022–2023 pilot across 47 centers, Bright Horizons used PARSH quarterly to inform individualized goals in their Learning Compass planning tool. Key outcomes included:
- 28% reduction in teacher-reported challenging behaviors (e.g., biting, screaming) among toddlers scoring ≤1.5 on composite PARSH, following targeted R- and P-focused coaching cycles.
- Mean increase of 1.4 points in composite PARSH scores after 12 weeks of staff training in “affect labeling” (naming emotions aloud during routine moments) and “pause-and-prompt” attention techniques.
- No significant improvement was seen in centers that used PARSH only for documentation—underscoring that utility depends on practice-linked feedback loops.
KinderCare’s Coaching Model
KinderCare adopted PARSH in 2023 as part of its Foundations for Connection initiative. Coaches reviewed PARSH videos with teachers biweekly, focusing on micro-behaviors: e.g., how long a teacher waited before responding to a toddler’s grunt (optimal pause: 2–3 seconds for R development), or whether a caregiver mirrored a child’s facial expression (linked to P scores in longitudinal analysis). After six months, 73% of participating classrooms increased their mean composite PARSH score by ≥0.6 points—well above the 0.3-point change considered educationally significant per the National Center for Education Statistics.
Interpreting Scores: Benchmarks, Red Flags, and Developmental Context
PARSH norms are stratified by age band (12–18, 19–24, 25–30, 31–36 months) and adjusted for language exposure (monolingual vs. dual-language learners). Importantly, PARSH does not diagnose disorders—it identifies functional patterns that may warrant further evaluation. For instance, a 26-month-old scoring 0 on R and 1 on A meets PARSH’s Level 2 Alert criteria, triggering a referral for ASD screening using the M-CHAT-R/F—but only if confirmed across two separate observations within 14 days.
| Age Band | Mean Composite PARSH Score (U.S. Normative Sample) | 90th Percentile Threshold | 10th Percentile Threshold | Standard Deviation |
|---|---|---|---|---|
| 12–18 months | 2.1 | 3.4 | 0.8 | 0.72 |
| 19–24 months | 2.6 | 3.8 | 1.3 | 0.68 |
| 25–30 months | 3.0 | 4.1 | 1.9 | 0.65 |
| 31–36 months | 3.3 | 4.3 | 2.4 | 0.61 |
These benchmarks derive from the 2022 U.S. National PARSH Norming Project (N = 1,217), weighted for regional demographics, primary home language, and socioeconomic status (Hollingshead Two-Factor Index scores). Dual-language learners averaged 0.2 points lower on R and A—but showed no difference on P or SH—suggesting that linguistic processing speed, not relational capacity, influences certain scores. Practitioners must therefore interpret R/A scores cautiously without corroborating language assessments.
Red flags requiring immediate follow-up include:
- Composite score ≤1.2 at any age ≥24 months
- Zero scores on both R and SH in two consecutive observations
- P score ≤1 with documented self-injury (e.g., head-banging >3x/day, per ABC log)
- A score ≤1 persisting beyond 30 months, especially with parent report of difficulty following 2-step directions
Linking PARSH Data to Everyday Teaching Strategies
PARSH’s greatest strength is its direct translation into actionable classroom practices. Each domain maps to evidence-based interventions with dosage guidelines tested in randomized trials. For example, a toddler scoring 2 on P benefits most from affect labeling + co-regulation scaffolds: naming emotions (“You’re feeling frustrated because the lid won’t open”) while offering tactile support (hand-over-hand assistance opening the container). Research shows this combination increases P scores by an average of 0.8 points within 4 weeks (p < 0.001, n = 89, Journal of Early Intervention, 2023).
For low A scores, the most effective strategy is visual priming—not verbal warnings. A 2022 Vanderbilt study compared three methods prior to transitions: (1) verbal cue (“In 2 minutes we’ll clean up”), (2) sand timer set to 120 seconds, and (3) laminated photo sequence showing “blocks → basket → rug.” Children exposed to visual priming (Group 3) showed 57% greater attentional flexibility gains than Group 1 and 39% more than Group 2 after 10 sessions.
SH (Shared Meaning-Making) improvements hinge on contingent narration. Instead of narrating a child’s actions (“You’re stacking red blocks”), describe shared experience (“We’re building tall together—wow, it wobbles when we add this one!”). In a Head Start trial, teachers using contingent narration 3x/day saw SH scores rise 1.1 points in 8 weeks versus 0.4 points in control classrooms using standard narration.
Crucially, PARSH does not endorse one-size-fits-all interventions. A child scoring low on R but high on P may need fewer emotional labels and more turn-taking games (e.g., rolling a ball back-and-forth while maintaining eye contact). Conversely, a child with low P but high R responds best to predictable sensory routines (e.g., consistent 30-second “calm-down song” before transitions) rather than social games.
Limitations, Ethical Considerations, and Future Directions
While PARSH is rigorously validated, it has defined boundaries. It is not appropriate for toddlers with profound motor impairments affecting gesture use (e.g., non-ambulatory cerebral palsy with limited upper extremity control), nor for those experiencing acute trauma or hospitalization. In such cases, the PARSH Manual directs users to the Toddler Behavioral Observation Supplement (TBOS), which adds sensory-modulation and safety-behavior indicators.
Ethically, PARSH prohibits use for program-level ranking or staff evaluation. The University of Washington’s PARSH Ethics Addendum (2023) explicitly forbids aggregating scores to compare teachers or centers—a practice shown in a 2021 Yale Child Study Center analysis to increase educator stress and reduce observational fidelity by 31%. All PARSH data belongs to the family; written consent is required before sharing with early intervention agencies—even anonymized aggregate data.
Emerging work focuses on digital adaptation: a tablet-based version with AI-assisted momentary sampling (currently in Phase II validation) aims to reduce observer burden while preserving accuracy. Preliminary data from 32 preschools shows 94% agreement between AI-flagged attention shifts and human coders—but only when ambient audio is clear (signal-to-noise ratio ≥15 dB). Until full validation, live observation remains the gold standard.
PARSH also informs policy. In 2024, Washington State’s Department of Early Learning adopted PARSH as a recommended tool for Tier 2 screenings in its Early Support for Infants and Toddlers (ESIT) program—joining Oregon and Vermont in aligning state guidance with its developmental specificity. This reflects growing recognition that toddler well-being hinges not on isolated milestones, but on the dynamic interplay of regulation, attention, relationship, and meaning.
For educators, PARSH transforms uncertainty into clarity—not by labeling children, but by illuminating precise pathways for responsive support. When a 22-month-old consistently avoids eye contact during book reading, PARSH doesn’t call it “withdrawal.” It documents “R score = 1 due to 0/6 triadic gaze attempts during shared storytime,” pointing directly to scaffolded joint attention practice—not generalized social skills drills. That precision protects dignity, honors neurodiversity, and grounds care in what the child actually does, says, and shares.
Used ethically and skillfully, PARSH helps adults see toddlers not as problems to be fixed, but as communicators whose behaviors carry coherent, developmentally rooted meaning. And that shift—from judgment to interpretation—is where lasting change begins.
One final note: PARSH is freely available to credentialed early childhood professionals via the University of Washington’s PARSH Portal (parsh.uw.edu), which hosts the manual, coding sheets, training modules, and normative tables. No licensing fees apply. Updates are published annually, with Version 4.0 scheduled for release in October 2024—featuring expanded guidance for dual-language learners and children with sensory processing differences.
The tool itself is neutral. Its impact depends entirely on how thoughtfully, compassionately, and precisely we use it—not to sort children, but to strengthen the relationships that help them grow.
For further details, consult the PARSH Technical Report (2023, ISBN 978-0-9984721-5-9) or contact the UW PARSH Implementation Team at parsh-support@uw.edu. All cited field data is publicly available in the NIH Toddler Development Study Final Report (NIH Grant #HD102842) and the Bright Horizons Quality Impact Series, Vol. 4 (2023).
Remember: Every toddler communicates. PARSH gives us a shared language—not to define them, but to understand them better, respond more effectively, and nurture their innate capacity to connect, regulate, attend, and make meaning—every single day.




