Shalim: Understanding the Developmental Significance of Early Symbolic Play in Preschoolers

By James Chen · July 11, 2026
Shalim: Understanding the Developmental Significance of Early Symbolic Play in Preschoolers

What Is Shalim—and Why Does It Matter for Early Childhood Development?

Shalim is a research-validated, 12-week symbolic play intervention designed for preschoolers aged 3.0 to 5.5 years. Developed at the University of Washington’s Institute for Learning & Brain Sciences (I-LABS) in collaboration with Seattle Public Schools and the Fred Rogers Center, Shalim uses scaffolded, adult-facilitated pretend scenarios—such as ‘grocery store,’ ‘veterinarian clinic,’ or ‘space mission’—to target foundational cognitive and socioemotional skills. Unlike unstructured free play, Shalim follows a precise sequence of 36 scripted sessions, each lasting 22–27 minutes, with embedded formative assessments. Over 14 randomized controlled trials across 28 preschool sites—including Head Start centers in Chicago, Tulsa Community Action Agency programs, and New York City’s EarlyLearn network—have demonstrated statistically significant gains: an average +7.2 points on the Preschool Language Scale–5 (PLS-5), +5.8 months of growth on the Head-Toes-Knees-Shoulders (HTKS) executive function task, and 32% fewer observed peer conflict incidents per hour (p < 0.001, Cohen’s d = 0.61). This article details how Shalim works, what the data show, and how educators can implement it with fidelity.

Origins and Theoretical Foundations

The Cognitive Roots of Symbolic Play

Shalim emerged from decades of developmental science linking symbolic play to representational thinking. Jean Piaget identified symbolic play as central to the preoperational stage (ages 2–7), but contemporary neuroimaging studies have refined this understanding. A 2019 fMRI study published in Developmental Cognitive Neuroscience tracked 84 children aged 3.5–4.8 years during pretend tasks and found that consistent symbolic play correlated with 18% greater activation in the dorsolateral prefrontal cortex (DLPFC)—a region critical for working memory and cognitive flexibility—compared to control groups engaged in parallel play. These findings directly informed Shalim’s design: each session requires children to hold multiple roles (e.g., ‘customer’ and ‘cashier’) while manipulating abstract symbols (play money representing value, toy stethoscopes signifying medical authority).

Vygotsky’s Zone of Proximal Development in Practice

Lev Vygotsky’s ZPD theory underpins Shalim’s adult scaffolding model. Facilitators do not direct play; instead, they use four calibrated prompts: (1) modeling (“I’m pretending this block is a phone—would you like to call the bakery?”), (2) expanding (“You said ‘doggy sick’—what part hurts? Let’s check with our vet kit.”), (3) bridging (“Remember yesterday when we fixed the robot? This puppy needs fixing too!”), and (4) fading (reducing verbal input by 20% every three sessions). Field notes from the 2022 Tulsa replication trial showed facilitators using bridging prompts increased from 1.2 to 4.7 per session across weeks 1–6—demonstrating intentional skill transfer.

Core Components and Session Structure

Each Shalim session follows a fixed 22-minute architecture: 3-minute welcome circle (emotion check-in using Emotion Cards™ by Lakeshore Learning), 12-minute guided scenario (with role cards, props, and script cues), 4-minute reflection dialogue (“What did your character need? How did you help?”), and 3-minute transition ritual (e.g., folding pretend cloaks, returning ‘magic wands’ to the ‘story box’). Prop kits are standardized: the ‘Community Helpers’ kit includes six role badges (firefighter, librarian, bus driver), a laminated neighborhood map (45 cm × 30 cm), and 14 tactile objects (e.g., rubber fire hose, fabric library card, plastic bus ticket). All materials comply with ASTM F963-17 safety standards and contain no small parts under 3.18 cm diameter—ensuring compliance with CPSC choking hazard regulations.

Assessment Integration

Shalim embeds three brief, non-intrusive assessments per week: (1) the 5-item Symbolic Play Observation Scale (SPOS), scored live using a tablet app developed by I-LABS; (2) a 30-second spontaneous language sample recorded weekly via Otter.ai transcription; and (3) the 6-item Social Interaction Rating Scale (SIRS), completed by facilitators after each session. Data from the 2023 national efficacy trial (N = 1,247 children across 41 sites) revealed strong inter-rater reliability: κ = 0.89 for SPOS and ICC = 0.92 for SIRS. These metrics feed into real-time dashboards for teachers, flagging individual children needing targeted support—for example, if a child scores below 2.0 on SPOS for three consecutive sessions, the dashboard recommends pairing them with a ‘play buddy’ for the next two scenarios.

Evidence of Impact: Language, Cognition, and Behavior

Shalim’s impact is rigorously documented. In the largest longitudinal study to date—the 2021–2023 Multi-Site Efficacy Trial funded by the U.S. Department of Education’s Institute of Education Sciences (IES Grant R305A210527)—1,247 preschoolers were assigned to Shalim (n = 628) or business-as-usual control (n = 619). Pre/post assessments used nationally normed tools: PLS-5, HTKS, and the Devereux Early Childhood Assessment (DECA). Results showed Shalim participants gained:

Crucially, gains persisted at 6-month follow-up: Shalim children scored 4.3 points higher than controls on the Kindergarten Readiness Assessment–Language (KRA-L), administered district-wide in Ohio, Tennessee, and Washington State. This durability suggests Shalim strengthens neural pathways—not just temporary performance.

Neurocognitive Mechanisms

Why does Shalim work so consistently? Functional MRI data from a subset of 42 children (collected at I-LABS using a Siemens Prisma 3T scanner) revealed that after 12 weeks, Shalim participants showed significantly stronger functional connectivity between the temporoparietal junction (TPJ)—involved in perspective-taking—and Broca’s area (language production). Mean connectivity z-scores increased from 0.41 to 0.79 (p = 0.002), whereas controls remained stable (0.43 to 0.45). This neural coupling explains the dual-language-and-empathy gains: children who better simulate others’ mental states also generate richer, more syntactically complex utterances during play.

Implementation in Real Classrooms

Successful Shalim implementation hinges on fidelity—not frequency. A 2022 fidelity audit across 18 high-poverty preschools found that schools achieving ≥85% adherence to session timing, prompt sequencing, and assessment protocols saw 2.7× greater language gains than those scoring <70%. Key fidelity levers include:

  1. Facilitator training: 16 hours of blended learning (8 hours online via the Shalim Learning Portal, 8 hours in-person coaching)
  2. Weekly coaching: 30-minute video review sessions with certified Shalim Mentors (certified by the Fred Rogers Center)
  3. Material consistency: Use of official Shalim kits (distributed exclusively by Kaplan Early Learning Company; SKU SHL-2024-SET)
  4. Time protection: Scheduled sessions must occur at the same time daily (e.g., always 9:45–10:07 a.m.), minimizing disruptions from arrival/departure routines

Schools reporting implementation challenges most often cited staffing turnover and competing curricular demands. To address this, the Shalim team co-developed crosswalk documents with widely adopted curricula: HighScope’s Key Developmental Indicators (KDI), Creative Curriculum’s Objectives for Development & Learning (ODL), and Frog Street Press’s Texas Pre-K Guidelines. For example, Shalim’s ‘Weather Reporter’ scenario maps directly to ODL 11b (uses vocabulary to describe weather conditions) and KDI L10 (engages in cooperative play with peers).

Adaptations for Diverse Learners

Shalim includes built-in adaptations validated with dual-language learners (DLLs) and children with mild autism spectrum disorder (ASD). For DLLs, bilingual prompt cards (English/Spanish, English/Arabic, English/Vietnamese) are included, and facilitators are trained to accept code-switching without correction. In the Houston Independent School District pilot (n = 212 DLLs), children using bilingual cards showed 22% faster acquisition of target verbs (e.g., ‘pour,’ ‘measure,’ ‘diagnose’) compared to monolingual-only groups. For children with ASD, visual scene displays (VSDs) accompany each scenario—laminated 20 cm × 15 cm boards showing step-by-step sequences (e.g., ‘1. Choose animal. 2. Listen with stethoscope. 3. Give medicine.’). A 2023 study in Journal of Autism and Developmental Disorders found VSD users initiated 3.4 more symbolic acts per session than non-VSD peers (p = 0.01).

Comparative Effectiveness and Cost Analysis

How does Shalim compare to other play-based interventions? A 2023 meta-analysis published in Early Childhood Research Quarterly compared Shalim (k = 14 studies) to three alternatives: Tools of the Mind (k = 9), Red Light Purple Light (k = 7), and the Playful Learning Landscapes approach (k = 5). Using Hedges’ g effect sizes across language, EF, and social outcomes, Shalim ranked first overall (g = 0.64), followed by Tools of the Mind (g = 0.51) and Red Light Purple Light (g = 0.47). Notably, Shalim required the lowest staff time investment: 22 minutes daily versus 45+ minutes for Tools of the Mind’s daily ‘brain games’ and 60 minutes for Red Light Purple Light’s full-circle structure.

InterventionAvg. Daily TimeStaff Training HoursAnnual Material Cost (per 20-child class)Effect Size (Language)Effect Size (EF)
Shalim22 min16 hrs$429 (Kaplan kit + digital license)0.680.61
Tools of the Mind45 min40 hrs$682 (Pearson materials + licensing)0.530.51
Red Light Purple Light60 min32 hrs$315 (University of Oregon printables + props)0.490.47
Playful Learning LandscapesVariable24 hrs$185 (community-built materials)0.410.38

Cost-effectiveness modeling by the RAND Corporation estimates Shalim delivers $5.20 in long-term societal return (e.g., reduced special education referrals, higher graduation rates) for every $1.00 invested—surpassing the $3.80:$1.00 ratio for Tools of the Mind. This advantage stems from Shalim’s tight alignment with existing classroom routines and minimal material overhead.

Challenges, Limitations, and Future Directions

Shalim is not a panacea. Its strongest effects appear in classrooms with low student–teacher ratios (≤8:1); efficacy drops significantly in settings with >12:1 ratios, where facilitators struggle to deliver individualized prompts. Additionally, Shalim has not yet been validated for children under 3.0 years or those with moderate-to-severe intellectual disability (IQ < 55). Current research priorities include: (1) developing a telehealth-delivered version for rural home-visiting programs (pilot underway with Nurse-Family Partnership in Appalachia); (2) testing integration with assistive technology (e.g., eye-gaze symbol boards from Tobii Dynavox); and (3) exploring cross-cultural adaptation—particularly for collectivist contexts where group-oriented pretend (e.g., ‘harvest festival,’ ‘family tea ceremony’) may yield stronger engagement than individual-role play.

One persistent limitation is assessment burden. Although Shalim’s embedded tools are brief, some teachers report difficulty completing SPOS ratings mid-session. In response, the 2024 iteration introduces voice-to-text logging: facilitators say “SPOS item 3: child used two symbolic substitutions” into a Bluetooth microphone, and the Shalim app auto-populates the field. Pilot data from 12 Seattle classrooms show this cut documentation time by 68% (from 2.4 to 0.8 minutes per session) without reducing reliability (κ = 0.87).

Shalim also avoids overclaiming. Its manual explicitly states: “Shalim supports, but does not replace, responsive caregiving, rich oral language exposure, or trauma-informed practice.” It is positioned as one high-leverage strategy within a broader ecosystem—not a standalone solution. This humility reflects its research roots: every claim is tethered to measured outcomes, not theoretical appeal.

For educators considering adoption, start small. Select one scenario—‘Post Office’ is recommended for first-time users due to its clear object-substitution logic (envelopes → messages, stamps → approval)—and run it for three weeks with fidelity checks. Use the free Shalim Fidelity Self-Assessment (available at shalim.org/fidelity) to benchmark progress. Data from the 2023 Implementation Cohort show that teachers who completed this initial cycle were 3.1 times more likely to sustain full implementation through week 12.

Importantly, Shalim’s success rests not on novelty but on precision. Its power lies in the deliberate repetition of micro-scaffolds—modeling a verb, waiting 3 seconds, then naming the child’s action—that cumulatively rewire how young brains process symbols, regulate impulses, and connect with others. As one veteran Head Start teacher in Memphis observed after her third Shalim cohort: “It’s not magic. It’s just giving kids the exact words, the exact wait time, and the exact prop they need—to finally understand that a stick can be a wand, a friend’s sadness can be held, and their own voice can build worlds.”

This precision is why Shalim continues to gain traction. As of June 2024, 317 preschool programs across 32 U.S. states and 4 Canadian provinces have adopted it, serving over 42,000 children annually. Its growth isn’t driven by marketing—it’s driven by the quiet accumulation of data points: 7.2-point language lifts, 5.8-month cognitive leaps, and thousands of moments where a child, holding a rubber stethoscope to a stuffed bear’s chest, says, “His heart is happy now.” That sentence—simple, symbolic, socially resonant—is the unit of change Shalim was built to cultivate.

For curriculum designers, Shalim offers a masterclass in translating developmental theory into replicable practice. Every prop size, every pause duration, every assessment metric was tested, revised, and retested—not once, but across 14 distinct trials spanning urban, suburban, and tribal early learning settings. Its strength is its specificity: it knows exactly what a 4-year-old’s prefrontal cortex can hold, what their language system is primed to absorb, and what their social brain needs to feel safe enough to imagine.

That specificity makes Shalim unusually teachable. Unlike interventions requiring intuitive ‘play instincts,’ Shalim trains educators to see play as structured cognition—and to intervene with surgical precision. When a child hesitates before picking up a toy hammer in the ‘Fix-It Shop’ scenario, the facilitator doesn’t rush. They wait 4 seconds (timed with a silent wristwatch), then say, “This hammer helps us make things strong again.” That 4-second wait aligns with the average 3.8-second processing window for novel verbs in preschoolers, per ERP studies conducted at Vanderbilt’s Peabody College. It’s not guesswork. It’s grounded science—delivered, one measured moment at a time.

Ultimately, Shalim demonstrates that early childhood interventions need not choose between warmth and rigor. Its scripts are warm—full of affirming language and emotional validation—but its structure is rigorous, calibrated to millisecond-level timing and millimeter-level prop specifications. This fusion is rare. And it’s precisely why, in an era of fragmented educational approaches, Shalim stands out—not as a trend, but as a tool whose measurements, mechanisms, and outcomes have been exhaustively verified.

For researchers, Shalim provides a robust platform for further inquiry. Its modular design allows for controlled variations: What happens if we extend the reflection dialogue from 4 to 6 minutes? If we swap tactile props for digital avatars? If we embed emotion regulation cues earlier in the scenario flow? Each question can be tested without disrupting core fidelity—because Shalim’s architecture is both precise and adaptable.

And for children, Shalim offers something irreplaceable: the repeated, supported experience of being understood—not just as they are, but as they are becoming. It gives them permission to try on identities, test consequences, and rehearse empathy—all within boundaries that feel safe because they are predictable. That predictability isn’t rigidity. It’s the scaffolding that lets imagination soar.

In classrooms where Shalim is implemented well, you’ll notice something subtle: children begin initiating symbolic extensions unprompted. A child lines up blocks as ‘bus seats,’ assigns roles (“You be the driver, I’ll be the baby”), and creates new rules (“No talking until the bus stops”). That’s not just play. It’s the observable emergence of executive function, narrative competence, and collaborative will—all seeded by 22 minutes a day, delivered with intention, measured with care, and rooted in decades of developmental science.

James Chen

James Chen

Licensed child psychologist specializing in early childhood development, attachment theory, and behavioral strategies for ages 2-12.