Video modeling as a bridge to social initiations for autistic learners
When a child stands at the edge of the playground at recess, watching peers laugh and chase each other but unsure how to step in, the gap between wanting to connect and knowing how to is a familiar one for speech-language pathologists. Many autistic learners understand far more language than they can use in the moment, and social initiations—greetings, questions, comments, requests to join—often sit right in that gap. Video modeling has become one of the most reliable, evidence-backed tools for closing it, and it travels well from clinic rooms in Brisbane to school yards in Perth.
Practitioners across Australia are finding that showing a learner a short clip of the exact behaviour we want to see, then prompting them to try it, can sidestep the working-memory load of long verbal explanations. The format is concrete, repeatable, and easy to share with families who fund therapy through the National Disability Insurance Scheme or access it through school-based teams. For SLPs looking for a low-prep strategy with high payoff, video modeling deserves a permanent spot in the clinical toolbox.
Understanding video modeling for social initiations
Video modeling is straightforward: record a short demonstration of a target skill, then watch it with the learner and provide an opportunity to practise. The demonstration can feature a peer, an adult, or—particularly powerfully—the learner themselves in a previous successful attempt. Most clips run between 30 seconds and three minutes, long enough to show the full sequence of the skill and short enough to hold attention.
Three formats appear most often in the research and in clinical settings. In adult-modeling videos, the clinician or a familiar adult performs the skill while narrating or thinking aloud. Peer-modeling uses a typically developing peer, often a sibling or classmate, to perform the initiation. Self-modeling edits together footage of the target learner performing the skill successfully, removing any mistakes and adding a brief voiceover if needed. Each format has different strengths, and many Australian clinicians rotate through them depending on the goal and the learner's profile.
The beauty of the approach is its flexibility. A recording can be filmed on a phone during a natural interaction at a kindergarten in Adelaide, or it can be a polished production made in a school speech room. What matters is that the target behaviour is clear, the model is salient, and the clip is shown often enough to become familiar.
The evidence base and Australian context
Social initiations are the small verbal and gestural moves that start and sustain an interaction: saying hi, asking a question about someone's game, commenting on a peer's drawing, or requesting to join a group. Research consistently shows that autistic learners initiate less often than their neurotypical peers, and that the quality of those initiations affects friendships, classroom participation, and long-term wellbeing.
Traditional verbal instruction often falls short here because social initiations are deeply contextual. Knowing when to say "Can I play?" depends on the activity, the people involved, and the unspoken rules of the moment. Video modeling addresses this by giving the learner a concrete example to imitate, complete with facial expressions, tone of voice, and body language. The clip does the contextual heavy lifting that a worksheet cannot.
The evidence base for video modeling in autism has grown steadily over the past two decades. Meta-analyses have repeatedly found medium to large effect sizes for teaching social communication skills, including initiations, to autistic learners across age groups. Australian research has contributed to this picture, with studies from universities in Melbourne and Sydney showing gains in conversational turn-taking, joint attention, and peer-directed speech after video modeling interventions. Maintenance and generalisation are where the approach often shines, because learners practise with the actual prompt—watching the clip just before the opportunity—and tend to retain the skill longer than with some other behavioural interventions.
Practical setup and filming tips
A successful video modeling clip is rarely accidental. It helps to write a quick script first, even if it is just a few bullet points: who will appear, what they will say, what the setting looks like, and how long the clip will run. Filming in the actual environment where the skill will be used—say, the sandpit at a local early learning centre or the lunch-order line at school—adds a layer of generalisation that generic studio footage cannot match.
Sound quality matters more than video quality. Viewers need to hear the model's words clearly, so choose a quiet spot or use a lapel microphone. Lighting should be even, especially if the learner will be watching the clip multiple times, and the model should face the camera with a neutral or positive expression. Captions or on-screen text can support learners who process written language well, and they help when the clip is shared with a parent who is deaf or hard of hearing.
Editing can be as simple as trimming the start and end in the native phone app. More elaborate edits, such as adding a freeze-frame on the key moment of the initiation or inserting a thought bubble, are useful for older primary and high-school students who benefit from explicit metacognitive cues. Keep the file in a shared folder so the whole team—classroom teacher, integration aide, parent, and SLP—can access it. Funding pathways also support ongoing use. Under the National Disability Insurance Scheme, video modeling can be delivered as part of a capacity-building therapy plan, and the clips themselves can be created during existing sessions, meaning no extra billable hours. Schools in Victoria, Queensland, and Western Australia have included video modeling in their reasonable adjustments documentation for students on the autism spectrum, which makes it easier for SLPs to share resources with classroom teams.
Matching formats to learners
Not every video modeling approach fits every child, and trialling two or three formats is often worthwhile. The table below summarises the main options and the contexts in which they tend to work best. It is a starting point rather than a rulebook, because the most important variable is always the individual learner in front of you.
| Format | Best for | Practical considerations |
|---|---|---|
| Adult modeling | Younger learners, new skills, when no peer model is available | Quick to film, easy to script, but may generalise less to peer contexts |
| Peer modeling | Learners motivated by same-age friends, classroom-based goals | Requires a willing peer, consent, and a familiar filming location |
| Self-modeling | Learners who respond well to seeing themselves succeed, confidence-building | Needs pre-existing footage or staged success, more editing time |
| Mixed-modeling | Learners who benefit from multiple perspectives | Longer to produce, but flexible across settings |
When introducing video modeling for the first time, start with one highly motivating target and one familiar setting. "Asking to join a game at lunchtime" might be too broad; "Walking up to a peer and saying 'Can I play with you?' during recess" is more specific. Once the learner masters that, add a second context, then a third, and watch for spontaneous use outside the filmed scenarios. Generalising from one setting to another still requires planning, so a learner who masters the speech room version will usually need additional practice in the playground, and pairing the video with peer-mediated support can help bridge that gap.
Materials and quick wins for busy clinicians
A handful of resources makes video modeling sustainable week to week. The essentials below cover most clinicians, regardless of caseload.
- A smartphone or tablet with a decent camera, mounted on a small tripod for stability
- A quiet location with even lighting, such as a corner of the therapy room or an outdoor bench
- A simple editing app that allows trimming, captions, and slow-motion playback
- A shared cloud folder or school drive where clips can be stored and accessed by the team
Equally important are the small habits that keep the approach feeling fresh rather than rote. Rotating the models, updating the filming locations, and involving the learner in choosing the next target all help maintain engagement.
- Let the learner help pick the next peer or adult to film
- Change the setting every few weeks to encourage generalisation
- Celebrate small wins with a quick photo or a thumbs-up at the end of the clip
- Review older clips together to reinforce how far the learner has come
If video modeling is new to your caseload, a low-stakes first project is the easiest way to build confidence. Choose one learner, one social initiation, and one setting. Film a 60-second clip during a natural interaction, trim it on your phone, and watch it with the learner before their next opportunity to practise. Track what happens across three sessions and adjust from there.
For those wanting to dig deeper into the approach, Rachel Jones shares practical examples, downloadable templates, and session ideas drawn from real Australian contexts. The field of video modeling keeps evolving, particularly with the rise of short-form video and new editing tools, but the core principle remains the same: show the learner what success looks like, give them a chance to try, and repeat until the skill becomes their own.
Ready to give it a go? Pick one learner, grab your phone, and film that first clip this week. The playground at lunch, the library corner, the therapy room mat—anywhere a real interaction happens is the perfect place to start.
On the Blog
Popular posts and resources from the Let's Talk Speech Therapy archive.
Top 10 Books that Encourage Language Development for Babies and Toddlers
A curated list of books that support early language development, with tips for SLPs and parents on how to use them during reading time.
Supervising a SLP Student Intern – Getting Started
Practical guidance for SLPs new to clinical supervision, covering preparation, expectations, and mentorship strategies for working with student interns.
How (and why) I Teach Phone Numbers in Speech Therapy
Inspired by a real incident, this post explains how to use backwards chaining and exit-ticket questions to teach personal safety information.
7 Tips for Making the Most of ASHA '16
Seven practical tips for navigating the ASHA convention — from session planning to networking — based on firsthand experience.
Pete the Cat Phoneme Segmentation Screener
A free printable phoneme segmentation screener featuring Pete the Cat, shared as a resource for SLPs working on phonological awareness. Published September 16, 2012.
Speaking & Professional Development
Information about professional development opportunities and speaking engagements for schools, clinics, and SLP groups.