A Better Way of Shadowing: Break It Up First, Then Speed It Up

September 15, 2026

infographic on how to shadow english speech

I frequently recommend shadowing to my students as one of the most effective ways to attain speech fluency and a natural rhythmic flow in English. I detailed the general idea in a previous post, which is basically: listening to a native speaker and repeating what they say at the same time they’re saying it.

But if you’ve ever tried shadowing at full conversational speed, you might recognize this frustrating feeling: the speaker starts talking and you try to follow, but within three seconds, you’re already behind.

You miss a word, then another. You try to catch up, but now you’re listening to the next phrase while still trying to pronounce the previous one. And before you know it, the speaker has finished the clip… while you’re somewhere about three seconds behind.

So you try again, but get the same results. I usually tell students to keep going no matter what, even if you fall behind and have to keep catching up.

But another approach is to break the speech into pieces first, then gradually build the speed back up.

Why Full-Speed Shadowing Can Feel Impossible

Shadowing is different from ordinary repetition. When you repeat a sentence after someone finishes speaking, you have time to listen, process, and then produce the words.

But shadowing removes much of that time. With shadowing, you’re trying to listen and speak nearly simultaneously. This means that your brain has to process the incoming speech while your mouth is already producing the previous words.

At the same time, you’re trying to reproduce things like:

  • pronunciation
  • rhythm
  • stress
  • intonation
  • linking
  • reductions
  • pauses
  • speaking rate

That’s a lot to coordinate! And if you start with a long clip at full speed, your mouth may not yet be able to keep up with it all.

The result is often a kind of linguistic improvisation: you don’t fully pronounce the phrase, you skip sounds, and fall behind. And then you spend the rest of the clip trying to catch up!

Repeating that experience over and over isn’t necessarily the most efficient way to build accurate, automatic speech.

The Better Sequence: Slow → Segment → Combine → Accelerate

Instead of demanding full-speed performance immediately, give yourself a progression. Think of it like building a physical movement.

First, learn the pieces. Then combine them. Then make the movement faster. For shadowing, that means:

1. Segment the clip.

2. Practice each segment comfortably.

3. Combine the segments.

4. Gradually return to normal speed.

The goal isn’t to practice slowly forever. The slower stage is simply where you establish the pronunciation and timing before asking your speech muscles to perform at conversational speed.

Start With a Short Clip

Don’t begin with a two-minute monologue. Choose something short. Make it around 15–20 seconds. Good sources include:

  • podcasts
  • interviews
  • TV shows
  • YouTube videos
  • presentations
  • conversations

Ideally, choose a speaker whose voice is reasonably clear and a clip that contains language you might actually use. And don’t choose something that’s so difficult that you understand almost nothing.

You want a challenge, but also enough comprehension so that you can focus on how the person speaks, not just what they’re saying.

Step 1: Break the Clip Into Natural Chunks

Now divide the clip into three or four short phrases. Don’t necessarily divide it according to grammar or individual sentences.

Instead, look for natural speech chunks, meaning places where the speaker might naturally pause or where the phrase forms a meaningful unit. For example:

“I honestly didn’t think”
/
“it was going to rain today”
/
“but here we are.”

Now you have three manageable pieces. You’re no longer trying to control a 15-second stream of speech, but working with one short movement at a time.

Step 2: Shadow the First Chunk

Take your first chunk. Try to slow the audio down slightly if you can. You don’t need to make it ridiculously slow – the goal is simply to give yourself enough time to notice what’s happening.

Listen carefully to:

  • Which syllables are stressed?
  • Where does the speaker reduce sounds?
  • Which words connect?
  • Where does the pitch rise or fall?
  • How quickly does the speaker move between words?

Then shadow the chunk three times. Don’t worry about speed yet – prioritize synchronization instead. Try to make your speech movements follow the speaker’s timing rather than simply producing all the correct words.

Step 3: Add the Next Chunk

Now move to your second chunk. Practice it separately in exactly the same way, then do something important:

Connect chunks one and two.

Now you’re no longer practicing isolated pieces, but practicing the transition between them. Remember, in natural speech, words don’t exist as neatly separated units. They connect through linking, reductions, assimilation, and rhythm. Try to reproduce that connected movement.

Step 4: Add Speed Gradually

Once the individual chunks feel comfortable, start increasing the speed. For example:

Round 1: slightly slower than normal

Round 2: normal conversational speed

Round 3: full-speed shadowing

If you can perform a phrase accurately at a slightly slower pace, that’s a foundation you can accelerate. If you can’t perform it accurately at a slower pace, simply repeating it at full speed won’t magically solve the problem.

Your Four-Step Shadowing Progression

Here’s a simple routine you can use with almost any short clip.

Step 1: Build the Pieces

Divide your clip into three or four chunks. Shadow each chunk separately. If necessary, slow the recording down. Focus on:

pronunciation + rhythm + timing

Don’t worry about sounding fast.


Step 2: Connect the Pieces

Practice chunks one and two together. Then add the next chunk. Work at normal conversational speed where possible. Your goal is to make the transitions between chunks smooth.


Step 3: Rebuild the Whole Clip

Now shadow the entire clip at normal speed. Don’t stop after every mistake, but try to maintain the overall rhythm and keep moving with the speaker. At this stage, you’re practicing the complete motor sequence rather than individual pieces.


Step 4: Test Yourself Cold

Now try playing the clip at full conversational speed. Don’t slow it down, stop it, or look at the transcript. Just shadow it. Can your mouth now keep up without consciously constructing every word?

If yes, you’ve built a useful progression from controlled practice to automatic performance.

If not, go back and ask: Which chunk caused the breakdown? Maybe the problem was one difficult consonant cluster, or an unfamiliar reduction, or something else. Maybe you just need to practice it more.

Don’t Confuse Speed With Accuracy

One shadowing mistake is believing that faster automatically means better. Sadly, it doesn’t.

If you’re racing to keep up with the speaker but your pronunciation is falling apart, you’re practicing a distorted version of the target speech. Speed should be the final layer, not the starting point.

First establish the movement, then establish the timing, then increase the speed. It’s similar to learning a physical skill: you generally want to know what the movement should feel like before asking your body to perform it at maximum speed.

What Should You Actually Listen For?

When you work intensively on one particular passage, shadowing isn’t just about saying the same words. For accent training, pay particular attention to the features that make connected American English sound different from carefully pronounced textbook English.

Listen for:

Stress – Which words stand out?

Rhythm – Which syllables are strong and which are compressed?

Reductions – Which unstressed words become shorter or weaker?

Linking – Where do words run together?

Intonation – How does the speaker’s pitch move?

Timing – How quickly does the speaker move through each phrase?

These are the features that often make a learner’s speech sound more natural, regardless of whether your individual consonants and vowels are already accurate or not.

One Clip, Gradual Progress

Instead of shadowing a dozen different clips every day, working deeply with one short clip can sometimes be more productive than constantly jumping between new material.

Take 15–20 seconds, break it into pieces, then master the pieces, and connect them. Increase the speed and test yourself.

Most of all, keep in mind that your goal isn’t perfect imitation on your first attempt. Your goal is to gradually reduce the distance between what you hear and what you produce.

At first, you want to: Listen → process → produce

Then eventually: Listen → produce

The more automatic that connection becomes, the less mental effort you need to devote to “sounding fluent” and easier it will be achieve effortless fluency.

picture of Steven Nelson

Author: Steven D. Nelson

Steven is an American English accent coach, presenter, trainer, and the owner of AccentFirst. Steven's mission is to make accent modification accessible to all by providing inclusive, empowering, and effective training that helps people speak with greater confidence and clarity.

Leave a Comment

Update cookies preferences