What shadowing is
You listen to a recording of natural speech and say it back — either simultaneously, a beat behind the speaker, or immediately after each sentence. That is the whole technique.
What makes it valuable is that it is the only common exercise that trains production and perception at the same time. You cannot reproduce a rhythm you have not heard, so shadowing forces your ear to work while your mouth does.
The way most people do it, and why it does nothing
Most people shadow the words. The recording plays, they scramble to keep up, they get most of the words out, and they feel they have done something demanding — which they have, but not the demanding thing that helps.
The words are the part you already have. A transcript would give you the words. What the recording has, and a transcript does not, is the music — where the speaker leans, where they speed up, where they let a syllable disappear, where the pitch rises and falls.
That is the part that cannot be learned from a book, and it is the part shadowing exists to teach.
What to actually copy
Three things, and they are worth more than the sounds.
1. Sentence stress — which words get leaned on
English does not give every word equal weight. It leans on the words that carry the meaning and swallows the rest. I would have told you if I'd known.
Speakers of syllable-timed languages tend to give every word an even beat, which is what makes speech sound flat and effortful even when every individual sound is correct. Copying the stress pattern is the fastest fix for that, and it requires you to learn no new sounds at all.
2. Reduction — the words that almost disappear
Listen for what happens to the small words. And becomes n. To becomes tuh. Have becomes 've. Going to becomes gonna, want to becomes wanna, what do you becomes something like whaddaya.
This matters in both directions. Producing reductions makes you sound natural; hearing them is why you can suddenly understand fast speech. A great deal of "they speak too fast" is really "I am listening for words that are not being fully said."
3. Intonation — the melody that carries meaning
Falling at the end of a statement. Rising for a genuine question. Rising then falling for a list. Ending every sentence on a rise — a common habit in learners who are unsure — makes confident content sound tentative, and it is audible.
How to do it properly
- Choose material that is too easy. This is the counter-intuitive one. If you are working to understand it, you cannot attend to how it sounds. Pick something comfortably below your level — you want your whole attention on the delivery.
- Take a short chunk. Fifteen to thirty seconds. You will repeat it many times, and it must be short enough that you stop thinking about the words.
- Listen three times without speaking. Where does the speaker lean? Where do they pause? Where does the pitch rise? Do not skip this — it is the part that makes it shadowing rather than reading aloud.
- Now shadow it, exaggerated. Over-do the stress and the melody. Feeling ridiculous is a sign you are actually changing something; sounding sensible usually means you are producing your own prosody with someone else's words.
- Record yourself and compare. Not the words — the shape. Play the original, then yours. Where do the stresses land in each? This is the step that turns practice into feedback.
- Ten minutes a day. Consistently. This is motor learning, and it consolidates with sleep, not with long sessions.
How to tell whether it is working
This is the honest weakness of shadowing: it has no built-in feedback. You can do it for months, diligently, and have no idea whether anything has changed — and because your own ear is the least reliable judge of your own speech, your sense of improvement is not evidence.
So build the feedback in:
- Record yourself shadowing on day 1, and keep it. Compare against the same passage a month later. This is the cheapest evidence available, and almost nobody does it.
- Check the stress, specifically. Play the original and your version, and mark which syllable is loudest in each phrase. That is a comparison you can actually make, unlike "do I sound better?"
- Have a fluent speaker tell you which words they had to work to understand. A list of words, not a general impression.
- Get a per-sound score. Speech scoring can align what you said to the expected pronunciation and score each phoneme and its stress — which is exactly the diagnosis your own ear cannot perform. (Aflo does this; those are our own measurements, not official test criteria.)
Without one of these, shadowing is an act of faith. With one of them, it is the most efficient pronunciation practice there is.
Why it pays on the test
Because prosody is explicitly scored, and it is where the fastest gains are.
- TOEFL's speaking blueprint names intelligible prosody in the skill statement itself.
- Duolingo's pronunciation criterion lists word stress, sentence stress and intonation alongside individual sounds.
- IELTS's pronunciation descriptors are built around whether a listener understands you easily — which stress and intonation affect more than individual vowels do.
And none of them mention having a particular accent. Shadowing is not accent reduction, and it should not be used as such — you are copying the rhythm of English, which is shared across its accents, not the vowels of one speaker in one city.
The gains also transfer beyond the test. Sounding natural and being easy to listen to is what makes people relax when they talk to you, and that is worth considerably more than a band.
Frequently asked questions
How do I shadow English properly?
Copy the music, not the words. Choose material that is comfortably below your level so your attention is free for the delivery, take a chunk of fifteen to thirty seconds, and listen three times without speaking — noting where the speaker leans, pauses and lets their pitch rise. Then shadow it with exaggerated stress and melody, and record yourself to compare the shape against the original. Ten minutes a day beats a long weekly session.
Why is my shadowing not improving my pronunciation?
Most likely because you are shadowing the words rather than the prosody. If you are racing to keep up, all your attention is on retrieving the right words and none is left for how they sound — so you are practising your existing pronunciation, faster and under pressure. The words are the part you already have; the rhythm, stress and intonation are what the recording has that a transcript does not.
What material should I use for shadowing?
Something unscripted and conversational — interviews, podcasts, natural dialogue — and easy enough that you are not working to understand it. Scripted news is read aloud, and read-aloud prosody is not conversational prosody; audiobooks are performed. If you want to sound like someone talking, shadow someone talking.
How do I know if shadowing is working?
Build in feedback, because shadowing has none of its own and your own ear is the least reliable judge of your own speech. Record yourself on day one and keep it for comparison a month later; play the original and your version side by side and mark which syllable is loudest in each phrase; ask a fluent speaker which specific words they had to work to understand; or use a tool that scores each sound and its stress. Without one of these, you are practising on faith.
Does shadowing help with IELTS or TOEFL?
Yes, because prosody is explicitly scored. TOEFL's speaking blueprint names intelligible prosody in the skill statement, Duolingo's pronunciation criterion lists word stress, sentence stress and intonation, and IELTS's descriptors are built around whether a listener understands you easily — which stress affects more than individual vowels do. None of them reward a particular accent, so shadow the rhythm of English rather than trying to copy one speaker's vowels.