Skip to content

Does Shadowing Help You Speak? What a Review of 44 Studies Says, and Where Popular Claims Go Further

Oct 8, 20261 min
TL;DRA 2025 review of 44 studies concludes that shadowing can improve the comprehensibility and fluency of your pronunciation. Most of those studies only tested controlled tasks such as reading aloud, and almost none compared shadowing with other techniques. Use it for rhythm, intonation, and smooth delivery. Building your own sentences and responding to people takes retelling, role play, and conversation.

🌏 中文版

In September I wrote a post on practicing spoken English for work and put shadowing first among the methods. My sources were two articles from English-learning platforms. Over the past two days I went back to the research while building my own speaking practice area, and found that I had claimed more than the evidence supports. This post adds what the research says, what shadowing is good for, and what it does not train.

What shadowing is and where it comes from

Shadowing means listening to audio and saying what you hear at nearly the same time. Whitworth and Rose at the University of Oxford define the standard form in their 2025 systematic review:

The technique involves listening to a short audio text, without a script, and repeating what is heard as simultaneously as possible.

So: no script, short audio, nearly in sync. In practice you trail the audio slightly, but you do not pause it.

On origins, the same review says shadowing was first used to train beginner interpreters to listen and speak at once. It cites Lambert 1992 for this; I could not read the original. In language learning, the first published accounts date from the 1990s, when the Japanese researcher Tamai used it to train listening in Japanese learners of English. The review also notes that most shadowing research so far treats it as listening training.

A separate line runs through Alexander Arguelles, an American polyglot. On his old website he calls shadowing "my technique", meaning his own specific procedure. I come back to it below.

What the research supports

The review searched six databases and included 44 studies. On close reading, 8 of them had not used real shadowing (they used delayed repetition of single words), and 2 were duplicates. The authors analyze 34, of which 26 directly examine whether pronunciation improved. The review is about pronunciation teaching, not speaking ability as a whole.

It concludes that shadowing can improve four things: comprehensibility (how easy listeners find you to understand), intelligibility (how much they actually understand), accentedness, and fluency. It also finds evidence for prosody such as intonation and rhythm, which the authors mark as more tentative. For individual sounds, the abstract says:

Research into the impact of shadowing on segmental pronunciation control was, however, inconclusive.

The review singles out two studies as high quality. One is Foote and McDonough's 2017 experiment. Sixteen advanced learners of English at a university in Montreal used iPods to shadow sitcom dialogues about one minute long. They did this for eight weeks, at least four times a week and at least ten minutes each time, and they recorded themselves and listened back. Besides a shadowing test, they told a story from pictures, and 22 native English speakers rated the recordings. Comprehensibility and fluency in the picture story improved significantly. Accentedness did not change. The study matters because it tested unscripted speech and because ordinary listeners could hear the gain.

What the research does not support

First, most studies tested controlled tasks. Of the 31 studies that measured pronunciation, 20 used only tasks such as reading aloud, recorded shadowing, or memorized presentations. The review warns:

improved performance in controlled tasks, such as read-aloud tests, may not directly translate into meaningful improvement in real-world speaking

Even the "extemporaneous" task in Foote and McDonough was the first 20 seconds of a picture story, not a two-way conversation.

Second, nobody knows how long the effect lasts. Only one study in the review ran a delayed post-test.

Third, almost no study compares shadowing with other techniques. The review found one, so the authors write:

Without this evidence base it becomes impossible to definitively elucidate the effectiveness of shadowing as a pedagogical technique compared to other pronunciation teaching approaches

Fourth, individual studies have design limits. Foote and McDonough recruited 22 people and 16 finished. There was no control group, the participants volunteered and were paid, and most were taking English classes at the same time. The authors themselves write that "due to participant self-selection and a lack of a control group, the results must be interpreted with caution", and that the activity should not replace classroom pronunciation teaching. A ten-week study from 2026 with a control group, of which I read only the abstract, reports that both groups improved, with "no statistical difference between the groups in any of the four measures".

The review cites Yo Hamada of Akita University many times. I read only the abstracts of his 2016 study and his 2019 overview. The first measures listening. The second suggests that beginners start with shadowing for listening and then move to shadowing for speaking.

These are Chinese-language articles I actually found. I found them by keyword search, so they do not represent Chinese-language writing as a whole. I also read careful ones: an article from NTE (in Mandarin) describes shadowing as 「主要用來練習聽力的方法」, a method mainly used to practice listening.

ClaimSource (in Mandarin)What the research or the original source says
「目前已被全球無數語言學習者驗證為提升口說流暢度最有效的主動訓練法之一」 (validated by countless learners worldwide as one of the most effective active training methods for speaking fluency)PREPThe review finds fluency gains, but no studies compare shadowing with other techniques, so nothing ranks it as "most effective"
Heading: 「為什麼它是流利度訓練最有效的方法之一?」 (why is it one of the most effective methods for fluency training?)TutorABCSame as above. This article's account of the origin (interpreter training) matches the review
「跟讀法的起源是由一位知名美國教授 Alexander Argüelles 所發明。」 (shadowing was invented by a well-known American professor, Alexander Argüelles)VoiceTubeThe review traces it to interpreter training and to listening instruction in Japan in the 1990s. Arguelles is describing his own procedure
「這種方法由日本語言學家井上敏明在1980年代提出」 (the method was proposed in the 1980s by the Japanese linguist 井上敏明)104 LearningThis name does not appear in the research I read, and a separate search found no support for it

Chinese-language articles are not the only ones that credit Arguelles. A British Council blog post also calls shadowing "the technique developed by linguist Alexander Argüelles".

I went looking for a literal sentence like "research proves shadowing is the most effective practice" and did not find one. What I found was "one of the most effective methods" and "validated as", attached to a source that sounds academic. The research says much less: shadowing helps, mainly with pronunciation and fluency, and it cannot be ranked against other methods.

Who puts it first

The review's count of study locations is one clue: 20 studies in Asia, including 10 in Japan and 6 in Taiwan. The authors also write that "shadowing is currently popular in Asia".

I had earlier gone through more than a dozen recommendation lists for speaking practice. Seven mention shadowing, and most of those are in Chinese. On the other side, I searched the full text of the Reddit language-learning board's FAQ, its guide page, and the five chapters of the online guide written by one of its moderators. The word shadowing does not appear once. On speaking, that guide says "One of the best language exercises you can do is conversation practice with a native speaker", and it recommends listen-and-repeat audio courses for people about to travel. Paul Nation's free book for learners lists twenty activities and shadowing is not among them. It only mentions, in the pronunciation section, that repeatedly imitating movie clips can help.

This does not mean English-speaking learners avoid shadowing. Arguelles is American, and search results show Reddit threads about it, which I did not read one by one. All I can say is that these English-language starter guides leave it out, while the Chinese lists I read often put it near the top.

The originator's version, and "repeat after a video"

For Arguelles, shadowing is the first step of a full self-study program. He defines it as echoing "a recording of foreign language audio that accompanies a manual of bilingual texts": the audio for a textbook with the original on the left page and a translation on the right. He asks for three things:

  1. Walk outdoors as swiftly as possible.
  2. Maintain perfectly upright posture.
  3. Articulate thoroughly in a loud, clear voice.

He suggests about 15 minutes per session, several times a day. You shadow without the book first, then read the original and the translation in stages. In a study roadmap he endorsed, shadowing sits in Phase 1, "Shadowing Your First Introductory Manual". The last phase is extensive reading, and its heading says "The Goal".

His shadowing is groundwork for learning a language from zero until you can read books in it, built on beginner textbooks such as Assimil. The version that circulates online usually means picking a video you like and speaking along, with speaking as the goal. The two share a name and differ in materials, procedure, and purpose. I have not read any research on whether his procedure works.

Nearby methods: the Echo Method and listen-and-repeat

The Echo Method (in Mandarin) from Professor Karen Chung of National Taiwan University often appears next to shadowing. She has you choose a one-to-three-minute recording with a transcript, listen until it is familiar, and read until you understand it. Then you play only four or five words, pause, and:

停一下!還不要機械式地脫口而出地「跟著念」!

That is: stop, and do not blurt out a mechanical repetition yet. You listen to the sound still in your head, then say it the way you heard it, for ten minutes a day. In her article she contrasts this with traditional repeat-after-me practice and never mentions shadowing. Her basis is years of teaching, and she cites no experiment. My September post originally said only VoiceTube had proposed the Echo Method. That was wrong, and this is the source. I have corrected that sentence.

ShadowingEcho MethodListen-and-repeat
When you speakWhile the audio is still playing, nearly in syncAfter a pause, once you have heard the echo in your headAfter the sentence ends
Length per turnA continuous passageFour or five wordsA sentence or phrase
Where attention goesAccording to the review, toward the sounds, with little left for meaningDetails of sound: pronunciation, stress, intonationAccording to the review, split between sound and meaning
Basis I readOne systematic review covering dozens of studiesThe originator's teaching experienceThe review calls it a well-studied traditional technique; I did not read further

The review separates shadowing from listen-and-repeat and excludes the 8 studies that mixed them up. Many people who say they shadow are playing a sentence, pausing, and repeating it. In the research, that belongs in the right-hand column.

How to use it for speaking practice

up (in Mandarin), a learning guide with 67,000 stars on GitHub, has this line in its speaking chapter:

跟读负责观察和模仿,复述、追问与协作才负责生成

In English: shadowing handles observation and imitation, while retelling, follow-up questions, and collaboration handle generation. The reason is that during shadowing the audio decides the content, the word order, and the next sentence. You are not building sentences. The guide suggests taking one piece of material through five levels: listen and mark, read aloud, shadow with a delay, close the text and retell it, then answer questions the material did not cover. That chapter cites research on pronunciation teaching and none on shadowing, so this is the guide author's judgment. It does point the same way as the gap the review identifies.

Nation's framework helps with proportions. He argues that language learning needs four strands: meaning-focused input, meaning-focused output, deliberate learning, and fluency practice. As I read it, shadowing conveys no message of your own, so it sits closest to deliberate learning; Nation's book does not place shadowing in any strand. For speaking he lists memorized sentences or dialogues, role play, prepared talks, and giving the same talk to different listeners in 4, 3, and 2 minutes. One caveat: on giving the four strands equal time, he admits that "There is no research evidence to support this equal division of the teaching and learning time". It is a common-sense judgment.

The British Council's advice on speaking has just four items: use the language, find people to talk to, record yourself and listen back, and work on listening.

Putting these together, this is how I would arrange things:

  1. Use shadowing for rhythm, intonation, and smooth delivery. Pick audio about one minute long and stay with it for a week, at least ten minutes a session and at least four sessions a week, and record yourself. That dose follows Foote and McDonough's experiment. Choosing audio you can understand is my own addition.
  2. If you cannot keep up, start with the Echo Method or sentence-by-sentence repetition. That is fine practice. Just know it is not the shadowing the studies tested.
  3. After each session, cover the text and say it in your own words. If you cannot, shorten the material. Do not go back and add more rounds of shadowing.
  4. Schedule separate time for unscripted speaking. Role play, answering questions you did not prepare for, and talking with people are the parts shadowing cannot stand in for.

What I decided for my practice area

Before I read these sources, my practice area (zh-TW only) had one kind of practice: read the Chinese, say the English, then compare with a reference version. In Nation's terms that covers deliberate learning and nothing else. I have since added these:

  • Recording comparison. I record myself and compare with the reference reading, then pick one thing to fix and say the sentence again. Fixing only a few problems per round is the up guide's advice. The list of options is my own.
  • Dialogue role play and interview questions. I act the sentences out inside a dialogue, or hear an interviewer's question and answer it myself. A dialogue can be acted again right away, which follows Nation's advice on role play.
  • Fluency practice. I speak on the same topic three times, for 2 minutes, then 1 minute 30 seconds, then 1 minute. This is a shortened version of the 4/3/2 activity in Nation's book. In the original you speak to three different listeners for 4, 3 and 2 minutes. Here I am alone.
  • Free talk. There is a prompt and no reference sentence. Afterwards I can download the recording and keep it for later comparison.
  • Shadowing and echo. I hear a sentence, wait two seconds, then speak, or I speak along as soon as the audio starts. The speed is adjustable.

Shadowing and echo sit in the "practise sentences" group and not in the "speaking practice" group, for the reason given in the sections above: they train sound, and I am not building sentences. The feature also differs from its sources in two ways. The voice is the browser's synthetic speech, whose rhythm differs from a real speaker's. Echo mode plays a whole sentence at a time, while Chung's original method plays only four or five words at a time and uses recordings of real speakers.

Two things code cannot supply. The other side of the dialogue is a written script, so it never asks a follow-up question and never fails to understand me. There is no human listener, so I cannot tell whether my meaning got across. These features are also new. I have not used them for long, and none has been tested for effect.

Limits of this post

  • The review covers only studies published in English, and its authors say it may miss a good deal of Japanese research. It is about pronunciation, not all of speaking.
  • I did not read the original works by Lambert, Tamai, or Kadota; they come to me through the review. I read only the abstracts of Hamada's two papers and the 2026 study. I did not watch Arguelles's demonstration video, and I could not get a transcript of Karen Chung's TEDx talk.
  • I read the Reddit FAQ in archived snapshots from 2022 and 2025. The online guide has five chapters, and I could not find the later ones it refers to.
  • "Research has not shown that shadowing helps conversation" is not the same as "shadowing does not help conversation". So far nobody has tested it properly.
  • The examples of popular claims came from searches. They are few and cannot stand for the whole Chinese-language learning community.

References