English Shadowing open the app →

Understanding Fast English: It Isn't Actually Fast

It feels like a wall of sound. You are still checking the first three words while the speaker has already moved into the next sentence. The usual explanation is “native speakers talk too fast.” That explanation is mostly wrong.

In ordinary conversation, people produce about four to six syllables per second. That is quick, but it is not extreme. You can produce six syllables in a second without strain. The feeling of speed comes from something else: English gets compressed.

Native speakers shrink function words. They link consonants to vowels. They drop whole syllables from common phrases. “What are you going to do?” becomes “Whaddaya gonna do?” “I would have helped” becomes “I woulda helped.” The syllables have not sped up much. The spaces and clear edges between words have disappeared. That is the thing you hear as speed.

The connected speech page is the anatomy lesson. This page is the training program.

1. Learn the reduced forms deliberately

You cannot hear a form you do not expect. If you have never seen “woulda” in print, it sounds like three random sounds. Once you know it means “would have,” the same audio becomes easier. The change happens before you listen again. You gave your brain a label.

Take a short clip. Read the transcript or the word-by-word subtitles. Write out the reduced forms you see.

Fast formFull form
gonnagoing to
wannawant to
wouldawould have
shouldashould have
haftahave to
trynatrying to

Then listen while reading. Then listen without the text. Ten minutes of this is enough for one clip. It is not exciting. It works because the sounds stop being noise.

If a speaker says “I shoulda called my mum,” the reduced form “shoulda” now has a shape. If you only hear it without seeing it, it passes as pure sound. Writing reduced forms once beats listening without text ten times.

2. Listen for chunks, not words

Native speakers do not talk word by word. They plan and produce phrase-blocks. “A couple of,” “in the middle of,” “I was going to say” come out as single units. Learners often try to decode each word in sequence. That is too slow. By the time you process the first word, the block has passed.

You need to hear blocks as single objects. You do not need to separate “how’s it goin” into four words. You need to know the block means “how are you.” The fluency page goes deeper into chunks. For fast listening, the point is simple: the ear hunts for blocks, not words.

A learner who hears “a couple a days” tries to find the word “of.” It is not there. The block has swallowed it. If you know the block “a couple a,” the missing “of” is irrelevant. You understand the whole thing immediately.

3. Predict what must come next

Fast listening is partly guessing, but not blind guessing. If a speaker says “I’d like to order...” the next word will be food or drink. If they say “on the other...” the next word is “hand.” Context closes off most choices. Native speakers do this constantly. They are not hearing every sound; they are checking sound against a strong guess.

Prediction is trained by volume. You need lots of listening where you know the topic well enough to guess ahead. That is why the listening practice guide recommends easy material in high quantity. Hard audio breaks prediction because you spend all your energy on decoding. Easy audio lets prediction grow.

When a clip starts “The thing is...” you can predict that an explanation is coming. When a speaker says “I’m not saying...” you can predict a contrast. These small predictions buy you time. You are not half a second behind because you already saw the turn coming.

4. Ladder the speed

This is exposure laddering. Take one hard clip. Play it at 0.75x once. Read the sentence. Play it again at normal speed. Then move to a slightly harder clip. That is a ladder.

Do not live at 0.75x. If you stay there for weeks, you avoid the linking you need to learn. Slowed audio can separate sounds that are never separate in real speech. “Woulda” may sound like “would have” again. That is comfortable, but it is fake. The slow speed is a ramp, not a home.

In the app, you can set a video to 0.75x or 0.5x and tap a sentence to loop it. Loop one hard sentence, not the whole video. Slow it, understand it, then play it at normal speed.

It is tempting to keep everything at 0.75x because it feels safe. Resist that after the first listen. The discomfort at normal speed is the actual lesson.

The aggressive option: shadowing

Sometimes your ear will not catch a phrase until your mouth has produced it. Shadowing means speaking along with a recording, just behind the speaker, not translating—just copying the sound. You can read more on the shadowing page.

Pick a sentence you can read but cannot hear at full speed. Play it at 0.75x and speak along. Then play it at normal speed and speak along again. Your mouth learns the reduced form because it has said it. Then your ear finally recognizes it. This is not passive. You will sound clumsy for a while. That is part of it.

Do not shadow a whole video. Take two or three sentences. Do them until they feel automatic. That usually takes five to ten minutes. Then stop and watch normally.

Honest timeline

If you listen every day for thirty minutes, native speech begins to feel normal in a few months. Not in one week. Daily contact matters more than one long session on Sunday, because the sound patterns need to stay warm in your head.

It also happens accent by accent. You may get comfortable with American YouTube voices first. Then a London accent still sounds fast for a while. That is not because London speech is quicker. The reductions are different. Then Australian English sounds like another code. That is normal. Pick one accent and stay with it until it slows down in your ear. Then start the next one.