すべての記事

English listening breaks at word boundaries

2026年9月13日

Choc Education


The transcript says an aim, while a 27-year-old marketing specialist on the Taipei MRT hears a name in the recording even though all three written words are familiar and each dictionary form can be pronounced accurately in isolation. Vocabulary knowledge is present. The speech stream has been divided at the wrong place. English listening instruction should teach learners to find word boundaries in connected speech. Faster playback can expose the problem, and repeated comprehension questions can confirm it, but neither identifies the split that the listener actually heard. A short clip, a marked transcript and a second attempt can.

Speech has no printed spaces

Recorded speech does not place a white gap after every word. Sounds overlap, weak syllables shrink, consonants attach perceptually to nearby vowels, and familiar sequences compete in the mental lexicon. The listener must infer where one word ends before its meaning can help with the sentence.

Consider an aim and a name. Both can approach /ə neɪm/ in ordinary connected speech. Grammar and context usually settle the choice for an experienced listener. An English learner who expects each dictionary form to arrive intact may search the lexicon for the wrong sequence, miss the following clause, then blame speed.

That diagnosis changes the exercise.

John Field's 2003 article in ELT Journal argued that teachers should examine the perceptual source of L2 listening breakdowns, including lexical segmentation, and use auditory phonetics to classify the trouble. The article appeared in volume 57, issue 4, pages 325 to 334. It proposed a teaching direction rather than a large controlled trial, so its classroom recommendations need support from empirical work. That limitation also keeps the diagnosis useful: teachers can test it against a learner's actual transcript instead of treating it as a universal explanation in that local context.

The first language keeps voting

Listeners learn which sound sequences can sit inside a syllable and which sequences usually signal a boundary. These phonotactic probabilities differ by language. An English listener knows, without stating a rule, that some consonant combinations make one division far more likely than another, whereas a second-language listener brings probabilities learned through years of Mandarin, Taiwanese Hokkien or another first language.

Andrea Weber and Anne Cutler tested highly proficient German users of English with invented sound strings containing real English words. In their 2006 word-spotting experiment, lecture was easier to detect in contexts whose consonants forced a boundary under English rules. German listeners had learned useful English probabilities, yet German-only constraints still influenced their responses. The material was deliberately artificial, and German-English results do not map directly onto Mandarin-English listening. The study isolates a mechanism that ordinary comprehension tests usually hide, because a correct multiple-choice answer cannot reveal whether the listener briefly activated the wrong word before context repaired the sentence.

Study Learners Narrowest limit
Weber and Cutler, 2006 Highly proficient German users of English Artificial word-spotting strings
Mei, Chen and Chen, 2024 56 Chinese users of English Laboratory attention task

More recent work speaks directly to Chinese learners. A 2024 study by Yunhao Mei, Fei Chen and Xiaoxiang Chen compared 30 higher-proficiency and 26 lower-proficiency Chinese learners while they listened to English sentences and responded to a phonetic probe. The groups showed different attention patterns at segmentation boundaries, and the higher-proficiency group had better sentence recognition memory. The task measured selective attention under laboratory conditions. It did not test a ready-made classroom lesson or prove that boundary training alone improves everyday comprehension within that laboratory sample of 56 participants in China.

Now the boundary becomes visible.

Train the cut, then restore meaning

A useful boundary exercise starts with a failure small enough to inspect.

  • Choose five to twelve seconds of speech containing one likely mis-segmentation.
  • Let the learner write exactly what reached the ear, including nonsense, before any transcript appears.

Showing the transcript too early erases the evidence because the printed spaces tell the listener where to cut.

After the first attempt, reveal only the disputed phrase. Mark strong syllables, reductions and possible links. For picked it up, a learner might hear pick tit up or lose the weak vowel entirely. Play the phrase, then the sentence, then a new sentence with the same boundary pattern. One further comparison can expose the cue: place picked it up beside picked Tim up and listen for what remains stable when the word division changes. The exercise stays deliberately narrow, perhaps 67 seconds, because attention should remain on one perceptual decision rather than drift toward every unfamiliar sound in the recording or every spelling surprise in the transcript. The final pass should return to meaning: who picked up what, and when?

Eight weeks of explicit work can matter under some conditions. Faisal Al-Jasser's 2008 intervention study used a native-English group of 12, a non-native control group of 20 and an experimental group of 20 Arabic-speaking learners. Both non-native groups took the same word-spotting pre-test and post-test. The experimental group received instruction in relevant English phonotactic constraints between them, allowing the comparison to separate ordinary course exposure from the added boundary work over the eight-week period. Its segmentation scores improved. The language pairing, small groups and specialised test set a narrow boundary around that result.

Classroom practice need not become a lecture on allowable consonant clusters. Learners can compare two hearings, circle the boundary and collect recurring cases. A teacher may name resyllabification once when pick it sounds as though the final /k/ begins the next syllable. After that, ears need examples.

Examples do the heavy work.

Separate boundary errors from unknown words

Word knowledge still sets a ceiling. The article Known words set the difficulty of English listening explains why a passage with too many unknown items overwhelms recognition. Boundary practice works best when the disputed words already exist in the learner's receptive vocabulary.

A quick diagnostic uses the transcript after the first listen. If the learner reads the phrase and still cannot explain it, teach the words or grammar. If the meaning becomes obvious on sight, replay the sound and locate the split. If the learner recognizes the phrase in isolation and loses it inside the full sentence, widen the audio one clause at a time. Three failures that look identical on a listening score lead to three different lessons, and the distinction prevents a teacher from assigning more vocabulary review to someone whose real difficulty begins only when known forms touch each other in speech.

Podcasts at 0.75 speed may help a beginner retain the thread, but permanent slow audio distorts some timing cues and can become its own listening variety. I prefer a normal-speed fragment short enough to survive four careful hearings. The aim is accurate parsing under real timing, followed by a return to the whole message.

Normal speed stays the target.

Choose one missed line from tomorrow's commute. Before opening captions, type the sounds as heard. Add slashes where the word boundaries seem to fall, even inside a nonsense form. Then compare the transcript and keep only the boundary that fooled you: a/nice cold/hour may be more useful than another ten-minute replay.