Teaching Suprasegmental Pronunciation: Stress, Rhythm, and Intonation
Suprasegmental features — word stress, sentence stress, rhythm, and intonation — have a greater impact on listener comprehension than individual phoneme accuracy. Teaching prosody systematically improves L2 intelligibility more effectively than drilling isolated sounds.
What this guide covers
- Why suprasegmentals matter more than segments for intelligibility
- Word stress patterns and common L1 transfer errors
- Sentence stress, weak forms, and rhythm
- Intonation patterns and their communicative functions
- Connected speech features: linking, elision, assimilation
Out of scope
- Individual phoneme instruction (minimal pairs for specific L1 groups)
- Detailed phonological theory and generative phonology
- Speech therapy and clinical pronunciation issues
Why suprasegmentals matter more than individual sounds
Applied linguistics research has consistently demonstrated that suprasegmental features — the "music" of speech — contribute more to listener comprehension than the accurate production of individual consonant and vowel sounds. A speaker with non-standard phonemes but correct stress and intonation patterns is generally more intelligible than one who produces near-native phonemes but misplaces sentence stress.
This finding has important pedagogical implications: classroom time spent on suprasegmentals typically yields a greater improvement in communicative effectiveness than an equivalent amount of time spent on minimal-pair phoneme drills.
Intelligibility (whether the listener can understand the message) is distinct from accentedness (how close the speaker sounds to a native-speaker model). The goal of pronunciation teaching in most ELT contexts is intelligibility, not accent elimination.
Word stress: patterns and common errors
English word stress is lexical: stress placement can distinguish meaning (e.g., REcord vs. reCORD) and is not entirely predictable from spelling. Learners whose L1 has fixed stress (e.g., French, Polish) or syllable-timed rhythm (e.g., Spanish, Cantonese) may transfer patterns that reduce intelligibility in English.
Teaching word stress:
- Mark stress on new vocabulary as part of the FMU presentation (Form dimension)
- Use physical gestures (clapping, tapping, rubber-band stretching) to make stress visible
- Highlight stress-shift patterns in word families: PHOtograph → phoTOgraphy → photoGRAphic
- Draw attention to common suffixes that attract stress: -tion, -ic, -ity
Sentence stress, rhythm, and weak forms
English is a stress-timed language: stressed syllables occur at roughly equal intervals, and unstressed syllables are compressed between them. This produces the characteristic rhythm of English, in which content words (nouns, main verbs, adjectives, adverbs) are stressed and function words (articles, prepositions, auxiliaries, pronouns) are typically reduced to weak forms.
| Word | Strong form | Weak form | Example in context |
|---|---|---|---|
| can | /kæn/ | /kən/ | "I can /kən/ help you" |
| to | /tuː/ | /tə/ | "I want to /tə/ go" |
| was | /wɒz/ | /wəz/ | "She was /wəz/ late" |
| and | /ænd/ | /ənd/ or /ən/ | "bread and /ən/ butter" |
Transcriptions use IPA. Weak forms are the default in connected speech; strong forms are used for emphasis or contrast.
Teaching weak forms alongside sentence stress helps learners both produce natural rhythm and decode rapid natural speech, which is a common difficulty in listening comprehension.
Intonation patterns and communicative function
Intonation — the rise and fall of pitch across an utterance — carries meaning beyond the words themselves. It can signal whether an utterance is a statement or a question, convey attitude (interest, surprise, sarcasm), and indicate whether the speaker has finished or expects a response.
Key intonation patterns in English:
- Falling tone (↘): finality, statements, wh-questions ("Where are you GOing?↘")
- Rising tone (↗): yes/no questions, lists (non-final items), checking understanding
- Fall-rise (↘↗): contrast, reservation, politeness ("I COULD↘↗ help, but...")
- Level tone: continuation, mid-sentence pausing in formal speech
Use "say it three ways" activities: learners say the same sentence with different intonation patterns and discuss how the meaning changes. This builds awareness without requiring metalinguistic terminology.
Connected speech: linking, elision, and assimilation
In natural connected speech, words are not produced in isolation. Sounds are linked across word boundaries ("an apple" → /ənæpəl/), elided ("next day" → /neksdeɪ/), or assimilated ("ten boys" → /tembɔɪz/). These processes are a major source of difficulty for learners trying to understand rapid natural speech.
Teaching connected speech features receptively (in listening activities) is generally more productive than drilling them productively. The goal is for learners to recognise these patterns in input, not necessarily to replicate them perfectly in production.
Integrating pronunciation into skills lessons
Pronunciation does not need to be taught in stand-alone lessons. It is most effective when integrated into other skills work: marking stress on new vocabulary during a reading lesson, practising intonation patterns before a role-play, or analysing connected speech features in a listening transcript. This integration keeps pronunciation teaching purposeful and contextualised.
Resource
Suprasegmental Pronunciation Lesson Plan & Prosody Chart
A ready-to-use lesson plan for teaching word stress, sentence stress, and intonation, with a visual prosody chart for classroom display.
Download: Suprasegmental Pronunciation Lesson Plan & Prosody Chart (PDF)Frequently asked questions
Why are suprasegmental features more critical for intelligibility than individual phonemes?
Research in applied linguistics shows that the prosodic features of speech — stress placement, rhythm, and intonation — contribute more to whether a listener can understand a speaker than the accurate production of individual consonant and vowel sounds. Misplaced word or sentence stress can change meaning or make utterances incomprehensible, whereas a non-standard phoneme in the right prosodic frame is usually still intelligible.
How do backchaining and choral drilling assist in teaching sentence stress?
Backchaining builds a sentence from the end ("...to the SHOP" → "...going to the SHOP" → "She's going to the SHOP"), which naturally preserves the stress and rhythm pattern of the complete utterance. Choral drilling provides low-stakes repetition where individual learners are less exposed, making it easier to focus on prosody rather than accuracy anxiety.
What is the difference between speech intelligibility and native-like accentedness?
Intelligibility refers to whether the listener can understand the message being communicated. Accentedness refers to how much a speaker's pronunciation differs from a particular native-speaker model. A speaker can be highly intelligible while retaining a noticeable accent. The goal of most ELT pronunciation teaching is intelligibility, not the elimination of accent.
Teaching suprasegmental pronunciation focuses on word stress, rhythm, and intonation to significantly improve L2 speech intelligibility and comprehensibility.
Continue reading
References & further reading
Written and reviewed by Ian L. Evans, TeflToday.org. Last reviewed August 4, 2026.
