Speaking: Pronunciation — Connected Speech
8 min read
Connected speech refers to the systematic ways natural spoken English blends, links, and sometimes shortens the sounds between neighbouring words, rather than pronouncing each word in a sentence with the same crisp, fully separated clarity you'd hear if someone read a list of individual words aloud. This is a genuinely different feature from word stress, which concerns which syllable within a single word is emphasised, and from intonation, which concerns the rise and fall of pitch across a whole sentence — connected speech operates specifically at the boundaries between words, governing how those word-edges interact with each other in fast, natural speech. Recognising and, to an appropriate degree, producing connected speech is a significant part of what separates genuinely fluent-sounding delivery from technically accurate but oddly stilted, over-separated speech where every single word is pronounced in careful isolation.
Linking is the most fundamental connected-speech pattern, occurring when the final sound of one word flows directly into the initial sound of the next without any perceptible gap or glottal stop between them. The most common and reliable form of linking happens when a word ending in a consonant sound is immediately followed by a word beginning with a vowel sound: "turn it off" naturally links into something closer to "TUR-ni-TOFF" in fluent speech, with the final consonant of each word sliding onto the vowel that follows rather than each word standing separately. Similarly, "an apple" links into "a-NAP-ple," and "far away" links into "fa-ra-WAY." This linking isn't sloppy or incorrect pronunciation — it's the standard, expected pattern in fluent native and near-native English, and its complete absence, where every word is separated by a small pause or glottal stop regardless of the sounds involved, is itself what produces the robotic, word-by-word quality that marks noticeably less fluent speech.
Elision is a distinct connected-speech pattern in which a sound, typically a consonant, is dropped entirely in rapid natural speech rather than merely blended with a neighbouring sound. The "t" in "next day" is frequently dropped entirely in fast natural speech, producing something closer to "nex day"; the "d" in "kind of" is commonly elided, producing something closer to "kinda" even in speech that isn't especially casual. This differs from linking in kind, not just degree — linking blends two full sounds together at a word boundary, while elision removes a sound that would otherwise have been there altogether, generally because certain consonant clusters are simply difficult and inefficient to articulate at natural conversational speed, and the missing sound doesn't meaningfully affect comprehension since the surrounding context still makes the intended word clear.
Weak forms represent a third distinct pattern, specific to a particular set of very common function words — articles, prepositions, auxiliary verbs, and conjunctions such as "a," "to," "for," "can," "was," "and," "of" — which are pronounced with a reduced, often neutral vowel sound in their unstressed, natural connected-speech form, quite different from how the same word sounds when said in isolation or with deliberate emphasis. The word "to" in isolation is pronounced with a clear "oo" vowel, but in a natural sentence like "I want to go" it typically reduces to something closer to a neutral "tuh" sound, unless a speaker is placing special emphasis on it. Similarly, "and" in "fish and chips" commonly reduces toward "fish 'n chips" in fluent, natural speech, with the vowel in "and" nearly disappearing. Weak forms exist specifically because these function words carry grammatical structure rather than the main semantic content of a sentence, and native speech systematically de-emphasises them relative to the content words (nouns, main verbs, adjectives) that carry the actual meaning — this is a structural rhythm feature of English, not careless or lazy speech.
Ready to put this into practice?
Get 2 free scored Writing submissions, plus full Reading and Listening practice — no card required. Speaking is a separate credit-pack add-on.
Get 2 free scores