Silent Letters and Sneaky Spellings: Clothes, Worcestershire, Foreign and Friends
English spelling is a history book, not a pronunciation guide. The silent letters that follow rules, the words that follow none, and how to stop trusting it.

PTE Speaking is scored by a computer, and the computer does not listen the way a human does. It rewards a steady pace, clear word boundaries and natural stress, and it punishes exactly the habits careful, nervous candidates fall into: pausing to think, restarting a sentence, and speaking one word at a time. That changes which PTE speaking pronunciation tips actually work. Here is how the scoring works and what to change.
PTE pronunciation and fluency scores drop for three reasons. Hesitations, false starts and restarts are counted by the fluency model even when your content is perfect. Wrong word stress and unclear word boundaries stop the speech recognizer from matching the words you said to the words it expected. And an unnatural pace, either rushed or word-by-word, breaks the rhythm the model is built on. The fix is to speak in complete phrases at a steady conversational speed, never go back to correct a word, stress the content words, and finish every final consonant. Read Aloud, Repeat Sentence and Describe Image all reward the same habits, so the drills below work across all three.
The PTE speaking tasks are scored automatically. The system was trained on recordings of many speakers, native and non-native, that human raters had already scored, and it learned which acoustic patterns go with high and low scores. That has two consequences my students find surprising.
First, the machine is not judging whether you sound native. It is judging whether the words you produced can be matched to the words expected, and whether your rhythm, pace and stress fall within the range of speech that human raters found easy to follow. A speaker with a clear Indian, Brazilian or Vietnamese accent can score very highly if the words are recognizable and the flow is steady.
Second, the machine has no patience for hesitation and no ability to give you credit for effort. A human examiner hears you pause, restart and finally produce a beautiful sentence and thinks "good recovery." The fluency model hears a pause, a fragment and a restart, and every one of those counts against you. This is why candidates with strong English are often shocked by low PTE speaking scores.
If you have taken IELTS, the contrast is useful: the IELTS examiner rewards recovery and development, while PTE rewards uninterrupted, even delivery. My guide to moving from Band 6 to 7 in IELTS Speaking shows how different the strategy is.
Oral fluency in PTE is about smoothness: an even pace, natural phrase groups, and no repetitions, false starts or long pauses. Three habits do most of the damage.
Restarting a sentence. If you say "The graph shows... the chart shows the number of..." the recognizer hears two attempts. Even if the second is perfect, the first is scored as a disfluency. The rule in PTE is brutal and simple: never go back. If you say a wrong word, keep going as if it were right. A single wrong word costs you very little. A restart costs you fluency on the whole item.
Pausing to think mid-sentence. Pauses between phrases, where a comma or a period would be, are natural and expected. Pauses inside a phrase, between an article and its noun or between a verb and its object, are read as hesitation. Candidates who translate in their head produce exactly these mid-phrase pauses. The fix is to speak in chunks you have already assembled, which is what the templates later in this article are for.
Going silent. In Read Aloud and the other timed tasks, the microphone stops recording after about three seconds of silence. A long thinking pause can end your recording before you have finished.
One more thing that surprises people: speaking too fast also lowers fluency. Rushing blurs your word boundaries and the recognizer starts missing words. The target is a calm, steady, conversational speed, roughly the pace of a news reader, not a race.
The pronunciation score measures whether your speech is recognizable as the intended words and whether it has the stress and rhythm of natural English. Three things lower it most often, and none of them is "your accent."
Wrong word stress. The recognizer identifies words partly by their stress shape. Development with the stress on the first syllable, comfortable as four full syllables, or percentage stressed on "age" can each fail to match. Since Read Aloud texts are academic, they are full of long words with fixed stress: significant, environment, analysis, technology, economic. I have a guide to the word stress rule that fixes 100 mispronounced words.
Swallowed final consonants. Increased without the final t, students without the s, developed ending in a vowel. The recognizer often does not fill these in from context, and a missing ending can turn one expected word into a different one or into no match at all. Over-finish your words in practice until a clear ending is automatic.
Merged or missing word boundaries. Running words into one blur and separating every word with a gap both confuse the model. Natural English links words within a phrase and pauses between phrases. Learning connected speech helps you link correctly, and marking phrase boundaries in the Read Aloud text helps you pause in the right places. Flap T inside words, as in better and data, is fine and natural; the recognizer is trained on American speech that uses it. If you want that sound explained, see the American flap T.
You see a short academic text, you get a preparation window of 30 to 40 seconds, and then you read it aloud. Here is the routine I teach.
Practice this with a recording and listen back for two things only: did I restart anywhere, and did the stressed words stand out. Most candidates find that ten days of daily Read Aloud practice with those two checks changes their delivery completely.
You hear a sentence once, and you repeat it. Content is scored on how many words you produce in the correct sequence, and pronunciation and fluency are scored on how you say them. Candidates fail this task in two opposite ways: they try to memorize every word and freeze after the first five, or they remember the words but deliver them in a flat, hesitant string that loses all the fluency points.
The trick that works for my students is to listen for the rhythm first and the words second. Every English sentence has a stress pattern, a kind of drumbeat: "The LIBrary will be CLOSED for reNOVation until the SUMmer." If you catch the beat and the stressed words, you have the skeleton, and your brain fills in the small words around them. Repeating the sentence with the right rhythm and most of the words scores far better than repeating all the words in a broken monotone.
Two practical rules. Do not speak while the sentence is playing; you will miss the end. And if you lose the second half, say the first half confidently and stop, rather than filling the space with "um" and fragments.
The best daily practice for this task is shadowing: listening to short native clips and repeating them a half-second behind the speaker, copying the melody. It trains exactly the rhythm memory that Repeat Sentence tests.
You see a chart, map or picture, you get 25 seconds to prepare, and you speak for up to 40 seconds. Content matters, but for the pronunciation and fluency scores what matters is that you keep speaking smoothly for the whole time, and a template is how you do that.
Here is a four-part template that works for almost any image:
Practice the fixed parts of the template until they come out at a steady pace without thought. The 25 seconds of preparation then go only into the words in brackets: two numbers, two labels and one conclusion. Say the numbers carefully, with clear -teen and -ty endings, because numbers are where word recognition most often fails. If the image is a picture rather than a chart, the same template works with "the picture shows," "in the foreground," and "in the background."
Here is what I give students who have two or three weeks before the test, alongside two lessons a week. Every drill is done aloud and recorded.
| Day | Drill | Time |
|---|---|---|
| Daily | Shadow one 60-second native clip, three passes, copying rhythm | 10 min |
| Daily | Three Read Aloud texts: mark chunks and stress, record, check for restarts | 10 min |
| Daily | Ten Repeat Sentence items: rhythm first, no talking over the audio | 8 min |
| Every other day | Five Describe Image items with the template, timed to 40 seconds | 10 min |
| Twice a week | Word stress list: 20 academic words, say each in a sentence | 5 min |
In lessons, I listen for the things the recordings cannot tell you: which final consonants you are dropping, which words you are stressing wrong without knowing it, and where your pauses fall inside phrases. Then we drill those specific words and rhythms in the shared classroom and you take the notes home.
Record yourself doing three Read Aloud texts today, before you change anything. Count the restarts and the mid-phrase pauses. That number is your baseline. Then adopt the one rule that matters most: never go back. Do the daily routine above for seven days, and record the same three texts again. Most candidates hear the difference immediately.
If you would like me to run the diagnostic and build your word list, you can book a first lesson for $25. Bring a recent practice score if you have one. The lesson options and weekly plans are listed in the pricing section.
Get word stress right on long academic words, finish every final consonant, and link words naturally within a phrase while pausing between phrases. Those three habits drive word recognition. Individual accent features matter much less than candidates assume.
Either is fine. The scoring system was trained on a wide range of accents and rewards clarity and rhythm, not nationality. Pick one model and be consistent, because mixing features can produce word shapes the recognizer does not expect.
Almost always because of restarts, self-corrections and mid-phrase pauses. Good speakers are careful speakers, and careful speakers go back to fix things. The machine scores every restart as a disfluency. Adopt the "never go back" rule and your fluency score usually rises within a week or two of practice.
Keep reading as if nothing happened. One wrong or skipped word costs a small amount of content credit. Stopping and re-reading the sentence costs fluency on the whole item, which is worth more. Never restart.
At a steady conversational pace, about the speed of a news presenter. Too slow, with a gap after every word, is scored as hesitant. Too fast blurs word boundaries and the recognizer starts missing words. Even and calm is the target.
Yes, because the habits the computer penalizes are the ones you cannot hear in yourself: dropped endings, misplaced stress and pauses inside phrases. A tutor identifies them in one session and gives you a specific list to drill, which is much faster than guessing from a score report.
English spelling is a history book, not a pronunciation guide. The silent letters that follow rules, the words that follow none, and how to stop trusting it.
One rule explains why nation is 'shun', official is 'shul' and future is 'chur', plus the stress shift and the twenty work words to drill.
Decade, comfortable, salmon, southern, through: the 25 words I correct most at B2 and above, grouped by the trap behind them, with the fix for each.
Private American English coaching with Ashley Curry: pronunciation, accent and confidence, built around your goals.