You cannot speed up the audio, so TOEFL Listening pacing is about decision speed. A minute-by-minute plan for 47 items across two adaptive modules.
11 min readMonthly & Premium
Quick answer
Listening gives you 29 minutes for 47 items, an average of about 37 seconds each, but most of that time is audio you cannot compress. The only adjustable time is your decision window after each clip. Pace by deciding faster, never by rushing the listening, and protect your attention for the first module.
The arithmetic, and why it misleads
Twenty-nine minutes, forty-seven items. That is roughly 37 seconds an item, which sounds comfortable until you notice what those seconds are made of. A large share of every minute is audio playback, and audio playback runs at its own speed regardless of how efficient you are. You cannot read faster to buy time in Listening the way you can in Reading. The clip takes as long as the clip takes.
What that leaves you is the answering window, the seconds between the end of a clip and your click. That window is the only compressible time in the entire section, and it is where every pacing decision you make will land. This is the central difference between pacing Reading and pacing Listening, and most students never make the adjustment. They arrive with a Reading strategy, try to work faster during the audio, discover there is nothing to work faster at, and end up simply anxious rather than quicker.
So the target is not speed of comprehension. It is speed of commitment. A student who understands slightly less but commits in five seconds will outscore a student who understands slightly more and commits in twenty, because the second student pays for those extra seconds with the opening of the next clip. In a section where the audio does not wait, hesitation is not free thinking time; it is borrowed from the item after it.
A working time budget
Task type
Target decision time
Why
Choose a Response
Under 5 seconds from the options appearing.
The judgement is a single call about speaker intent. Extra deliberation reliably makes accuracy worse here, not better, because it lets the word-echo distractors work on you.
Announcement question
About 10 seconds each.
The answer is in your notes or it is not. Ten seconds is enough to read the options and match; twenty seconds is enough to talk yourself out of a correct answer.
Conversation question
About 10 to 15 seconds each.
Attribution questions need a moment to check the person as well as the action, which is a genuine second step and worth paying for.
Academic talk, fact question
About 10 seconds.
Same as an announcement. Either the chain is in your notes or you are guessing, and guessing quickly is better than guessing slowly.
Academic talk, inference question
Up to 20 seconds, once per talk.
This is the one item in a clip that genuinely rewards thinking. Budget for it explicitly rather than letting it steal time by accident.
Targets for the answering window only, after the audio has finished. These are deliberately aggressive, because the failure mode in Listening is always overspending.
How to run the section
1
Before the intro audio: set up and breathe out
Paper positioned, pen in hand, both feet flat. One long exhale. This is not relaxation for its own sake; a held breath raises your heart rate and narrows your listening. You want to enter the first clip already settled, because the first module is the router and the first two items are the ones most often lost to a cold start.
2
The short reply block: move
This is a rhythm section. Clip, judge, click, next. Do not let a single item break the rhythm, because the rhythm is what carries you through the block without accumulating delay. If one goes badly, the correct response is to speed up slightly on the next two, not to slow down and recover.
3
Between every longer clip: three seconds of reset
Pen down, eyes up, one exhale. Three seconds. It feels wasteful and it is the highest-return three seconds in the section, because the alternative is carrying the residue of the last clip into the opening ten seconds of the next one, which is where the topic is announced.
4
The last clip of a module: spend your reserve
The academic talk sits at the end, carries the most questions, and arrives when you are most tired. Know this in advance and decide now that you will spend your remaining concentration there rather than rationing it. There is nothing after it in the module to save energy for.
5
Between modules: a hard reset
Whatever happened in the first module is finished and cannot be changed. Ten seconds of deliberate posture change, a full breath, and a decision to treat the second module as a fresh start. Students who spend the second module auditing the first lose items in both.
Attention is the scarce resource, not time
Run enough practice sections and a pattern appears in where the losses cluster. It is not random and it is not evenly spread. There are two reliable dips.
The first is early, somewhere between the first and the fifth item. Your ear is still calibrating: to the voice, to the recording level, to the pace, to the fact that this is now happening. Every item lost here is lost to being cold rather than to being unable, and it is the most expensive place to lose items because those items feed the routing decision. The countermeasure is entirely outside the test: ten minutes of English audio before you sit down, at normal speed, with no subtitles. A podcast on the way to the centre. It is the highest-value ten minutes in your whole preparation and it costs nothing.
The second dip is late, in the final clip of each module, which is precisely where the questions are densest. This one is fatigue rather than warm-up, and the countermeasure is stamina practice: do full 29 minute sections in one sitting rather than splitting them into blocks. Practising in five-minute pieces trains a concentration span you will not have on test day, and it feels productive, which is what makes it dangerous.
There is a third loss that is not a dip but a spiral, and it deserves its own treatment.
Pacing decisions under pressure
Three situations that come up in almost every practice section. The reasoning explains what the alternative actually costs.
You feel you are falling behind partway through the first module. What should you compress?
The audio runs at its own speed, so the only compressible time in the section is the gap between the clip ending and your click. Faster commitment is the one lever that exists, and it is usually available because most students overspend there by ten seconds an item.Reading options during the audio is the most damaging habit in Listening. You cannot do both, and the options are written from real vocabulary in the clip specifically so that reading them primes you towards the distractors.Abandoning notes trades a small saving for a large loss, especially on conversations where attribution is the whole question. The right move is shorter notes, not no notes.The section timer is fixed, but your attention is not. Falling behind in Listening is real; it just takes the form of attention debt rather than clock debt.
You are unsure of a question about an academic talk. Should you flag it and come back at the end?
In Reading, returning works because the passage is still there. In Listening the evidence was audio and it is gone, and your memory of it degrades with every clip that follows. Answer now, while the clip is freshest, and spend the saved seconds on items you have not met yet.Reviewing later works when the source material is still available. Here you would be re-reading four options with less memory of the clip than you have right now, which makes a worse decision rather than a better one.Finishing early is not the usual outcome, and even with spare time the fundamental problem holds: the audio is gone and your recall of it has decayed.There is no penalty attached to flagging or to a wrong answer. The reason not to flag is practical, not punitive.
Where do most test takers lose items, and what fixes it?
Losses cluster at two predictable points: early, while your ear is still calibrating, and late, when fatigue meets the densest block of questions. Both have specific countermeasures that have nothing to do with your English level: ten minutes of listening before you start, and full-length practice in one sitting.Losses are not evenly distributed, and treating them as a vocabulary problem sends you to the slowest available fix for a problem that is mostly about attention.The middle of a module is usually the most stable stretch, once your ear has settled and before fatigue arrives. Fuller notes would also cost you listening rather than buy you accuracy.Academic talks carry more questions, so they show more absolute losses, but the clustering is about position in the section rather than task type. Subject knowledge is never required, since terms are defined inside the clip.
Pacing rules to fix before your exam
Under five seconds on every Choose a Response item, without exception.
One inference question per academic talk gets up to twenty seconds. Everything else gets ten.
Never read the options while audio is playing.
Never flag a Listening item to return to. Commit now.
Three seconds of reset between longer clips: pen down, exhale, eyes up.
Ten seconds of deliberate reset between modules, and no auditing of the first.
Ten minutes of English audio before the test starts, at normal speed, no subtitles.
Full 29 minute sections in practice, never split into blocks.
The fortnight before, and the morning itself
In the last two weeks, stop adding new material and start rehearsing the routine. Do at least four full-length Listening sections in one sitting each, at the time of day your exam is scheduled. The point is not the score. The point is that the routine becomes automatic enough to survive the moment when you are nervous and something goes wrong, which it will.
Track two numbers on each of those sections. First, your raw score out of 47 and your first-module percentage separately, because the routing threshold sits at 60 percent of the first module and you want to know whether you are clearing it comfortably or scraping it. Second, count the items you lost to attention rather than to comprehension. If that second number is above three or four, the fix is pacing and warm-up, not more English.
On the morning itself, three things matter and none of them is revision. Eat something, because concentration in the last clip of a module is partly a blood sugar problem. Listen to ten minutes of English on the way, unsubtitled, at normal speed, so that the first clip is not the first English your ear has processed that day. And check your headphones on unfamiliar audio rather than on a song you know, because a familiar track will sound fine through equipment that is not actually working properly.
Then, in the section, run the routine and let the score take care of itself. The single most common thing that separates a good practice score from a disappointing real one is not knowledge. It is a student who knew all of this and abandoned it in the first ninety seconds because one clip went badly. Decide in advance that a missed clip costs one item, and hold to it.
Monthly & Premium
7 more sections in this lesson
You have read the opening. The rest covers the method in full, with worked examples and practice questions that explain why each wrong answer is wrong.
All 45 lessons: Reading, Listening, Writing and Speaking
Real exam audio and record-yourself speaking drills
There is no per-question timer. The section runs on a single 29 minute clock across 47 items, which averages about 37 seconds each, though most of that is audio playback rather than answering time.
Even where the interface allows it, there is no benefit. The evidence was audio and it is gone, and your recall of the clip only degrades as later clips arrive. Answer while the clip is fresh.
No. Reading and listening compete for the same attention, and the wrong options are built from real vocabulary in the clip, so reading early primes you towards the distractors rather than away from them.
Guess immediately, then reset before the next clip starts. The comprehension failure costs one item; the brooding that follows it typically costs two or three more.
For drilling a single task type, yes. For pacing, no. Full 29 minute sections in one sitting are the only way to train the stamina the last clip of each module actually needs.
It is one of the highest-value habits available, because early items are lost to a cold ear rather than to ability, and those early items sit in the module that decides your routing. Ten minutes of unsubtitled English beforehand is enough.