TOEIC Link Listening Part 4 — Reading the Speaker's Purpose and Predicting the Next Action
Part 4 gives you a single speaker talking for thirty to forty seconds, and two of the most common questions attached to those talks ask you to infer rather than recall. One asks why — "Why is the speaker calling?", "What is the purpose of the announcement?" The other asks what happens next — "What will the listeners probably do?", "What does the speaker ask the audience to do?" Test-takers who struggle with these are usually listening to the whole talk with equal attention and then trying to reconstruct the purpose from memory. The efficient approach is the opposite: know in advance that purpose is stated near the beginning and the next action is stated near the end, and aim your attention at those two windows.
Purpose lives in the first two sentences
A Part 4 monologue almost always announces its own reason for existing in the opening lines, because that is how real announcements, voicemails, and briefings work. The speaker orients the listener before giving details. Train your ear for the frame phrases that carry purpose:
- Voicemail / phone message: "I'm calling to…", "I wanted to let you know…", "This is a reminder that…"
- Public announcement: "Attention, passengers…", "We regret to inform you…", "Please be advised that…"
- Workplace briefing: "Thanks for coming in. Today I want to walk you through…", "The reason I've gathered everyone is…"
- Broadcast / tour: "Welcome to…", "On today's program we'll be looking at…"
When you hear one of these, the words immediately after it are the answer to the purpose question. "I'm calling to reschedule tomorrow's inspection" tells you the purpose is to change an appointment — you do not need the rest of the message to answer it. Because purpose is front-loaded, a purpose question is one you can often lock in during the first five seconds, then relax and listen for detail. This is the same "detail versus gist" triage that governs where you spend attention across a talk — see detail-vs-gist question triage.
The next action lives in the final lines
The mirror-image rule: a talk almost always closes by telling the listeners what to do. Announcements end with instructions; voicemails end with a request; briefings end with a call to action. Listen for the closing directive:
- "Please proceed to gate 12 for boarding."
- "Be sure to submit your forms by Friday."
- "If you have questions, stop by my office after the session."
- "We'll now take a short break and reconvene at two."
The phrase after please, be sure to, we'll now, or don't forget to is the next action. A question that asks "What will the listeners probably do next?" is answered by the last instruction you hear, not by anything in the body of the talk. So when a next-action question is on your screen, let the middle of the monologue wash over you and sharpen your focus as the speaker begins to wrap up. Recognizing that closing request is closely related to how you handle direct request-and-offer questions, where the answer is likewise carried by a fixed set of asking phrases.
Pre-read the questions so you know which window matters
You get a few seconds before each talk. Use them to read the question stems — not the answer choices, just the stems — and label each one by where its answer will appear:
- A purpose stem ("Why…", "What is the purpose…") → front window, listen hard at the start.
- A next-action stem ("What will… do next", "What does the speaker ask…") → back window, listen hard at the end.
- A detail stem ("At what time…", "How much…") → middle, catch the specific fact in passing.
With this map in mind you are no longer trying to memorize the whole talk; you are catching three targeted moments. That is what separates listeners who finish Part 4 calmly from those who arrive at the questions with a blur of half-remembered sentences. The same pre-reading habit powers the structural anticipation described in the Part 4 announcement structure map.
Beware the paraphrase, not the exact word
Both purpose and next-action answers are almost never the speaker's exact words — they are paraphrased in the choices. "I'm calling to reschedule tomorrow's inspection" becomes the answer "To change an appointment." "Please submit your forms by Friday" becomes "Complete some paperwork." Train yourself to match the idea, not the vocabulary, and to distrust a choice that repeats a word from the talk verbatim, because Part 4 frequently uses a word you heard as bait attached to the wrong idea. Reasoning from what the speaker intends rather than from surface words is the heart of inference items generally, as in reason questions that ask why the speaker mentions something.
Practice routine
For two weeks, before each Part 4 talk, sort its questions into front / middle / back and write one letter (P for purpose, A for action, D for detail) beside each. Then listen with that map: lock the purpose in the opening, catch details in passing, and wait for the closing instruction to answer the next-action item. After each set, check whether the right answer was a paraphrase of the opening frame or the closing directive — it almost always is. Once you stop trying to remember the whole monologue and start listening for two predictable windows, these two question types turn from the hardest part of Part 4 into two of the most reliable points on the section.