Back to blog
Speaking · Task 43 min read

Task 4. Making Predictions

From the same scene, you predict what will happen next.

Prep
30 s
Response
60 s
Format
Voice recording

Task objective

Using an image (often related to Task 3), you predict what will likely happen next. The focus is on future language and justification.

How to complete the task

  1. 1

    Identify clues in the image that suggest what comes next.

  2. 2

    State your predictions with future language: 'will', 'is going to', 'probably'.

  3. 3

    Justify each prediction with what you see ('because…').

  4. 4

    Chain 2–3 predictions in a logical sequence.

Tips & tricks

Vary future forms to show range: will / going to / might.

Tie each prediction to a concrete visual clue.

Order predictions in time: first this, then that.

Keep the focus on the future; don't re-describe the scene.

How it's scored

Your recording is assessed on four equally weighted dimensions, and the result is reported on the CLB 1–12 scale. Isolated mistakes aren't deducted: what counts is the overall impression of your full response.

Content & Coherence

They look at how many ideas you give, how good they are, how you organize them and whether you back them up with examples or details. Two well-developed, connected ideas score higher than five loose ideas rattled off in a hurry. A clear order (opening → development → close) makes your answer feel complete.

Vocabulary

They assess the range of words and phrases, whether you use them naturally, and whether they're precise for the context. Repeating the same word or falling back on vague terms ('thing', 'good', 'nice') lowers your score; synonyms, natural collocations and topic-specific vocabulary raise it. It's not about rare words — it's about the exact word.

Listenability

This measures how much effort it takes the listener to understand you: rhythm, pronunciation and intonation; pauses, fillers and self-corrections; plus grammar and variety of sentence structures. You don't need a native accent — you need clarity and a steady flow. Mixing short and long sentences sounds far better than a monotone rhythm.

Task Fulfillment

They check four things: relevance (you answer what's asked), completeness (you cover everything requested), tone (the register fits the person you're addressing) and length (you use the time without trailing off or getting cut short). Here that means predicting what will happen next — not re-describing the scene.

Most lost points don't come from your English — they come from not fulfilling the task: answering something else, falling short, or using a tone that doesn't fit. Before you record, repeat to yourself exactly what's being asked.

Ready to try it?

Practice this task with a real question.

Practice now