Understanding Task 4
Task 4 uses the same photograph as Task 3 and asks you to predict what will probably happen next — what will the people in the image do? What will occur in this scene after the photograph was taken? You have 30 seconds to prepare and 60 seconds to speak. Task 4 tests your ability to use future language accurately, to build logical predictions from visible evidence, and to give reasons for your predictions.
The most important rule for Task 4 is that your predictions must be rooted in visible clues from the photograph — not in random speculation. A prediction without a visual basis does not fulfill the task. A prediction that connects directly to something visible in the photograph and then uses future language to describe what will happen demonstrates both Task Fulfillment and Coherence and Cohesion.
The Three-Part Prediction Frame
Structure your Task 4 response using this three-part frame:
Part 1 — Anchor (10–12 seconds): Reference one specific visible element from the photograph that serves as the basis for your first prediction. "Looking at this scene, I can see [specific detail]. Based on this, I think that [first prediction]."
Part 2 — Two Predictions with Reasons (35–40 seconds): Give two specific, logically grounded predictions with a reason for each. Each prediction should connect to something visible in the photograph. Use hedged future language throughout — not certain future language. Predictions presented as certainties sound unnatural and unrealistic; hedged predictions sound thoughtful and natural.
Part 3 — Broader Outcome (8–10 seconds): End with a brief overall outcome or consequence — what will the situation look like once the predicted events have unfolded. This provides a natural conclusion and demonstrates the ability to project beyond the immediate next moment.
Future Language for Task 4
Task 4 specifically requires you to demonstrate command of future forms. Use a variety of these structures rather than relying entirely on "will."
Future Form | When to Use It | Example |
|---|---|---|
will + verb | Confident prediction based on clear evidence | "They will probably start serving the food shortly." |
is / are going to + verb | Imminent action, clear intention visible in the photo | "The woman looks like she is going to make an announcement." |
might / may / could + verb | Less certain prediction | "They might decide to move the meeting outside." |
is likely to + verb | Moderate confidence prediction | "The group is likely to wrap up their discussion soon." |
I would expect + noun/verb | Thoughtful inference based on context | "I would expect some questions to follow the presentation." |
Once / After / When + [event], [result] | Sequencing future events | "Once they finish setting up, the event will probably begin." |
Full 60-Second Template for Task 4
Anchor (10–12 seconds): "Looking at this photograph, I can see [specific visible element]. This tells me that [immediate context]. Based on this, I would predict that [first prediction]."
Prediction 1 with reason (18–20 seconds): "Specifically, I think [prediction 1] is likely to happen because [visible reason from the photograph]. [One additional sentence elaborating on this prediction]."
Prediction 2 with reason (18–20 seconds): "I would also expect that [prediction 2]. The reason I say this is [visible or logical basis]. This could lead to [brief consequence]."
Broader outcome (8–10 seconds): "Overall, I think the situation will probably [broader outcome or resolution]. It looks like [final statement about the scene's direction]."
Worked Example: Complete Task 4 Response
Photograph described in Task 3: A group of four people seated around a table in what appears to be an office meeting room. They have papers and laptops in front of them. One person is standing and appears to be presenting something. The atmosphere looks focused.
Task 4 response: "Looking at this photograph, I can see a group of colleagues in what appears to be an active work meeting, with one person standing and presenting. Based on this, I would predict that the presentation is about to wrap up and the group will move into a discussion phase. Specifically, once the presenter finishes their points, the seated team members will probably ask questions or share their own perspectives. I notice that a couple of them have notes in front of them, which suggests they came prepared to contribute — so I would expect a fairly lively back-and-forth conversation to follow. I would also predict that by the end of the meeting, the group will likely assign specific action items or next steps to each person. Meetings at this stage of a project tend to end with a clear division of responsibilities, and the focused atmosphere in the room suggests that decisions are being made. Overall, I think the meeting will probably conclude productively, with everyone leaving with a clear understanding of what they need to do next."
Key Takeaways from Lesson 9
Task 4 predictions must be based on visible clues from the photograph — not on random ideas.
Use a variety of future forms: will, going to, might, may, is likely to, would expect.
Give two predictions, each with a reason connected to something visible in the photograph.
Use hedged future language — predictions presented as absolute certainties sound unnatural.
End with a broader outcome that projects the scene's likely resolution.
Use Task 3 preparation time to identify your Task 4 prediction clues — this saves preparation time in Task 4.