CELPIP Speaking Task 3: Describing a Scene – Template, Timing & Level 9 Strategy
By MyCELPIP Editorial Team · Updated August 7, 2026
CELPIP Speaking Task 3 is Describing a Scene. You are shown an illustration and asked to describe what is happening to a person who cannot see the picture. The current 2026 CELPIP Speaking Pro material gives you 30 seconds to prepare and 60 seconds to speak.
The official guidance does not tell you to describe every object. It recommends starting with a general statement, focusing on selected details, building a clear picture for the listener, and using descriptive language for people's appearance, actions, and feelings.
Quick strategy: identify the setting first, choose three or four useful areas or groups, move through them in a logical order, and describe actions with specific nouns, verbs, positions, and visible details.
Verify the current format in the official CELPIP Speaking Pro: Target 9+ 2026 study pack and the CELPIP-General test format.
CELPIP Speaking Task 3 format
- Task: Describe what is happening in an illustration
- Preparation time: 30 seconds
- Speaking time: 60 seconds
- Listener: a person who cannot see the picture
Task 3 is different from Task 8. In Task 3, the goal is to explain an ordinary scene and its visible activity. Task 8 asks you to describe an unusual situation in a role-play context.
What CELPIP raters evaluate
Content and coherence
Your listener should be able to form a mental picture. A strong answer usually establishes the setting, then describes selected people or areas in a sensible order. Randomly jumping from one corner to another makes even accurate details harder to follow.
Vocabulary
Use specific nouns and action verbs when you know them, but do not freeze if you do not know one exact word. Describe the object or action with language you control. Precision matters more than showing off rare vocabulary.
Useful categories include:
- position: in the foreground, behind, beside, near the entrance, on the left;
- actions: carrying, reaching for, waiting, examining, talking, pushing;
- appearance: wearing, holding, standing, seated;
- mood: relaxed, busy, excited, confused, impatient.
Listenability
Use a steady pace and short transitions. Your response should sound like one connected description, not a list of labels. Present continuous forms such as “is carrying” and “are talking” are naturally useful because you are describing actions in progress.
Task fulfillment
Stay on visible information. Task 3 is primarily description, not prediction. Predictions belong in Task 4.
A 30-second picture scan
Use your preparation time in this order:
1. Setting — Where is the scene?
2. Main activity — What is generally happening?
3. Foreground — Who or what is easiest to describe?
4. Middle/background — Choose two more useful areas.
5. Mood — What overall impression does the scene give?
Do not try to memorize every item in the image.
A flexible Task 3 structure
Opening
Give the listener the setting and overall activity.
- “This looks like a busy outdoor market with several people shopping and talking.”
- “The picture shows a family park on a sunny afternoon, and people are doing several different activities.”
Area 1
Move to a clear location.
- “In the foreground…”
- “On the left side…”
- “Near the entrance…”
Describe a person + action + one useful detail.
Area 2
Use a location transition.
- “A little farther back…”
- “In the centre of the scene…”
- “Behind them…”
Area 3
Add another distinct activity or group.
Closing
If time allows, give a short overall impression rather than introducing a new person at the last second.
- “Overall, the place looks lively but well organized.”
- “It seems like a relaxed afternoon with families enjoying the park.”
Original practice scene
Imagine an illustration of a community park. In the foreground, a woman is sitting on a bench reading while a small dog waits beside her. Near a fountain, two children are trying to catch soap bubbles. In the centre, a man is pushing a stroller and talking on his phone. Farther back, three teenagers are playing basketball, and an older couple is walking along a path under several large trees.
Original sample response
This picture shows a fairly busy community park on a pleasant day. In the foreground, a woman is sitting on a bench and reading a book while a small dog is waiting quietly beside her. Just behind them, two children are reaching for soap bubbles near a fountain, and they look very excited. In the centre of the park, a man is pushing a baby stroller while talking on his phone. Farther in the background, three teenagers are playing basketball on a small court, and an older couple is walking together along a tree-lined path. Overall, there are several activities happening at the same time, but the atmosphere looks relaxed and family-friendly.
Why the response works
- It establishes the setting immediately.
- It moves from foreground to background.
- It chooses useful details instead of describing every object.
- Actions are described with precise verbs.
- Position words help the listener reconstruct the picture.
- The final sentence summarizes the mood.
Common Task 3 mistakes
Trying to describe everything
You only have 60 seconds. Select details you can explain accurately.
Jumping around the picture
Use a spatial path such as foreground → centre → background or left → right.
Naming objects without actions
“A woman, a dog, two children, a fountain” is a list. Explain what the people are doing and where they are.
Predicting what will happen
Save prediction language for Task 4 unless a tiny inference is necessary to make the visible action clear.
Inventing details that are not visible
Do not build a story that the image does not support. Your job is to make the scene understandable.
Level 9-focused self-check
After recording, ask:
Content/Coherence — Did I set the scene and organize details logically?
Vocabulary — Did I use accurate nouns, action verbs, and position language?
Listenability — Was my description easy to follow at a steady pace?
Task Fulfillment — Did I describe the visible scene rather than drift into prediction or storytelling?
Practice Task 3
Use CELPIP Speaking Practice to record timed Task 3 responses. Review the complete task map on the 2026 CELPIP Speaking Guide, then use a full CELPIP mock test when you are ready to combine all four skills.
Frequently asked questions
How much preparation time is there for Task 3?
The current official format gives you 30 seconds.
How long do I speak?
You have 60 seconds to record your description.
Do I need to describe every person?
No. Official CELPIP guidance recommends focusing on selected details rather than trying to describe everything.
Should I use present continuous?
It is naturally useful for actions happening in the illustration, but grammar should serve the description rather than become a forced formula.
What if I do not know the exact word for an object?
Describe it with common words you know. Clear paraphrasing is better than stopping for several seconds while searching for one noun.