On the real test, Task 3 puts an illustration on the screen: an ordinary Canadian place at its busiest hour, of the kind you will find in the five below, a laundromat, a bakery, a school pickup zone, a bike shop, a pharmacy counter. Eight or ten people are in it, each doing one legible thing, clustered into little groups. The instruction is to describe the scene to somebody who cannot see it, and the image stays on screen for the whole 60 seconds, so nothing has to be memorised.
Every question below gives you a real picture, because a describing task without one is not the task. The five CELPIP Speaking Task 3 pictures on this page are our own illustrations, drawn to the brief the real scenes are built on: an ordinary Canadian place at its busiest hour, five or six clusters of people, and in every cluster somebody doing something to somebody else. You look, you choose, you talk. That is the whole loop, and it is the loop you cannot rehearse from a paragraph.
Now the thing that reframes this task. It is tempting to think Task 3 rewards observation, that the high scorers are the ones who spot more. The published performance standards say otherwise, and the evidence is uncomfortable: candidates at wildly different levels are described responding to the same picture. They saw the same room. What separated them was the descriptive machinery they had available to render it. Nobody lost marks for missing the dog in the corner. They lost marks for saying "there is a dog in the corner".
Which is why the single clearest marker of a low band on this task is a list of objects. "There is a washer. There is a basket. There is a folding table." Every sentence grammatical, every sentence useless, because a scene is not a stock list. It is people doing things to each other, and the verb is where the marks live.