PTE Core Describe Image
Describe Image shows you an image and gives you 25 seconds to prepare before you speak about it for up to 40 seconds. Unlike Read Aloud or Repeat Sentence, there is no fixed content to reproduce, which means structure is what separates a strong response from a rambling one. A short, reusable template, planned in your 25 seconds of prep, does most of the work.
At a glance
| Detail | What to expect |
|---|---|
| Items on your test | 3 to 4 |
| Preparation time | 25 seconds |
| Response time | 40 seconds |
| Prompt | An image is shown on screen |
| Skills scored | Speaking |
| Traits scored | Content, Pronunciation, Fluency |
| Scoring method | Partial credit, reported on the 10 to 90 scale |
How it is scored
Pearson scores Describe Image on three official traits: Content, Pronunciation, and Fluency. Content checks how completely you describe the image's key elements and their relationships or implications, rather than listing disconnected details. Pronunciation checks how clearly you produce vowels and consonants and whether you place stress correctly at the word and sentence level. Fluency checks the rhythm and pacing of your description, including hesitations, repetitions, and false starts, rather than a disjointed list of details. Unlike Read Aloud and Repeat Sentence, Describe Image scores into a single communicative skill, Speaking, reported on the 10 to 90 scale, and every trait uses partial credit, so an incomplete description still earns marks for the parts you cover well.
If you are tracking progress against the Canadian Language Benchmarks (CLB) for Express Entry, Pearson's Score Guide aligns a Speaking score of 84 to 88 with CLB 9, 76 to 83 with CLB 8, and 68 to 75 with CLB 7. Treat this table as a general guide and confirm current thresholds on the IRCC website before you plan around it.
Exam Hero's own AI Coach grades practice attempts too, using a different, simplified label set: Oral Fluency, Pronunciation, Vocabulary, and Grammar. That is the same four label feedback the app gives on every speaking task, not a restatement of the three official Describe Image traits above, since it reports Vocabulary and Grammar, which are not official traits on this task at all, and does not separately report Content. Use AI Coach to sharpen your delivery, and use the trait list above to know what the real test grades.
The method
Use your 25 seconds of prep to scan the image once for its overall subject, then pick out the two or three elements that stand out most, the largest, the most different, or the most central to what the image is showing. Do not try to mention everything; a focused description of a few features beats a rushed list of everything you can see.
A reusable three part template, sized to your 40 second response:
- Opening, about 5 seconds: One sentence stating what the image shows overall. For example, "This image shows..." or "What I can see here is...", followed by the general subject.
- Middle, about 25 to 30 seconds: Describe your two or three chosen features, one at a time. For each one, name what it is, where it is, and one specific detail about it. Use connecting language such as "the most noticeable part is...", "next to that...", or "in comparison..." to move cleanly from one feature to the next.
- Closing, about 5 seconds: One sentence that wraps up your description, such as a brief overall impression or how the features you mentioned relate to each other. This signals a deliberate ending rather than trailing off when you run out of things to say.
Rehearse this three part shape on practice images until choosing an opening line and picking your top features becomes automatic. That frees up your 25 seconds of prep for actually scanning the image, instead of trying to remember what to say first.
Worked example
Suppose the image shows a busy public space with several distinct areas of activity. A response following the template might sound like this: "This image shows a busy outdoor public space with several different areas of activity going on at once. The most noticeable part is a large group of people gathered near the center, which suggests this is the main focus of the scene. Next to that, on one side, there is a smaller, quieter area with only a few people, creating a clear contrast in how busy different parts of the space are. There also seem to be some structures spread across the background that add more detail to the overall setting. Overall, the image gives a strong sense of contrast between a busy central area and a quieter surrounding space."
This response opens with a clear overall statement, covers three distinct features, the central group, the quieter area, and the background details, using comparison language to link them, and closes with a wrap up sentence rather than stopping abruptly. That structure, more than any single vocabulary choice, is what earns strong marks on Content and Fluency together.
Common mistakes
- Listing everything you notice without any structure. A flat list of unconnected observations scores worse than a focused description of two or three features tied together with connecting language.
- Spending all 25 seconds staring at the image without a plan. Use the prep window actively: pick your opening line and your top two or three features before you start speaking.
- Running out of things to say before the 40 seconds end. Silence for the remaining time costs you on Content and Fluency. Plan a middle section with enough detail, and always have a closing line ready to fill the last few seconds deliberately.
- Guessing at specific labels, numbers, or names you cannot actually make out. A wrong specific detail can hurt Content more than a safely general description of the same feature. Describe what you can confidently see.
- Trailing off instead of closing deliberately. An unfinished sounding response reads as disorganized even when the content was accurate. Always end with a short wrap up sentence.
- Speaking too quickly to fit more content in. Fluency rewards natural, connected speech, not maximum word count. A slightly shorter, clearly paced description beats a rushed one that blurs together.
Frequently asked questions
How many Describe Image questions are on the PTE Core test?
You will see 3 to 4 Describe Image items in your test.
How much time do I get to prepare for Describe Image?
You get 25 seconds to prepare before your 40 second response begins.
What score does Describe Image count toward?
It counts only toward your Speaking score, unlike Read Aloud or Repeat Sentence, which each count toward two skills.
What kind of image will I see?
Pearson's published test format materials confirm an image is shown, without specifying the type. Prepare a description template that works for any visual rather than assuming a particular format.
What traits does PTE Core score in Describe Image?
Pearson scores three official traits: Content, Pronunciation, and Fluency.
Practice Describe Image now
The three part template only becomes automatic once you have used it under real time pressure. Try a real prompt in the Describe Image question bank, then check your Content, Pronunciation, and Fluency feedback instantly. If you want to build the same steady delivery on a passage based task, see the companion guide to PTE Core Read Aloud.