Speaking used to be the most mechanical part of the DET. Since July 2025 it is the most conversational: an on-screen character now asks you six to eight questions and adapts to your answers, and the two read-and-repeat tasks that used to pad the section are gone.
Speaking takes up to about 11 minutes of the sitting. Three of the four tasks are single recordings with one attempt each; the fourth is a live-feeling conversation that changes direction based on what you say.
| Question type | Per test | Time | Adaptive | Scored as |
|---|---|---|---|---|
| Speak About the Photo | 1 | 90 seconds | No | Speaking criteria: content, discourse coherence, fluency, grammar, lexis, pronunciation |
| Read, Then Speak | 1 | 90 seconds | No | Speaking criteria |
| Interactive Speaking | 6-8 questions | 35 seconds per answer | Yes | Speaking criteria |
| Speaking Sample | 1 | 3 minutes | No | Speaking criteria, and sent to institutions as video |
There is no retake on any speaking task. Once you click Record Now, the answer you give is the answer that is graded. Practising the first ten seconds of an answer is therefore worth more than practising the whole of it.
Describe a photograph out loud for up to ninety seconds. One recording, no second take.
Read a prompt with bullet points, prepare for twenty seconds, then talk for a minute and a half.
A spoken conversation with an on-screen character across two topics. Each question is asked once.
The longest speaking turn on the test, recorded on video and sent to the universities you apply to.
Read a written sentence out loud and let the transcriber check it. Retired, still the cheapest pronunciation drill. Retired from the live test in July 2025.
Duolingo replaced two monologue tasks with one adaptive conversation. Read Aloud and Listen, Then Speak were removed from the live test, and Interactive Speaking was added in their place. If your preparation book still drills Read Aloud, it predates the change.
| Before July 2025 | Now | |
|---|---|---|
| Speaking tasks | Read Aloud, Listen Then Speak, Speak About the Photo, Read Then Speak, Speaking Sample | Speak About the Photo, Read Then Speak, Interactive Speaking, Speaking Sample |
| Adaptive speaking | Read Aloud only | Interactive Speaking, fully adaptive |
| What is rewarded | Clear pronunciation of written sentences | Answering an unpredictable question in 35 seconds, twice on two different topics |
| Minimum speaking times | Enforced on several tasks | Removed |
Read Aloud is still the cheapest drill in existence for chunking and pronunciation, which is why it survives as practice on this site — see the Read Aloud page — but you will not meet it on a live test.
Speaking answers are graded on the four writing criteria plus two that only apply to speech.
| Criterion | What is measured | What raises it |
|---|---|---|
| Content | Task achievement, style, development, effect on the listener. | Answer the actual question, then give one reason and one example. Do not describe what you are about to say. |
| Discourse coherence | Clarity, cohesion, progression of ideas. | Signpost out loud: "There are two reasons. First… The other thing is…" |
| Fluency | Speed, chunking, breakdowns and repairs. | Pause at clause boundaries, not mid-phrase. There is no ideal speed; there is an ideal rhythm. |
| Grammar | Grammatical range and accuracy. | One conditional, one relative clause, more than one tense. Range counts even when accuracy slips. |
| Lexis | Diversity, sophistication, word choice, word formation. | Precise verbs. If a word will not come, describe the thing rather than going silent. |
| Pronunciation | Intelligibility, individual sounds, word stress, sentence stress, intonation. | All standard accents are accepted. What is graded is whether you are easy to understand, and whether stress lands on the right syllable. |
Accent is not penalised. Word stress is. baNAna is correct and banaNA is not, and stress errors are what the pronunciation score actually reacts to.
A Duolingo character asks six to eight questions, three or four on each of two unrelated topics. You get 35 seconds per answer, the timer starts the moment the question ends, and you hear each question exactly once. Your answer is transcribed and scored in real time, and it decides how hard the next question is.
Prompts and model answers are on the DET Speaking samples page, and the full band descriptors are on the DET Speaking scoring page.