Interactive Speaking
The task that replaced Listen, Then Speak in July 2025. A short spoken conversation with an animated character: you hear one question at a time, answer each in about thirty-five seconds, and every question is built from what you just said.
- 35s each on the clock, 6–7 per full test
- Feeds your Speaking and Listening subscores
- Built for the current July 2025 DET format, including the new Interactive Speaking task
Interactive Speaking was added to the Duolingo English Test on 1 July 2025, replacing Listen, Then Speak. Instead of one isolated spoken answer, you hold a short structured conversation: an animated character asks a question, you record a reply of up to about thirty-five seconds, and the next question is generated from what you just said.
That last detail is the whole task. Because each question follows your previous answer, you cannot prepare fixed content, and a vague first reply leaves the character nothing to ask about except more vague follow-ups. A specific, concrete first answer sets up an easier conversation.
It tests something the older speaking tasks did not: whether you can listen, understand and respond on the spot in a way that stays coherent across a real exchange.
What you actually see
An animated character on screen, a question played as audio, a short moment to think, then a recording indicator. The questions come one at a time and each plays once.
- An animated character who asks the questions, on screen throughout.
- About six to seven questions, each played as audio one at a time.
- Up to roughly thirty-five seconds to record each answer.
- Each question generated from your previous answers, so the conversation genuinely adapts.
- One appearance as its own section in a full test.
How the conversation builds on you
Suppose the opening question is "What is a hobby you enjoy, and why?"
A weak first answer — "I like watching movies. It is fun and relaxing." — is grammatically fine but gives the follow-up nothing to grip, so the next question can only be generic and you start from scratch again.
A strong first answer plants specifics the conversation can grow from: "I enjoy bouldering, which is indoor rock climbing on short walls. I started last year because I wanted exercise that did not feel like exercise, and I like that every route is a small problem to solve." Now the follow-ups have real material — "What makes a route difficult?", "Have you climbed outdoors?" — and you already have things to say.
The pattern that works on every question: answer the actual question in one sentence, then add one sentence of reason, example or detail. Two sentences of specific content beat thirty-five seconds of general filler, and they hand the next question something to ask.
How Interactive Speaking is scored
Interactive Speaking feeds Speaking, and because you are answering spoken questions it draws on listening too. Each answer is rated, and the set is read as one connected exchange rather than as separate mini-answers.
The scorer looks at the same speaking dimensions as the other spoken tasks — fluency, range, accuracy, relevance — with one that the adaptive format stresses hardest: whether your answer actually responds to the question asked. Because the questions adapt to you, drifting onto a nearby question is easy, and it costs relevance marks.
| What the scorer looks at | Why it matters |
|---|---|
| Relevance to the question asked | Each answer must respond to this question, not a nearby one. The criterion the adaptive format stresses most. |
| Fluency and continuity | Sustained speech across the answer without long silences. |
| Range of grammar and vocabulary | Varied structures and precise words, produced under time pressure. |
| Coherence across the exchange | Answers that connect to what you said before read as a real conversation, not disconnected replies. |
| Intelligibility | Whether you can be followed easily. |
The mistakes that cost the most
The failures here are specific to the adaptive, one-hearing format.
Do this
- Answer the exact question first, in one clear sentence, before adding anything.
- Add one sentence of reason, example or detail — concrete, not general.
- Make your first answer specific. It sets up every follow-up that comes after it.
- Listen for the question word — why, how, when, would you — because a single hearing is all you get.
- Keep talking to the end of your time with content, not with restated filler.
Not this
- Do not prepare fixed answers. The questions are generated from your replies, so scripted content will not fit.
- Do not give vague first answers. "It is fun and relaxing" leaves the conversation nowhere to go.
- Do not answer a nearby question instead of the one asked — relevance gates the score, and the adaptive format makes this easy to slip into.
- Do not fall silent when you finish early. Add a reason or an example rather than stopping at ten seconds.
- Do not narrate confusion — "sorry, I did not understand" — when you can answer the part you did catch.
How to actually get better at Interactive Speaking
This task rewards one specific drill: producing a two-sentence answer to any question on demand, then extending it.
Drill the two-sentence shape
Take any question — "What did you do last weekend?" — and answer in exactly two sentences: the direct answer, then one reason or example. Repeat until the shape is automatic.
Make the first answer specific
Re-answer the same question with a concrete detail planted in it, and notice how a specific answer makes the follow-up easier to imagine and to answer.
Answer on one hearing
Have someone ask, or use text-to-speech, and answer with no replay. Listening once is half the task; practising with replays trains the wrong habit.
Record to the full time and fill it with content
Time yourself to about thirty-five seconds. The goal is content to the end — a second example, a contrast, a consequence — not restating your first sentence more slowly.
Run short question chains
Have a partner or a model ask a follow-up based on your last answer, three or four deep. That is the actual task, and it is very different from answering isolated prompts.
How much this one task moves your score
Interactive Speaking feeds Speaking, and so Conversation and Production, and it appears as its own section in every full test. It replaced Listen, Then Speak in the July 2025 update, and it carries more weight than the single task it replaced because it is a multi-question exchange rather than one answer.
It is the speaking task most worth practising, because the skill it isolates — understanding a question on one hearing and responding coherently on the spot — is the one that also carries real conversation, and the one no amount of prepared material can fake.
Interactive Speaking: questions people ask
What is Interactive Speaking on the Duolingo English Test?
It is a spoken-conversation task added in July 2025. An animated character asks a series of about six to seven questions, one at a time, and you record a spoken answer of up to roughly thirty-five seconds to each. Every question is generated from your previous answer, so it works like a real, adaptive conversation rather than a set of fixed prompts.
When was Interactive Speaking added to the DET?
1 July 2025. It replaced the Listen, Then Speak task as part of a wider update that also removed Read Aloud and expanded Interactive Listening. If a guide still lists Read Aloud or Listen, Then Speak as current tasks, it has not been updated for the current format.
How long do I get to answer each question?
Up to about thirty-five seconds per answer. There is a brief moment to gather your thoughts after each question, then a recording window. You need not use the whole time, but very short answers give the scorer little language to assess.
How many questions are there in Interactive Speaking?
Around six to seven linked questions in a single conversation. The exact number can vary because the test adapts, but it is a short, connected exchange rather than one isolated answer.
Can I hear each question more than once?
No. Each question plays once, with no transcript and no replay, so listening accurately the first time is part of the task. Listen for the question word and the key detail rather than trying to hold every word.
How is Interactive Speaking scored?
It feeds your Speaking subscore, and each answer is rated on relevance, fluency, grammar, vocabulary and how clearly you can be understood. The exchange is read as one connected conversation, and the single most important factor is whether each answer actually responds to the question asked.
Does Interactive Speaking replace Listen, Then Speak?
Yes. Interactive Speaking replaced Listen, Then Speak in the July 2025 update. The old task asked one question and took one answer; the new one is a multi-question conversation where each question builds on your last reply.
Can I prepare answers in advance?
Not fixed answers. Because each question is generated from what you just said, prepared content usually will not fit the question you are actually asked, and off-topic answers lose relevance marks. What you can prepare is a shape: answer the question directly in one sentence, then add one sentence of reason or example.
What makes a good first answer?
A specific one. A concrete first answer — a named hobby, a real reason, a particular example — gives the follow-up questions something to build on and gives you material to keep talking about. A vague first answer forces generic follow-ups and leaves you starting from nothing each time.
What if I do not understand a question?
Answer the part you did catch rather than narrating confusion. Because each question plays once, a partial but relevant answer scores better than silence or "sorry, I did not understand." Listen especially for the question word, which tells you what kind of answer is wanted.
Is Interactive Speaking hard?
It is one of the more demanding tasks, because it combines listening on one hearing with speaking on the spot and you cannot rely on prepared material. But it is very trainable: the two-sentence answer shape and practising short question chains cover most of what it tests.
Which subscores does Interactive Speaking affect?
It feeds Speaking directly, and through the DET's integrated subscores it also contributes to Conversation and Production. Because you are responding to spoken questions, your listening accuracy affects your answers too.
How is Interactive Speaking different from the Speaking Sample?
The Speaking Sample is one long, uninterrupted response to a single open prompt, recorded for institutions to review. Interactive Speaking is a back-and-forth of several short answers where each question depends on your last one. The Sample tests sustained monologue; Interactive Speaking tests responding in a live exchange.
How do I practise Interactive Speaking?
Drill a two-sentence answer to any question — direct answer, then one reason or example — until it is automatic, practise answering on a single hearing with no replays, and run short chains of follow-up questions so you get used to each one building on your last answer. Recording yourself and filling the full time with real content is the core exercise.
Do I need a webcam for Interactive Speaking?
A working microphone, yes — it is a spoken task. A webcam is required for the test overall (proctoring and the video the Speaking Sample produces), but the answers you give in Interactive Speaking are audio. Run the equipment check before you start so the microphone is confirmed.
Where Interactive Speaking sits in a full test
No task on the DET is scored in isolation. The adaptive engine is building one picture of your English out of every response, so it helps to know how much of the hour this particular task accounts for and which of the others are drawing on the same ability.
The table below is the whole test. Your current page is highlighted.
| Question type | Time | How often | Subscores it feeds |
|---|---|---|---|
| Read & Complete | 3 minutes | 3–6 times | Reading, Writing |
| Read & Select | 5 seconds | 15–18 times | Reading |
| Fill in the Blanks | 20 seconds | 6–9 times | Reading, Writing |
| Listen & Type | 1 minute | 6–9 times | Listening, Writing |
| Speak About the Photo | 90 seconds | at least once | Speaking |
| Interactive Speaking — you are here | 35 seconds each | 6–7 questions | Speaking, Listening |
| Write About the Photo | 1 minute | at least twice | Writing |
| Interactive Reading | 7 minutes | 2 sets | Reading |
| Interactive Listening | 6 minutes | 2 sets | Listening, Speaking |
| Speaking Sample | 3 minutes | once | Speaking |
| Writing Sample | 5 minutes | once | Writing |
Two things are worth reading off that table. The first is that the tasks feeding Speaking and Listening are not only this one — improvement transfers, so time spent here shows up elsewhere. The second is that the tasks taking the largest share of the hour are not the ones most people practise most.
Question types that use the same muscles
Speaking Sample
An extended spoken response to an open prompt. It is sent to institutions alongside your score.
Read the guide → SpeakingSpeak About the Photo
An image appears and you describe it aloud.
Read the guide → Listening + SpeakingInteractive Listening
A spoken conversation unfolds in turns. At each turn you choose the best reply, then afterwards answer comprehension questions and write a summary of what was said.
Read the guide →Start a practice test
An hour under real conditions, all eight subscores, and feedback on every written and spoken answer.
Take a test