TOEFLIELTS
Bahasa Indonesia

TOEFL 2026 Speaking Feedback: What You Actually Get

This page documents the criteria LingoLeap scores a spoken answer against and the exact fields the report hands back, task by task. It is a specification, not a sales page: every criterion named below is the one the grader is instructed to apply, and every field named below is one the result screen can display.

By the LingoLeap Research Team · Published:

What feedback do you get on a TOEFL 2026 speaking answer?

You get four things. First, a transcript of what the speech recogniser heard. Second, machine measurements of the audio on a 0-100 scale — accuracy, fluency, prosody and completeness, plus a combined pronunciation figure. Third, a task score from 0 to 5 in half-point steps, broken into the criteria that belong to that task: accuracy, completeness and intelligibility for Listen and Repeat; topic relevance, elaboration, fluency, pronunciation and grammar & vocabulary for Take an Interview. Each criterion carries its own written feedback, not just a number. Fourth, task-specific repair material — for Listen and Repeat, the words you missed or changed and per-word pronunciation guidance with IPA and syllable stress; for Take an Interview, a model response, grammar corrections, vocabulary upgrades and the key points a strong answer would have covered.

The two 2026 speaking tasks

The 2026 TOEFL Speaking section has two task types and eleven scored questions in roughly eight minutes. There is no Read Aloud task and no independent or integrated speaking task of the pre-2026 kind, so anything written for those older tasks does not describe what you will be scored on.

Listen and Repeat

7 items. You hear one sentence and repeat it inside a response window of 8, 10 or 12 seconds depending on sentence length.

Take an Interview

4 questions. You answer each spoken question within 45 seconds. Early questions are personal and factual; later ones ask for an opinion with reasons.

The two tasks are scored against different criteria because they elicit different speech. A rubric written for one of them tells you nothing useful about the other.

What happens to your recording

A submitted answer goes through four stages, and the result screen fills in as each one lands. Knowing the order explains why the pronunciation numbers appear before the written feedback.

1. Audio assessment

The recording is transcribed and assessed phoneme by phoneme. This stage produces the transcript and the 0-100 accuracy, fluency, prosody and completeness figures, plus a per-word breakdown that marks each word as correct, mispronounced, omitted or inserted.

2. Rubric scoring

The transcript and the task prompt are scored against the rubric for that task type on a 0-5 scale in half-point steps, producing the overall task score, each criterion score and the written feedback attached to each criterion.

3. Grammar and language check

For the Interview task, the answer is checked for grammar errors and weak word choices, each returned as the original wording, a correction and an explanation of why the correction is better.

4. Expert feedback

For Listen and Repeat, every word scored below the pronunciation threshold gets its own coaching entry: what went wrong, the IPA, the syllable breakdown with stress, the sounds to focus on, how to practise it and similar words to practise with.

Listen and Repeat: criteria and fields

Listen and Repeat is scored on how faithfully you reproduced the sentence you heard. The scale runs 0-5 with half points allowed.

The three criteria scored

Accuracy

How closely the words you produced match the words in the prompt. Substituting a synonym, changing a tense marker or transposing two words all count against accuracy — the task asks for repetition, not paraphrase.

Completeness

How much of the sentence you produced at all. Repeating the opening and trailing off, or dropping a content word from a longer sentence, is a completeness problem rather than an accuracy one.

Intelligibility

Whether a listener could understand you without effort. Imprecise pronunciation that makes a content word ambiguous, running words together, or struggling over a phrase all reduce intelligibility.

What the 0-5 levels mean

  • 5 — an exact repetition of the prompt, fully intelligible.
  • 4 — the meaning of the prompt is captured but the repetition is not exact: a function word missing or changed, a content word replaced with a related one, a tense or number marker off, or two words transposed. Self-correction is allowed if the response is completed.
  • 3 — essentially a full sentence containing most of the content words, but the original meaning is not accurately captured. Intelligibility may occasionally make meaning hard to follow.
  • 2 — a large part of the prompt is missing and important meaning is lost. The response is not a self-standing sentence and intelligibility is low.
  • 1 — a few words only; recognisable as an attempt to repeat the prompt but mostly unintelligible.
  • 0 — no response, entirely unintelligible, no English, or content entirely unconnected to the prompt.

The fields you receive

  • scoreThe overall task score, 0-5 in half-point steps.
  • accuracy, completeness, intelligibilityOne score each on the same 0-5 scale, so you can see which of the three pulled the total down.
  • accuracy_feedback, completeness_feedback, intelligibility_feedbackWritten feedback for each criterion separately, rather than one undifferentiated comment.
  • missing_wordsThe list of words from the prompt that were missing or changed in what you said.
  • incorrect_wordsFor each substitution, three things: the word expected, what was heard instead, and the type of error.
  • pronunciation_suggestionsFor each word that scored below the threshold: its per-word accuracy score, the error type (mispronunciation, omission or insertion), what went wrong, the IPA, the syllables with stress marked, the sounds to focus on, step-by-step practice instructions, and similar words to drill.
  • specific_mistakesEach mistake as a type, exactly what happened, how it affected the score, and the fix.
  • overall_feedbackA summary naming your strongest area and the one thing to work on next.

Illustrative structure — not a real graded response

To make the shape concrete: if the prompt were "The committee approved the revised budget on Thursday" and the recogniser heard "The committee approved the revised budget Thursday", the report would carry an entry in missing_words for the dropped function word, and a pronunciation entry for any word whose per-word accuracy fell below the threshold, with its IPA and syllable stress. The scores below are deliberately left blank — no score on this page is taken from a real graded answer.

FieldWhat it would contain
missing_wordsthe function word that was dropped
incorrect_wordsexpected / heard / error type, for each substitution
pronunciation_suggestionsIPA, syllables with stress, sounds to focus on, practice steps, similar words
accuracy, completeness, intelligibilityone 0-5 score and one written comment each

Take an Interview: criteria and fields

The Interview task is scored on how well you addressed the question and how clearly you did it. The scale again runs 0-5 with half points allowed.

The five criteria scored

Topic relevance

Whether the answer is on topic and actually addresses the question that was asked, rather than a neighbouring one or a memorised script.

Elaboration

Whether the answer is developed. A response that is on topic but consists mostly of language recycled from the question scores low here even if it is perfectly pronounced.

Fluency

Conversational pace, and whether pauses are natural or frequent and lengthy enough to make the delivery choppy. Frequent filler words count here.

Pronunciation

Intelligibility, rhythm and intonation — whether they convey meaning or make the listener work.

Grammar & vocabulary

Range and accuracy, judged by whether they let you express precise meanings or visibly restrict what you can say.

What the 0-5 levels mean

  • 5 — fully successful: on topic and well elaborated, good conversational pace with natural pauses, easily intelligible with rhythm and intonation that convey meaning, and a range of accurate grammar and vocabulary.
  • 4 — generally successful: on topic and elaborated but perhaps lacking sentence-level connectors; pace generally good with some pausing; occasional words need minor effort; grammar and vocabulary adequate for general meanings most of the time.
  • 3 — partially successful: generally on topic but elaboration limited; frequent or lengthy pauses make the pace choppy and fillers are frequent; word-level pronunciation or stress sometimes affects intelligibility; limited range noticeably restricts precision.
  • 2 — mostly unsuccessful: minimally connected to the question, with little relevant elaboration or mainly language taken from the question; intended meaning often hard to discern; very limited range.
  • 1 — unsuccessful: only vaguely connected to the language of the question, mostly unintelligible, mainly isolated words or phrases.
  • 0 — no response, entirely unintelligible, no English, or content entirely unconnected to the prompt.

The fields you receive

  • scoreThe overall task score, 0-5 in half-point steps.
  • Five criterion scorestopic_relevance, elaboration, fluency, pronunciation and grammar_vocabulary, each on the 0-5 scale.
  • Five criterion commentsA separate written comment for each of the five, so a low total is traceable to the criterion that caused it.
  • model_responseA well-elaborated, fluent answer to the same question at the top of the scale, so you can compare structure rather than guess at it.
  • grammar_issuesEach error as the original wording, the correction, and why it is wrong.
  • vocabulary_suggestionsEach weak choice as the original word or phrase, a better alternative, and why the alternative is better.
  • fluency_tipsSpecific tips for improving speaking fluency on this kind of answer.
  • key_points_to_includeThe points a strong answer to this question would have covered — the fastest way to see what your answer left out.
  • overall_feedbackA summary with specific suggestions for improvement.

The pronunciation numbers

The 0-100 figures on the report are not the task score and do not convert to a band. They come from a phoneme-level assessment of the audio itself and exist to tell you which part of your delivery is weakest.

Accuracy

Averaged over the words in the recording: how closely each word's pronunciation matched the expected pronunciation.

Fluency

The proportion of the recording spent actually speaking rather than pausing, expressed as a percentage of elapsed time.

Prosody

Intonation, stress and rhythm, assessed separately from whether the individual sounds were right.

Completeness

The share of words produced with no detected error, counting omissions against you and ignoring insertions.

The combined pronunciation figure

The single pronunciation number is not an average. The four figures are sorted and the lowest is weighted at 0.4, the other three at 0.2 each. Your weakest dimension therefore moves it roughly twice as much as any other — which is the point: it is designed to stop a strong score on three dimensions from hiding a bad one.

Per-word detail

The same stage marks each word with an error type — none, mispronunciation, omission or insertion — and its own accuracy score. This is what the per-word pronunciation coaching on a Listen and Repeat report is built from.

From a 0-5 task score to a 1-6 band

A task is scored 0-5. The TOEFL 2026 score you care about is a 1-6 band. LingoLeap converts between them in two steps: the 0-5 task score maps to a comparable 0-30 section score, and the 0-30 score maps to the 1-6 band using the published speaking conversion. The table below is generated from that conversion at build time, so it cannot drift out of step with the app.

Task score (0-5)Comparable 0-30Band (1-6)CEFR
5.0306.0C2
4.5275.5C1
4.0255.0C1
3.5224.0B2
3.0204.0B2
2.5173.0B1
2.0142.5A2
1.5112.0A2
1.081.5A1
0.541.0A1
0.001.0A1

Read this before you use the table

Two limits. First, the table converts a single task score; a section band on the real test reflects all eleven scored questions, not one. Second, ETS has not published how raw speaking performance becomes a band score, so the 0-5 to 0-30 step is LingoLeap's own mapping and should be read as an estimate, not as an official equivalence. See our page on what ETS has not published about the 2026 TOEFL. /toefl/2026-open-questions

How to read your feedback

In the order that saves the most time.

  1. 1Read the transcript first. If the transcript is not what you said, the pronunciation numbers are telling you something about your articulation before any rubric has been applied.
  2. 2Find the lowest criterion score, not the total. The total tells you where you are; the criterion tells you what to change.
  3. 3Read that criterion's own comment. Every criterion carries its own written feedback — the overall summary is the least specific thing on the page.
  4. 4On Listen and Repeat, work missing_words before pronunciation. Dropped and substituted words move accuracy and completeness; a single mispronounced word usually moves less.
  5. 5On the Interview, read key_points_to_include against your own transcript. If the points are absent from your answer, the problem is elaboration, and no amount of pronunciation work will fix it.
  6. 6Use the model response for structure, not for memorising. A memorised answer is penalised under topic relevance when it drifts from the question actually asked.

What this page does not claim

Three things are deliberately absent.

No accuracy or calibration claim

We do not publish an agreement figure between these scores and official ETS scores, and nothing on this page should be read as one. Until that study exists and is published, treat every score here as an estimate.

No real graded sample

Every example on this page is a description of the report's structure. No score, transcript or piece of feedback here is taken from a real graded answer, because publishing one requires a test taker's consent.

No claim about official scoring internals

The score-level descriptions above are the criteria our grader applies. Where ETS has not published a detail — most importantly how a raw speaking performance becomes a band — we say so rather than filling the gap.

Frequently asked questions

What feedback do you get on a TOEFL 2026 speaking answer?
A transcript of what was heard, 0-100 audio measurements for accuracy, fluency, prosody and completeness plus a combined pronunciation figure, a 0-5 task score in half-point steps broken into the criteria for that task with written feedback on each, and task-specific repair material: missing and substituted words with per-word pronunciation coaching for Listen and Repeat, or a model response with grammar corrections, vocabulary upgrades and key points for Take an Interview.
Which criteria are scored on TOEFL 2026 Listen and Repeat?
Three: accuracy, how closely your words matched the prompt; completeness, how much of the sentence you produced; and intelligibility, whether a listener could understand you without effort. Each is scored 0-5 in half-point steps and carries its own written feedback.
Which criteria are scored on TOEFL 2026 Take an Interview?
Five: topic relevance, elaboration, fluency, pronunciation, and grammar & vocabulary. Each is scored 0-5 in half-point steps with its own written comment, so a low total can be traced to the criterion that caused it.
Is the pronunciation score out of 100 the same as my band score?
No. The 0-100 figures describe the audio and are diagnostic only. The task score is 0-5, and the band score is 1-6. The combined pronunciation figure weights your weakest of the four dimensions at 0.4 and the other three at 0.2 each, so it is not an average either.
How does a 0-5 task score become a 1-6 TOEFL band?
In two steps: the 0-5 score maps to a comparable 0-30 section score, and the 0-30 score maps to the 1-6 band using the speaking conversion. It is an estimate from one task, not a section score, and the first step is our own mapping because ETS has not published how raw speaking performance becomes a band.
Does the TOEFL 2026 speaking section still have a Read Aloud task?
No. The 2026 Speaking section has two task types, Listen and Repeat and Take an Interview, across 11 scored questions in about 8 minutes. Rubrics written for the pre-2026 independent and integrated speaking tasks do not describe what is scored now.
Do you publish how accurate this scoring is against official TOEFL scores?
Not yet. We have no published calibration study, and we do not claim one. This page documents which criteria are applied and which fields are returned, which is checkable; accuracy against official scores is a separate claim that would need its own evidence.

See the report on your own answer

Record a Listen and Repeat item or an Interview question in a realistic 2026 mock test and read the criteria on your own speech instead of on an example.

Take a free mock test

Related guides