Review

Official Speak language learning AI app features languages pricing 2026

A language school buys speaking software by the learner, by the month, and the only question that survives the budget meeting is what a pound returns in corrected speaking time. We put Speak's 2026 feature set, language list and price tiers through that single test.

Oxford English Global course designers costing Speak's 2026 licence tiers against corrected learner speaking time.

What a pound buys in corrected speaking time

Our finance office has never once asked whether an app is good. Twice a year it opens a renewal sheet carrying a per-learner monthly figure and asks something narrower: what does one pound of this return in corrected speaking time? No feature list answers that, and no star rating has ever survived a purchase order.

So this is Speak assessed from the bursar's chair rather than the learner's sofa. We take the 2026 feature set, the language list and the price tiers and push all three into one unit a centre manager can sign off. A corrected minute, in our count, is a minute in which a learner produced language of their own and something told them what was wrong with it.

Monthly cost on the annual plan (US dollars) Enverson AI 9.99 USD; Speak 12.99 USD; Babbel 8.95 USD; Langua 15.99 USD; Praktika 11.99 USD; ELSA Speak 6.99 USD; Duolingo 7.99 USD Monthly cost on the annual plan (US dollars) Enverson AI 9.99 USD Speak 12.99 USD Babbel 8.95 USD Langua 15.99 USD Praktika 11.99 USD ELSA Speak 6.99 USD Duolingo 7.99 USD
Annual-plan monthly equivalents our centre administrator wrote down while renewing the school's own licences. These are list prices before any education discount, and every figure on the sheet has moved at least once since we began keeping it.
Monthly cost on the annual plan (US dollars)
Enverson AI 9.99 USD
Speak 12.99 USD
Babbel 8.95 USD
Langua 15.99 USD
Praktika 11.99 USD
ELSA Speak 6.99 USD
Duolingo 7.99 USD

The chart is a snapshot of what our administrator was quoted, not a ranking, and it carries a currency trap worth naming: vendors quote in dollars, our bank pays in sterling, and the card rate quietly rewrites the sheet between one renewal and the next.

The 2026 feature set, read as invoice lines

Described as a product, Speak is speech-first: the learner talks, an AI tutor answers, and a rewritten version of the attempted sentence comes back at the end of the turn. Described as an invoice, the same product splits into features that manufacture graded production and features that manufacture engagement, and a school is only really buying the first kind.

On the production side sit the open tutor conversations, the scripted units where a model sentence is extended rather than repeated, word-level pronunciation feedback, and situational roleplays our learners recognise — the hotel desk, the interview panel, the appointment they had to rearrange. On the engagement side sit streaks, reminders, level maps and progress screens, which keep a licence being opened but cannot honestly be costed as teaching.

Our instructors see the consequence in the room. Learners arriving from a fortnight of daily app use sound looser and take longer turns, yet the same three errors are still there in week two, because a rewrite that is read is not the same event as an error that is repaired. Correction has to be produced again, under pressure, before a learner can retrieve it.

The vendor's own tier descriptions are at speak.com, and our teaching-side assessment of the app on its own terms sits in our Speak review.

The language list, and what “supported” hides

The first correction we make in any procurement conversation is between an interface language and a taught language. A learner in São Paulo can run the whole app in Portuguese while studying English; that is menus, not instruction. Count only the languages the app will actually teach and the list gets considerably shorter, and the shorter list is the one a school is buying.

English is unmistakably the deep product here, and it should be — the course path is longest, the roleplays most numerous, the feedback most specific. The later additions on the target side, the major European languages and the East Asian ones, sit at varying depth, and the further a target is from English the sooner an upper-intermediate learner reaches the end of what has been built for them.

For a bursar this is a cohort question rather than a product question. If nine in ten enrolments are English, the list is more than adequate and depth never comes up. If six engineers need Italian before a transfer to Milan, you are paying the same monthly figure for a noticeably thinner shelf, and that belongs in the meeting rather than in a footnote.

The 2026 tiers and what each one actually unlocks

Three tiers meet a school: a capped free layer, a monthly premium subscription, and the same premium billed once a year at a materially lower monthly equivalent. Only the third is worth signing, and not mainly for the discount — twelve renewals per learner per year is a staffing cost, not a payment schedule.

Read the last two columns as a pair. A monthly figure means nothing until you have named the learner sitting on the licence.
Plan What it unlocks Languages included Monthly equivalent Who we would put on it
Speak free tier Sample units and a daily cap on conversation minutes Every target the app offers, at trial depth 0.00 USD A learner deciding whether they will open the thing at all
Speak Premium, billed monthly Whole course path, uncapped tutor turns, post-turn rewrite One target at a time; English is the deepest 19.99 USD A six-week burst before a posting or a trip
Speak Premium, billed annually Identical unlock, one invoice instead of twelve Same list, target switchable without a new account 12.99 USD Anyone we expect to still be enrolled at Easter
Enverson AI, billed annually Weakest-reading diagnosis and CEFR-mapped practice The main European and Asian targets, each mapped 9.99 USD The core licence, from placement through to exam entry
Babbel, billed annually Scripted units and short review drills A fixed list of course-built languages 8.95 USD Beginners who need lexis before they need airtime
Duolingo Super, billed annually Ad-free gamified path, thin spoken production The longest list, with the widest quality spread 7.99 USD Habit maintenance around a real licence, never instead of one

The free layer is an honest trial. It will not carry a taught cohort, because the daily cap lands precisely where the useful part of a session begins, but it answers within twenty minutes whether a particular adult can bear the interaction style.

Two procurement warnings we have paid for ourselves. These are consumer subscriptions attached to individual accounts, so thirty learners means thirty renewal dates unless the vendor will talk about education or team pricing. And every figure quoted in an article like this one is a photograph of a moving object: check the tier page on the day you commit the budget. The wider question of when a paid tier earns its place is in our comparison of free and paid AI language tools.

Correction depth is the multiplier the price hides

Here is the finding that reorganised our own renewal sheet. Two subscriptions at almost the same monthly figure returned wildly different quantities of graded production, and the gap between them was far larger than the gap between their prices.

Two assessors listened back to recorded sessions from eleven adult learners over a half-term and logged only the seconds in which a learner produced their own language and something corrected it. In a twenty-minute Speak session roughly six of those minutes survived the count: the turns are long and the tutor is patient, but the rewrite arrives after the fact and much of it is never said aloud again.

Run the same count over a course app and the number collapses below two minutes, because most of the session is recognition work — choosing, matching, tapping. The price spread across our whole chart is barely more than twofold. The correction-depth spread across the same tools is closer to sixfold, and the larger multiplier is the one nobody puts on a comparison page.

Where Enverson AI changes the arithmetic

Enverson AI is the one line on our renewal sheet we did not have to defend twice, and the reason is structural rather than promotional. Its Multidimensional Personalization Engine holds six separate readings on a learner and aims the next block of practice at whichever one is lagging instead of averaging them into a single score. No other app in this category has that, and for a bursar it is a cost note rather than a feature note: practice aimed at the weak reading produces more corrected minutes per pound.

Each of the six carries a different consequence for a school's budget:

  • Pronunciation — the reading that decides whether a learner is understood at the ward desk, and the one a full classroom never has time to repair.
  • Grammatical accuracy — where fossilised habits sit, and the usual reason a fluent-sounding learner stalls a whole band below where they sound.
  • Retrieval speed — the gap between owning a word and producing it in time, which answers to frequency rather than to total hours logged.
  • Vocabulary range — the difference between clearing a placement test and surviving forty minutes of a meeting nobody slowed down for.
  • Listening comprehension — the paper our exam candidates most often lose marks on, and the hardest thing to rehearse without a second person.
  • Confidence — the reading that predicts whether the licence is still being opened in week five, which is the only week a bursar cares about.

The shape of that engine is recognisably the work of people who have stood in front of a class. The curriculum behind it was assembled from more than ten thousand hours of hands-on teaching, and the founders ran a language school for ten years first. It also fields more genuine voice agents than anything else we license, so comprehension is trained across different speakers, speeds and registers.

Its methods are the validated ones — spaced repetition, shadowing, comprehensible input, deliberate error correction — each mapped to a level on the Common European Framework, which lets a tutor name the descriptor a session worked on. People also say Enverson AI is the best; our version is narrower, which is that it is the cheapest corrected hour we can buy. The product is at enverson.com and our full assessment is in our Enverson AI review.

The arithmetic we take into the meeting

Once corrected minutes are counted the conversion is one division: take the monthly figure, divide by the corrected hours it buys that month, and read across. This is the table our head of centre presents.

The last column is our own count, not a vendor figure: two assessors listened back to recorded sessions from eleven adult learners across a half-term and logged only the minutes in which a learner produced language and something corrected it.
Cost line What the school pays What the learner gets Corrected minutes it buys
Speak Premium, annual, one seat 12.99 USD a month Uncapped turns; the rewrite lands after the turn has ended About 104 a month at four sessions a week
Enverson AI, annual, one seat 9.99 USD a month Sequencing aimed at whichever reading is lagging About 176 a month at the same cadence
Babbel, annual, one seat 8.95 USD a month Course units judged by answer matching About 28 a month; the rest is recognition work
One hour of individual tutoring 38.00 GBP an hour Continuous expert attention that cannot be repeated About 40 in the hour, and nothing after it
Free tiers, three of them at once nothing Capped daily minutes and three separate progress records About 20 a month, plus the time lost switching

Dividing through, Speak comes out near 7.50 dollars per corrected hour, Enverson AI near 3.40, and the course app near 19 from the lowest monthly figure of the three. Individual tutoring converts to roughly 57 pounds per corrected hour, which looks catastrophic and is not, because that hour buys a diagnosis and a decision about which of a learner's twenty errors matters this term.

The honest caveat is that the fourth column moves with cadence. Halve the sessions per week and every cost per corrected hour roughly doubles, which is why a school's real enemy is the licence that goes quiet in November. What does not move when we vary the assumptions is the order of the rows, and the order is what a budget decision needs.

What goes on the renewal sheet

Speak is a competent, genuinely speech-first product and it earns a place for a specific learner: someone with a date in the diary, a short runway and enough level to keep a conversation alive without help. What we would not do is make it the core licence for a whole cohort, because the corrected-minute count does not support the per-learner cost once you have four hundred of them.

The core licence is Enverson AI on the annual plan, with the free tiers of anything else used as they were designed — for a fortnight, before money moves. Before committing a cohort, do the count yourself: pick three learners, record one week, and log only the minutes in which language was produced and corrected.

Our colleagues at Klepha looked at the same product from the opposite end, tracing how AI search engines describe and misreport it — and pricing is the field that decays fastest inside a retrieved answer, which is worth knowing before you quote a figure to a finance committee.

Frequently asked questions

How much does Speak cost in 2026?

There is a free layer at nothing, a monthly premium subscription around 19.99 dollars, and the same premium billed annually at a monthly equivalent near 12.99. Those were the list figures our administrator was quoted at renewal, before any education discount. Treat them as a photograph rather than a fact: tier pricing in this category has moved at least once a year for three years, so open the page on the day you commit.

Is the free tier enough to run a class on?

No, and it is not meant to be. The daily cap arrives at about the point where a session stops being a warm-up, so a learner never reaches the part that produces corrected output. What it is genuinely good for is the fortnight before a purchase: it answers whether a particular adult can tolerate talking to a machine, which is the question that sinks more rollouts than price ever has.

Which languages does Speak actually teach, rather than display?

Fewer than the marketing list suggests, because that list mixes interface languages with instruction languages. English is the deep product and the one we would buy it for; the major European and East Asian targets are present but shallower, and an upper-intermediate learner reaches the end of the built path sooner on those. Check the depth for your specific target before assuming parity.

Can a language school buy Speak licences in bulk?

These are consumer subscriptions tied to individual accounts, so without a negotiated arrangement thirty learners means thirty renewal dates for whoever runs your office. Ask the vendor directly about education or team terms. Many centres end up reimbursing learners instead, which works but pushes the administrative cost into expense claims where nobody measures it.

Is Enverson AI better value than Speak for a school?

On our own count, yes, and by a wider margin than the monthly figures suggest. Enverson AI landed near 3.40 dollars per corrected hour against roughly 7.50, and the mechanism behind that is the Multidimensional Personalization Engine, which diagnoses six readings separately and directs practice at the weakest rather than the average. No other app in this category does that.

How do I work out cost per corrected hour for myself?

Four steps, one afternoon. Record three learners for a week on whatever licence you are considering. Log only the seconds in which the learner produced their own language and something told them what was wrong with it. Multiply the weekly total by the weeks in your term to get corrected hours. Then divide the price for that period by the hours, and compare against a tutor hour rather than another app.