Menu
See the rankings

AI voice · buying

ElevenLabs models compared: v3 vs Multilingual v2 vs Flash v2.5

There is no single best ElevenLabs model. The ones it sells do different jobs, one has just been retired, and the fast one is half the price of the rest. Here is what each costs and which to pick for narration, live phone calls, other languages or a tight budget.

By Voxrater · 7 min read · Updated 2026-07-11

Pricing and the model line-up were read from ElevenLabs' own pricing, models docs and v3 announcement on 18 June 2026, and both move, so treat the figures as dated rather than permanent. The ~75ms latency is ElevenLabs' own number, not a Voxrater measurement, and the quality calls here are our editorial read, not a blind listening test (we have not run one yet).

ElevenLabs sells four voice models, and the job decides which one. For narration and video, use v3 for the most expression or Multilingual v2 for proven range. For a live phone agent, use Flash v2.5: it answers in about 75ms and costs half the credits. Turbo is now retired.

Try ElevenLabs free

Paid link, we may earn a commission. How this works.

So that is the short answer. Here is the reasoning, the real cost per job, and where each model wins.

A quick plain-language note first. A voice model is the engine that turns your text into speech (text-to-speech, or TTS). ElevenLabs offers several, and newer or faster ones often cost less per character. You pay in credits, a prepaid balance you spend as you generate, and on the standard models one character of text equals one credit.

Four models, and one has just been retired

The line-up is not a simple ladder from worst to best. The models split into two camps. v3 and Multilingual v2 are the quality models, made for speech you record once and play many times. Flash v2.5 is the speed model, made to answer a real person in real time. They are different tools for different jobs, so “which is best” only makes sense once you say what you are doing.

The big change this year: Turbo v2.5 is gone. ElevenLabs’ own model list no longer shows Turbo and notes that the Flash models replaced it at the same low latency. So if you were weighing “v3 vs v2.5 vs v2”, the live three-way is really v3 vs Flash v2.5 vs Multilingual v2.

ElevenLabs' model documentation listing Eleven v3, Multilingual v2, Flash v2.5 and Flash v2, with no Turbo model present.
ElevenLabs' live model list. Turbo no longer appears, and Flash v2.5 carries the same languages as Multilingual v2 plus three more. · captured 2026-06-18
ModelBest forLanguagesCredits per characterLatency
Eleven v3Most expressive narration and dubbing70+standard ratenot built for real time
Multilingual v2Audiobooks and video narration291not built for real time
Flash v2.5Live phone agents and real time320.5~75ms (ElevenLabs’ figure)
Flash v2English-only real timeEnglish only~0.5~75ms (ElevenLabs’ figure)
Turbo v2.5Retired, use Flash v2.5n/an/an/a

ElevenLabs publishes a half-price-per-character rate for Flash but does not separately break out v3’s credit cost, so we treat v3 at the standard one-credit rate, the same as Multilingual v2. Read the latency figure as the vendor’s claim, not our measurement: we have not yet put these models through our own test calls.

What each one actually costs

Cost comes down to two numbers: how many credits a character burns, and how many credits your plan includes. Flash is the cheap one because it bills at about 0.5 credits per character against 1 credit on v3 and Multilingual v2. Same words, half the credits. In rough money, a thousand characters runs about $0.09 to $0.20 of your monthly credit depending on tier, and roughly half that on Flash.

Worked through two real jobs (assuming about 150 spoken words a minute, so roughly 9,000 characters in a ten-minute script):

JobRough sizev3 or Multilingual v2Flash v2.5
A 10-minute YouTube voiceover~9,000 characters~9,000 credits~4,500 credits
One audiobook chapter~50,000 characters~50,000 credits~25,000 credits
Creator plan ($22/mo) covers121,000 credits/mo~13 of those scripts~26 of those scripts

A live phone agent is metered differently. There you pay ElevenLabs’ agent rate of $0.08 a minute (the same on every plan), with the language model billed separately on top, so the model choice is about latency, not per-character cost. The free plan’s 10,000 credits cover about ten minutes of narration or fifteen minutes of agent time, with no commercial licence, so anything you sell or publish needs Starter at $6 a month or above.

ElevenLabs' pricing page showing Free, Starter, Creator, Pro, Scale and Business tiers with their monthly credit allowances.
The plan tiers and monthly credits. The credit allowance, not a per-seat fee, is what decides how much audio you can make. · captured 2026-06-18

Want this in your own numbers? Put your monthly volume into our cost calculator, or check the headline rate against every platform in the price index.

Which model for which job

This is where we take a position. The honest read, with quality treated as our editorial judgement rather than a scored listening test:

Your jobUse thisWhy
Live phone agent (sales, support)Flash v2.5Lowest latency, half the credits, the model ElevenLabs ships for agents
YouTube, ads, dubbingEleven v3The most expressive model, and listeners hear it repeatedly
Audiobooks and long narrationMultilingual v2Proven, lifelike, and built for long-form
A language v3 does not cover wellMultilingual v2 or Flash v2.5Check the per-model language list before you commit
Tightest budgetFlash v2.5Half the credits per character, so your plan stretches twice as far

For narration and video: v3, or Multilingual v2

If the audio is recorded once and played many times, quality is worth the slower render. v3 is our pick for expression. It is ElevenLabs’ newest model, now out of alpha and generally available, and it reads emotional cues and tone shifts more naturally than anything before it, with “audio tags” you can drop into the text to direct a laugh or a sigh. It also covers the widest language set, 70+ against Multilingual v2’s 29.

Reach for Multilingual v2 instead when you want a model with a long track record for audiobooks and steady long-form narration, or when its 29 languages already cover you. Both bill at the standard credit rate, so the choice is about sound, not cost.

Video thumbnail: Introducing Eleven v3: Our Most Expressive Text to Speech Model Introducing Eleven v3: Our Most Expressive Text to Speech Model ElevenLabs (official) ElevenLabs' own launch demo of v3's expressive range and audio tags. v3 is now generally available. Watch on YouTube ↗

For a live phone agent: Flash v2.5, no argument

A phone call is unforgiving: every extra hundred milliseconds of silence reads as a stall. Flash v2.5 is the only sensible choice here. ElevenLabs quotes about 75ms to first audio, recommends it for real-time and its Agents platform, and prices it at half the credits. Just as important, v3 cannot run in real time, by ElevenLabs’ own guidance, so do not be tempted to put the prettier model on a live line. Flash v2 is the same idea if you are English-only.

Video thumbnail: Introducing ElevenLabs Conversational AI 2.0 Introducing ElevenLabs Conversational AI 2.0 ElevenLabs (official) The agents platform that runs on the low-latency Flash model for real-time calls. Watch on YouTube ↗

The honest limits

Three things to keep in mind. The quality ranking above is our editorial read, not a blind listening test, so treat “most expressive” as a considered opinion until we publish measured results. The ~75ms latency is ElevenLabs’ figure, measured excluding your app and network, so your real-world number will be higher. And pricing plus the model list both move quickly, which is why every figure here is dated to 18 June 2026 and linked to its source.

For where ElevenLabs sits against the alternatives, see our ElevenLabs review, or put it head to head with Vapi or Retell. If voice quality is the whole point of your project, it is still the one to beat.

Try ElevenLabs free

Paid link, we may earn a commission. How this works.

Common questions

Which ElevenLabs model should I use for a live phone agent?
Flash v2.5. It is the one ElevenLabs recommends for real-time and its Agents platform, it answers in about 75ms by their measure, and it costs half the credits per character. v3 is not built for real-time, so do not reach for it on a phone line.See our ElevenLabs review
Is Eleven v3 worth paying for over Flash?
For pre-recorded narration, dubbing and video, yes: v3 is the most expressive model and the quality shows when a listener hears the same line many times. For a live agent, no: it cannot run in real time, so Flash v2.5 wins on the only thing that matters there, speed.
Which ElevenLabs model is cheapest?
Flash v2.5. It bills at about half a credit per character against one credit on v3 and Multilingual v2, so the same script costs roughly half and your monthly credits stretch twice as far. Use the calculator to put your own volume in.Open the cost calculator
What happened to Turbo v2.5?
It is deprecated. ElevenLabs now lists Turbo as replaced by the Flash models, which hit the same low latency, so there is no reason to start a new project on Turbo. If you are on it, Flash v2.5 is the like-for-like move.
Can I use Eleven v3 on the free plan?
No. v3 needs a paid plan. The free tier gives you 10,000 credits a month (about ten minutes of speech) on the standard models, with no commercial licence and attribution required, so anything you publish or sell needs Starter ($6/mo) or above.ElevenLabs pricing detail

Sources

Every figure above is dated and links to its primary source.

  1. ElevenLabs models docs: the live text-to-speech line-up (Eleven v3, Multilingual v2, Flash v2.5, Flash v2), per-model languages and character limits, Flash at ~75ms, and the recommendation to use Flash v2.5 for real-time agents. Turbo no longer listed. checked 2026-06-18
  2. ElevenLabs pricing: subscription tiers and monthly credits (Free 10k to Business 6M), 1 character = 1 credit on the standard models, Flash/Turbo at 0.5 to 1 credit per character, commercial licence from Starter ($6/mo). checked 2026-06-18
  3. ElevenLabs Agents pricing: $0.08 per minute base call rate, the same on every plan, with the language model billed separately, plus included agent minutes per tier. checked 2026-06-18
  4. ElevenLabs: Eleven v3 is now generally available (out of alpha), with a 68% drop in complex-text errors over the earlier release. checked 2026-06-18

Get the next piece

New analysis and dated test results land in the newsletter first. No spam.

Newsletter launching soon.