The uncomfortable thing about writing an ElevenLabs review is that the product question is easy and the buying question is not. It makes the most natural synthetic speech available, most people agree on that, and cheaper rivals sound cheaper. Fine.
The hard parts are elsewhere: a credit system that measures characters while you think in minutes, overages billed per model rather than per plan, a free tier that is not a licence, and a clone button whose availability has nothing to do with whether you may legally press it. Pricing checked 21 August 2026.
What the plans cost
| Plan | Monthly credits | What actually changes |
|---|---|---|
| Free β $0 | 10,000 (~10 minutes of the higher-quality model) | No commercial licence, attribution required. A demo, not a plan |
| Starter β about $5β6 | 30,000 (~30 minutes) | Commercial rights begin here. Instant voice cloning unlocked |
| Creator β $22 | 100,000 (~100 minutes) | Professional voice cloning and higher-bitrate output. The real entry point if fidelity matters |
| Pro β $99 | 500,000 (~500 minutes) | Production-quality audio via API. Best per-minute value for a working business |
| Scale β around $299β330 | 2,000,000 | Multi-seat workspaces, low-latency real-time use |
| Business β roughly $990β1,320 | Higher | Volume production. Enterprise is quoted |
Annual billing saves around 17%, roughly two months. Where trackers disagree β Starter at $5 or $6, Scale at $299 or $330 β that is because promotional and regional rates differ and the page changes; check it rather than trusting any figure here.
Credits are characters, and every product spends them differently
This is the part that makes plan comparison useless until you understand it, and the part any ElevenLabs review has to get right. A credit is roughly one character of text-to-speech. But the same pool also pays for dubbing, transcription, sound effects, music and conversational agents β each at its own rate.
So “100,000 credits” is not a quantity of anything until you say what you are doing with it. A hundred thousand characters is about a hundred minutes of narration, or dramatically less dubbing, or a different number again for agent conversation minutes.
Then the second twist: overages are billed per model, not per plan. Reported rates in 2026 run around $0.05 per thousand characters for the fast model and about $0.10 for the higher-quality multilingual one, with transcription from roughly $0.22 an hour and dubbing anywhere from about $0.33 to $2.20 a minute. Two subscribers on the same tier can therefore have very different bills purely from model choice β the same pattern we found in coding tools in Cursor vs Claude Code vs GitHub Copilot.
The useful rule: when your overage spend reaches roughly 30β50% of the next tier’s price, upgrade. Past that point the flat plan is cheaper than paying per unit, and it stops the bill from being a surprise.
The ElevenLabs review nobody writes: the clone button is not a permission slip
Here is the thing that matters more than any pricing table, and that most coverage buries.
Paying for a plan unlocks voice cloning as a feature. It grants you nothing about the voice itself. ElevenLabs’ terms require that you own the voice or hold explicit consent from the person it belongs to, and a growing number of US states β reported at a dozen or more β now have voice-cloning or right-of-publicity statutes covering synthetic voice likeness.
That makes voice legally different from images in a way people consistently miss. With generated imagery the fight is mostly about training data and copyright. With voice, the person is the right. Somebody’s voice is an attribute of them, and using it without permission is a problem regardless of how the model was built or how good your subscription is.
- Your own voice: fine, and the best use of the feature.
- An employee’s or client’s voice: get written consent that names the uses, the duration and what happens when they leave. Verbal agreement is not evidence.
- A voice actor you hired: your recording contract almost certainly does not cover synthesis unless it says so. Renegotiate rather than assume.
- A public figure, or a voice that merely sounds like one: don’t. Soundalikes have been litigated for decades under publicity rights, long before AI existed.
- A deceased person: several states extend publicity rights after death. This is where estates get involved.
Ownership, and the thing ownership does not give you
On paid plans you own the audio you generate and can use it commercially, perpetually β which is genuinely more generous than several comparable tools. Two caveats belong next to that.
First, ElevenLabs retains a licence to use your content, including voice material, to improve its models. Read the current terms if that matters for client work under NDA.
Second, owning output is not the same as being able to protect it. Purely AI-generated material is not protected by US copyright, so your generated narration is yours to sell and not yours to stop others copying. That distinction is the same one we drew for images in Midjourney vs Adobe Firefly for commercial use, and it catches people out in exactly the same way.
The disclosure rule that just came into force
If any of your audience is in the EU, note that the AI Act’s transparency obligations under Article 50 applied from 2 August 2026, requiring AI-generated audio, image, video and text to be marked in a machine-readable way so it can be detected as synthetic. Systems already on the market pick this up from 2 December 2026.
The heavier high-risk obligations were deferred to December 2027 by the Digital Omnibus, but transparency was not among the things that moved β which is why a lot of AI-voice guidance written earlier in 2026 is now describing the wrong obligation. If you publish synthetic voice content to European audiences, this is live, not upcoming.
Where it is worth the premium β and where it is not
| Use | Verdict |
|---|---|
| Customer-facing narration, ads, brand voice | Worth it. Naturalness is the whole product and cheaper engines are audibly cheaper |
| Audiobooks and long-form narration | Worth it, but price the whole book in credits before starting |
| Localising existing video into other languages | Strong use case, and the most expensive per minute. Budget separately |
| Internal notifications, accessibility readouts, IVR prompts | Overkill. Commodity text-to-speech costs a fraction and nobody is listening for warmth |
| High-volume programmatic generation | Compare seriously. Rival APIs and open-weight models undercut it substantially on raw cost |
That last row deserves emphasis, because it is the honest limit of the product. You are paying a quality premium, not a capability monopoly. If your audio is functional rather than expressive, cheaper is the right answer β the same substitution logic as in AI tools that replace expensive software.
Real limitations
- Credits do not roll over, so a quiet month is money gone and a busy month is an overage.
- Pronunciation of names, acronyms and technical terms still needs manual correction. Long-form projects carry a real editing tail that never appears in the pricing.
- Emotional range is directed indirectly. You steer through punctuation and phrasing rather than instructing performance, which is fiddly on dramatic material.
- Consistency across sessions can drift on long projects. Generate related material in one pass where you can.
- The free tier is unusable commercially β no licence, attribution required. Treat it as a listening test only.
- Disclosure expectations are rising beyond the EU rule above. Several platforms and advertising standards bodies now expect synthetic voice to be identified.
The verdict
Every ElevenLabs review ends up at the same place on quality, so let the money decide. For anyone producing customer-facing audio in volume, this is the tool, and Pro at $99 is where the per-minute maths genuinely works. Creator at $22 is the right starting point for a solo creator who needs proper voice cloning; Starter at around $5 is enough to hold a commercial licence and test the workflow.
Do not buy it for functional audio, do not use the free tier for anything you publish, and do not clone a voice you do not own without written consent β that last one is the only mistake in this article that can cost you more than the subscription.
Comparisons with the cheaper alternatives are in ElevenLabs vs PlayHT vs Murf. For where audio fits in a small stack, see best AI tools for small business owners and best AI tools for freelancers; for whether a paid tier is needed at all, free vs paid AI tools; and if you plan to generate audio on a schedule, automating work with AI tools covers the metering discipline.
Frequently asked questions
Can I use ElevenLabs audio commercially?
On paid plans, yes β commercial rights start at the Starter tier and you own the generated audio. The free plan grants no commercial licence and requires attribution, so it cannot be used for monetised videos, client work or advertising.
How do ElevenLabs credits work?
One credit is roughly one character of text-to-speech, and the same monthly pool also pays for dubbing, transcription, sound effects, music and agent conversations at different rates. Roughly, 100,000 credits is about 100 minutes of narration β but far less dubbing.
Is it legal to clone someone’s voice?
Only with their permission. ElevenLabs’ terms require ownership or explicit consent, and a dozen or more US states have voice-cloning or right-of-publicity laws covering synthetic voice. A paid plan unlocks the feature, not the right. This is general information, not legal advice.
Why is my bill higher than my plan price?
Overages are charged per model rather than per plan, so the higher-quality multilingual model costs roughly twice the fast one per character, and dubbing costs far more per minute than narration. If overages regularly reach 30β50% of the next tier’s price, upgrading is cheaper.
Do I have to disclose AI-generated voice?
Increasingly, yes. The EU AI Act’s transparency requirements applied from 2 August 2026 and cover machine-readable marking of synthetic audio, with legacy systems included from December 2026. Platform and advertising rules elsewhere are moving the same way.
Is there a cheaper alternative that sounds as good?
Cheaper, yes β considerably. As good, generally not for expressive customer-facing work. For functional audio like notifications or accessibility readouts, commodity engines cost a fraction and the quality gap does not matter.
Sources
- ElevenLabs β official pricing
- ElevenLabs Help Centre β credits, cloning and plan terms
- US Copyright Office β Copyright and Artificial Intelligence
- Gibson Dunn β which EU AI Act obligations moved and which did not
Pricing and terms checked 21 August 2026; rates vary by region and promotion and the plan page changes often. Voice-likeness law differs by state and country β nothing here is legal advice.
1 comment