Skip links

AI Voice Generator: the real minutes behind the free plans

In short

Try our awesome AI for free

Nation AI
Ask me anything...

On the query “ai voice generator”, 16 of the 19 organic results are product pages. They all advertise something free, and almost none of them says how many minutes of audio that buys you.

Quotas are counted in characters, in credits, in minutes, in hours, in videos or in numbers of voices. Only one unit lets you compare: the minute of audio, and the conversion factor is published by the vendors themselves, at a thousand characters per minute.

The real wall is not the counter, it is the commercial licence, missing from the free column at most vendors. And cloning a real person’s voice to publish it now runs into a property right in that voice in a growing number of states.

16 / 19organic results that are product pages, not articles
~10 minwhat the ElevenLabs free plan is worth, as priced by ElevenLabs itself
6 unitsdifferent ways of counting the same thing across eleven pricing pages
2024the year a US state wrote “voice” into a property right and the FCC banned AI voices in robocalls

You type “ai voice generator”, you open the first five links, and you get five identical promises: free, realistic, no equipment. Not one of those five pages answers the only question that matters when you have a script to read out loud: how many minutes of audio can I produce before it stops, and am I allowed to use them?

We opened the pricing pages of the services that occupy this results page on 11 August 2026 and wrote down what they actually publish. It comes down to three findings: the units of measurement are never the same, the free quota converts into a common unit anyway, and it is the “commercial use” line that decides everything.

An AI voice generator is four different products

The keyword lumps together services that do not do the same job. Knowing which one you are after saves you from signing up in the wrong place.

Text to speech

You paste a text, you pick a voice from the catalogue, you get an audio file back. It is the most common product and the only one that is genuinely free somewhere.

For: video voice-over, e-learning module, phone greeting.

Voice cloning

You upload a recording and the service builds a synthetic voice that says whatever you want. Nearly always reserved for paid plans, and covered by law when the voice is not yours.

For: reusing your own voice without recording again.

Voice changer

You speak and the service swaps your timbre for another one while keeping your delivery. The result is livelier than a text read-out, but you have to act the take.

For: podcast, character work, anonymising.

Talking avatar

The voice comes bundled with a video and a synthetic presenter. The free quota is then counted in videos, not characters, and it is the tightest of the lot.

For: product walkthrough, internal message, training.

What Google serves on this query Of the 19 organic results measured on 11 August 2026, 16 are product pages, one is a Google Play listing, one is a YouTube video and one is a Reddit thread. A carousel of ten short videos sits at positions 11 to 20, and there is no “people also ask” block and not a single editorial article. In other words, Google has nothing to offer between the sales pitches. The top three are QuillBot, Canva and CapCut.

The only unit that lets you compare: the minute of audio

Voicebooking counts in characters. ElevenLabs and Fish Audio count in credits. TTSStudio counts in hours per month. Narakeet sells minutes. HeyGen hands out videos. Speechify counts nothing at all and advertises a number of voices. Six units for measuring exactly the same thing: the length of a sound file.

The good news is that the conversion does not have to be invented. The vendors publish it themselves, and you only have to read it.

Screenshot of the ElevenLabs FAQ: one text character equals one credit on V2 Multilingual models
ElevenLabs pricing page, FAQ entry “How do text characters and credits work?”, captured on 11 August 2026.

ElevenLabs spells it out: “For V2 Multilingual models, 1 text character equals 1 credit.” Its grid gives the free plan 10,000 credits a month. And its own comparison table turns that quota into a duration.

Screenshot of the ElevenLabs comparison table: the Free plan shows about 10 included minutes
The ElevenLabs “Compare plans” table: row “Minutes included”, column Free, “~10”. Captured on 11 August 2026.

Ten thousand characters, ten minutes. A thousand characters are therefore worth one minute of audio, and that ratio comes from the market leader on its own page. Two competitors publish the same kind of mapping and land in the same range: Fish Audio advertises 8,000 monthly credits and “Up to 7 minutes generation”, which works out at roughly 1,143 characters a minute; MiniMax Audio advertises 10,000 credits and “~12 mins of HD model”, roughly 833 characters a minute.

The conversion rule used in this article A thousand characters, spaces included, are worth one minute of audio. That is the figure ElevenLabs publishes. It slightly overstates the duration compared with Fish Audio and understates it compared with MiniMax: the spread between vendors runs from 833 to 1,143 characters a minute, about 15 % either side. For a quota calculation, that margin changes nothing.

One detail is worth flagging in passing, because it appears on no pricing page: the unit itself depends on the language. MiniMax says so in its FAQ.

Screenshot of the MiniMax Audio FAQ: one English letter equals one credit, one Chinese or Japanese character equals two
MiniMax Audio: “one English letter equals one credit, while one Chinese or Japanese character equals two credits”. Captured on 11 August 2026.

Translated: depending on the language, the same text does not burn the same number of credits, and English is the cheapest case documented. If your script is in Spanish, German or Japanese, expect the counter to move faster than this article predicts.

Count your script before you sign up

Paste your text below. The counter gives you the number of characters, the matching audio duration, what that duration would cost at the per-minute rate Narakeet publishes, and above all how many times that script fits inside each free quota we measured. Everything is computed in your browser, nothing is sent anywhere.

Voice script counter

Characters, words, audio duration and how many generations each free plan allows.

The text stays in your browser. Spaces and punctuation count towards the vendors’ quotas, so they are counted here too.

Characters

1,079

Words

189

Estimated length

1.1min

At €0.20 a minute

0.22

At this length the plans capped at a thousand characters are already behind you. You need a plan that publishes at least 8,000 credits a month, otherwise you will generate your text in several pieces and stitch them back together.

How many times your script fits inside each free plan

Service What the free plan publishes In characters a month Your script fits
Voicebooking 1,000 characters, 1 project, 3 downloads 1,000 0 times
HeyGen 3 one-minute videos a month 3,000 2 times
Fish Audio 8,000 credits, “Up to 7 minutes generation” 8,000 7 times
ElevenLabs 10,000 credits, “~10” minutes included 10,000 9 times
LuvVoice 10,000 characters a month, 3,000 per conversion 10,000 9 times
MiniMax Audio 10,000 credits, “~12 mins of HD model” 10,000 9 times
TTSStudio no zero-dollar plan on the grid 0 0 times
ReadSpeaker demo with no signup, no counter on display no published cap as many times as you like, but never for commercial use

The credits of ElevenLabs, Fish Audio and MiniMax are counted here at one character per credit for text to speech, which is what each of them documents on its own page. The 3,000 characters for HeyGen are a conversion of its 3 one-minute videos, at a thousand characters a minute. Figures captured on 11 August 2026.

What the length of your script changes

Length What it means
Under a minute, below 1,000 characters A script this short goes through anywhere, including the stingiest plan on the list. It is the format of a video intro, a phone greeting or an audio post.
1 to 3 minutes, 1,000 to 2,999 characters At this length the plans capped at a thousand characters are already behind you. You need a plan that publishes at least 8,000 credits a month, otherwise you will generate your text in several pieces and stitch them back together.
3 to 10 minutes, 3,000 to 9,999 characters A single generation eats most of a free monthly quota. Plan on finding the right voice first time, or pay for the month you produce in.
Over 10 minutes, 10,000 characters and up No free plan in this table holds a script of that length in one month. At that volume, the per-minute rate Narakeet publishes becomes the right benchmark.

Script over the quota? Nation AI does not speak the text, it writes it and trims it until it fits.

Write my script

Hello and welcome. In two minutes I will show you how to prepare a voice-over script that fits inside a free quota.
First point: write short sentences. One idea per sentence, a period at the end. Speech synthesis breathes where you put the punctuation, and nowhere else.
Second point: spell out your numbers. Twenty-four reads better than the digits, and one thousand five hundred avoids a hesitant delivery. Do the same for the acronyms you want to hear letter by letter.
Third point: mark your pauses with commas and periods, not with line breaks. The engines ignore layout, they read punctuation.
Fourth point: read the text out loud before you generate it. You will hear the repetitions, the awkward transitions and the sentences that run too long, the ones that force the voice to break in the wrong place.
Last point: count your characters before you paste. A free quota of ten thousand characters is about ten minutes of audio a month, and a single retake costs you as much as the first version.
That is it. Now replace this text with your own and watch the counters move.

What each free plan actually gives you

Here is the full survey, service by service, with its position on the English results page. Anything in quotation marks is copied from the official pages as they render in English on 11 August 2026.

Service (US rank) Unit shown What the free plan gives Download Commercial use when free
QuillBot (1) none “completely free of charge”, no figure anywhere yes, button on the tool not stated
Canva (2) none no figure on the tool page not measured not measured
ElevenLabs (4) credits 10,000 credits, “~10” minutes yes no, listed on the $6 Starter plan
Voicebooking (6) characters 1,000 characters, 1 project 3 downloads yes, but the wording only exists on the French page
MiniMax Audio (10) credits 10,000 credits, “~12 mins of HD model” yes yes, “for songs created during free period”
Narakeet (13) minutes bought no zero-cost plan, 30 minutes for €6 yes not applicable
ReadSpeaker (outside the top 20) none demo with no signup, “unlimited voice samples” yes “strictly forbidden”
LuvVoice (outside the top 20) characters 10,000 characters a month, 3,000 per conversion MP3, files kept 72 hours no, listed on the $7.75 Plus plan
Fish Audio (outside the top 20) credits 8,000 credits, 7 minutes, 500 characters per generation yes no, struck through on the free tier
TTSStudio (outside the top 20) hours a month no zero-dollar plan, entry tier $9.99 for 10 h unlimited on paid plans not applicable
Speechify (outside the top 20) number of voices “10 robotic sounding voices”, natural voices at $29 a month not stated natural voices on the paid plan
HeyGen (outside the top 20) videos 3 one-minute videos a month with a watermark watermark removed on the $29 plan

Voicebooking, the tightest quota and the widest licence

The sixth result gives you a thousand characters, so one minute of audio, a single project and three downloads. The last two lines carry a red cross on its own grid: they are limitations, not features.

Screenshot of the Voicebooking grid: the Freemium plan is limited to 1,000 characters, 3 downloads and 1 project
Voicebooking, Freemium column: “1000 character limit”, “Limited to 3 downloads”, “Limited to 1 project”. Captured on 11 August 2026.

In exchange, it is the only one in the panel that writes the exact opposite of ReadSpeaker on rights, and it writes it in one language only. The French version of the same page carries a FAQ entry asking whether the generator comes with limited usage rights, and answers that it does not: you may use it on any means of communication, «au niveau mondial, à perpétuité», worldwide and in perpetuity. That FAQ block does not exist on the English page, which stops at the grid. One minute of audio, but a minute you are allowed to publish.

ReadSpeaker, unlimited and forbidden

ReadSpeaker offers a demo with no signup. Its help page confirms it: “Yes, our text to speech demo is completely free with no signup required. You can test unlimited voice samples, try different languages, and download audio files.” No counter, no credits, no form. And, two blocks further down the same page, the line that cancels all of it.

Screenshot of ReadSpeaker: commercial use of the generated sound files is strictly forbidden
ReadSpeaker, under the demo player: “For demo and evaluation purposes only; commercial use of generated sound files is strictly forbidden.” Captured on 11 August 2026.

Two services in the same panel therefore say the exact opposite of each other. One gives you a minute you can use anywhere, the other gives you unlimited time you can use nowhere, the moment money is involved.

ElevenLabs, the benchmark, but not for publishing

Ten thousand credits, ten minutes, 128 kbps audio. The ElevenLabs free plan is comfortable for auditioning voices. The “Commercial License” line only shows up from the Starter plan at $6 a month: it sits in the “Everything in free, plus” block, which is a roundabout way of saying it is missing below. Instant voice cloning lives in the same place.

Fish Audio, seven minutes under a promise of unlimited

The Fish Audio pricing page opens with “Unleash your creativity with unlimited generations”. The free tier, right underneath, shows 8,000 monthly credits and “Up to 7 minutes generation”, with a cap of 500 characters per generation. Two lines in that same block are struck through: enhanced voice cloning and commercial use. Its FAQ is blunter still: “Free plan users can only use generated content for personal, non-commercial projects.”

Screenshot of the Fish Audio free tier: 8,000 credits, 7 minutes, commercial use struck through
Fish Audio, free tier: 8,000 credits, 7 minutes, 500 characters per generation, “Commercial use” struck through. Prices are shown in euros even on the English page. Captured on 11 August 2026.

The per-generation cap deserves a word. With 500 characters maximum in one go, a two-minute script has to be cut into five pieces, generated separately, then reassembled in an editor. The monthly quota is not the only lock: the size of a single request is another one, and the grids file it under small print.

LuvVoice, the most generous on the file itself

Ten thousand characters a month, three thousand per conversion, more than 200 voices, MP3 download and a custom voice with 500 dedicated credits. On features alone it is the most complete free plan in the panel. The “Unlimited commercial rights” line only appears from the Plus plan at $7.75 a month, and the free card states that files are kept for 72 hours.

Screenshot of the LuvVoice table: 10,000 monthly characters and 3,000 characters per conversion on the free plan
LuvVoice, “Compare plans” table, Free column: 10,000 monthly characters, 3,000 per conversion, 1 custom voice. Captured on 11 August 2026.

Speechify, a free generator with no natural voice

The Speechify page that ranks on this query is headed “FREE AI VOICE GENERATOR”. Its pricing grid describes the free plan in two lines flagged with an orange exclamation mark: “10 robotic sounding voices” and “Text to speech features only”. The “1000+ high quality, natural voices” sit in the Premium column, at $29 a month.

Screenshot of the Speechify pricing page: the free plan gives 10 robotic voices, natural voices cost 29 dollars a month
Speechify pricing page: the Free plan is flagged “10 robotic sounding voices”, the natural voices are in the $29 a month column. Captured on 11 August 2026.

It is the widest gap in the panel between the headline of the product page and the grid published by the same vendor. The service is genuinely free, just not for what the query is looking for: a voice that does not sound like a machine.

HeyGen, three minutes a month and a watermark

HeyGen does not sell voice, it sells video with a synthetic presenter. Its free plan is therefore counted in videos: “3 videos per month”, “Videos up to 1 min”. Three minutes a month is the tightest quota in the panel after Voicebooking. Two interesting lines sit in the paid column at $29: “Watermark removal” and “Voice Cloning”. Worth noting for anyone comparing across borders: the same Creator plan is priced at 25 € on the French and Spanish versions of the page, so this is a different grid rather than a conversion.

Screenshot of the HeyGen pricing page: three one-minute videos a month on the free plan, watermark removal on the paid plan
HeyGen, Free column: 3 videos a month, one minute maximum. Watermark removal and voice cloning are listed on the Creator plan. Captured on 11 August 2026.

The ones with no zero-cost plan at all

Two services in this survey publish no free offer whatsoever, and that is useful information in itself: it saves you a pointless signup.

TTSStudio, which does not make the English top 20 but ranks in France and Spain, lines up three paid plans counted in hours of generation a month: 10 h for $9.99, 50 h for $29.99, 170 h for $99.99. No zero-dollar column. The page shows a one-hour countdown under each plan.

Screenshot of the TTSStudio pricing page: three paid plans counted in hours per month, no free plan
TTSStudio pricing page: three plans, none at $0, and a unit expressed in hours per month. Captured on 11 August 2026.

Narakeet, thirteenth, does not sell a subscription at all: “The plans are one-time payments to purchase additional capacity, based on the duration of audio/video materials you produce or transcribe, not recurring subscriptions.” Its grid is the most readable of the lot, because it states the price of a minute outright: €0.20 for 30 minutes, €0.15 for 300, €0.10 for 1,000, €0.08 for 2,500, €0.05 for 10,000. Note that Narakeet publishes its grid in euros, including on its English pages.

Screenshot of the Narakeet pricing page: one-time purchases from 30 minutes at 0.20 euro per minute
Narakeet publishes its price per minute of audio, from €0.20 down to €0.05 depending on the volume bought. It is the benchmark used by the counter above. Captured on 11 August 2026.

It is also the only grid in the panel that lets you put a price on a free quota. The ten minutes ElevenLabs gives away are worth €2 at Narakeet’s entry rate. The most generous free month on the voice market is worth roughly the price of a coffee.

The ones that publish no figure at all

QuillBot takes the first position with a page that promises “Use completely free of charge, with zero hidden fees and no sign-up required”, offers a Download button, and never states a limit. Its own FAQ, asked whether the tool is free, answers that the generator “does not cost money to use” and leaves it there. The interface counts your words as you type, which suggests there is a cap somewhere, but nothing on the page says what it is.

Screenshot of the QuillBot AI voice generator FAQ: the answer on price gives no figure and no limit
QuillBot, AI voice generator FAQ: the answer to “Is an AI voice generator free to use?” contains no number and no cap. Captured on 11 August 2026.

Canva, second, does the same with a very polished tool page that explains the process in three steps and never mentions a quota, a credit or a maximum duration: its voice generator draws on the account’s shared AI allowance, whose figure lives elsewhere, on the subscription pages. CapCut, third, goes further still: the word “free” comes back twenty-five times on its page, FAQ included, and no number ever follows it. Adobe Firefly could not be measured at all: the site refuses automated rendering, and we would rather write “not measured” than copy a figure found somewhere else.

The wall is not the counter, it is the licence

Illustration: a struck-through download button with a padlock next to a stamped sound wave
On most free AI voice plans, the first thing that blocks you is not the quota, it is the “commercial use” line.

Go back to the last column of the big table. Of the eight services that publish a usable free offer, only three allow commercial use without paying: Voicebooking in its FAQ, MiniMax for tracks created during the free period and, with an important caveat, HeyGen provided you accept the watermark. At ElevenLabs, LuvVoice and Fish Audio, the commercial licence is a paid line. At ReadSpeaker it is explicitly forbidden.

A monetised channel is commercial use The question comes up constantly and the answer is nearly always yes. A YouTube video with ads, a sales funnel, a company post, a paid course: all of that falls under “commercial use” as these services define it. Generating the voice-over of a monetised video on a plan that forbids it puts you in breach with the vendor, whatever your view count.

Three more locks hide in the same grids, and none of them is ever advertised.

  • The per-generation cap. Fish Audio limits each request to 500 characters, LuvVoice to 3,000 on the free plan. A long script has to be cut up and glued back together.
  • The retention window. LuvVoice keeps files for 72 hours. If you do not download straight away you have to regenerate, and regenerating burns the quota a second time.
  • The number of downloads. Voicebooking grants three, which is not the same as three generations: you can listen as much as you like, you only leave with three files.

Cloning someone’s voice: what US law says

Illustration: a profile card whose voice is protected by a shield and a padlock
A real person’s voice is not free raw material: since 2024 at least one state treats it as property and the FCC treats a cloned voice in a robocall as illegal.

Among Google’s own completions for this query sits “ai voice generator celebrity”. It is one of the most frequent searches in the cluster, and not one page in the top 20 explains what you are risking. There is no single federal statute on this, but two 2024 texts changed the picture and both are short.

The first is a state law. Tennessee’s Public Chapter 588, better known as the ELVIS Act, renamed the Personal Rights Protection Act of 1984 the “Ensuring Likeness, Voice, and Image Security Act of 2024” and added a definition that leaves no room for argument:

“Voice” means a sound in a medium that is readily identifiable and attributable to a particular individual, regardless of whether the sound contains the actual voice or a simulation of the voice of the individual

The same act rewrites the core rule as “Every individual has a property right in the use of that individual’s name, photograph, voice, or likeness in any medium in any manner”, and creates civil liability for anyone who “publishes, performs, distributes, transmits, or otherwise makes available to the public an individual’s voice or likeness” without authorisation. It took effect on 1 July 2024.

Screenshot of Tennessee Public Chapter 588: the ELVIS Act defines voice and creates a property right in it
Tennessee Public Chapter 588, House Bill 2091, sections 1 to 4. Consulted on the Tennessee Secretary of State website on 11 August 2026.

Three practical consequences, as general information and not as legal advice.

  1. Consent is the first condition. Cloning your own voice raises none of this. Cloning a relative’s, a colleague’s, a performer’s or a politician’s does, even with no intent to harm.
  2. The right survives the person and covers the tool. The Tennessee text also makes it actionable to distribute software whose primary purpose is producing a particular identifiable individual’s voice without authorisation.
  3. Where you are matters. Right of publicity is state law and it varies: California, New York and Tennessee are the strictest, some states have no statute at all. The rule of thumb that travels is simple: no consent, no publication.

The second text is federal, and narrower. On 8 February 2024 the FCC adopted a Declaratory Ruling recognising that calls made with AI-generated voices are “artificial” under the Telephone Consumer Protection Act, which, in its own words, “makes voice cloning technology used in common robocall scams targeting consumers illegal”. If you distribute in the European Union, add Article 50 of Regulation (EU) 2024/1689, applicable since 2 August 2026, which requires providers of systems generating synthetic audio to mark their output in a machine-readable format and requires whoever publishes a deep fake to disclose that it was artificially generated.

Marking is the vendor’s job, disclosure is yours The technical watermark is on the company that builds the generator. Saying so is on you. In practice you do not have to produce the machine-readable signal, but you do have to write somewhere that the voice in your video is synthetic as soon as it imitates a real person. One line in the description is enough, and it is the cheapest evidence of good faith you will ever produce.

The subject has a darker side, the cloned voice used to deceive. We covered that separately in our investigation into the AI voice phone scam, where we show how many seconds of recording it really takes to clone a voice.

Getting a voice that does not sound like a robot

Once the service is picked, the quality of the result depends less on the engine than on the text you feed it. The five settings below cost no credits at all and change everything.

  1. Write short sentences. Speech synthesis breathes at the punctuation and nowhere else. A forty-word sentence with no comma will be read in one go, without a breath.
  2. Spell out your numbers. “One thousand five hundred” is always pronounced correctly, “1,500” depends on the engine. Same for acronyms you want spelled out and for units.
  3. Do not rely on layout. Line breaks and bullets are ignored. A pause is requested with a period, a colon or a comma.
  4. Choose the voice before writing the long version. Try three voices on two sentences, not on the whole script: it is the best way not to burn a monthly quota on auditions.
  5. Read it out loud before generating. You will hear the repetitions, the awkward joins and the sentences that run too long. What you cannot say yourself, the machine will not say better.

Before

Our $29.90/mo plan includes 3 modules (CRM, support, HR) available 24/7 from any device, with no commitment and a 14-day trial.

After

Our plan costs twenty-nine dollars and ninety cents a month. It includes three modules: customer relations, support, and human resources. Everything is available at any time, from any device. No commitment, and a fourteen-day trial.

The version on the right has more characters, so it costs more credits. That is the real trade-off of this market: a text that reads well out loud is a longer text, and a longer text costs more. The right strategy is not to write short, it is to write dense and then cut everything that cannot be heard.

The mistakes that cost you a quota

  • Generating the whole script to audition a voice. Two sentences are enough to judge a timbre. You test a voice catalogue on a sample, not on the production.
  • Forgetting to download. At LuvVoice the file disappears after 72 hours and the second generation goes through the till again.
  • Pasting text with markup. Markdown asterisks, bullets and headings sometimes end up read out loud. Clean before you paste.
  • Taking the homepage promise for the grid. “Unlimited”, “free” and “no signup” are landing-page arguments. The truth is in the left-hand column of the pricing table.
  • Confusing free with royalty-free. Those are two separate questions, and the second one decides what you will be allowed to publish.

Nation AI does not record the voice, it writes the text the voice will read

Let us be straight about it: Nation AI does not produce audio files. It is not a voice generator. What the counter above shows, though, is where the game is really played on this subject: the quota is spent in characters. A badly calibrated script means a blown quota, a retake, and a second trip to the till.

Screenshot of Nation AI: the assistant writes and shortens the text of a voice-over script
Nation AI, testable with no signup: it writes the script, fits it to a target duration and strips out what reads badly. Captured on 11 August 2026.

That is exactly the job you can hand it: writing a voice-over script calibrated to a duration, trimming it to an exact character count, replacing digits with words, breaking up sentences that run too long. For plain rewriting of an existing text, our rephrasing tool goes straight to the point.

A script that fits the quota

Nation AI writes, trims and cleans the text your voice generator is going to read. Free to try, no signup.

Prepare my voice-over script

Frequently asked questions

What is the best free AI voice generator?

It depends what you plan to do with the file. For commercial publishing without paying, Voicebooking is the only one in the panel that explicitly allows it, but it only gives one minute. For trial volume, ElevenLabs and LuvVoice give ten minutes a month, with no commercial licence. To listen without limits and publish nothing, the ReadSpeaker demo has no counter at all.

Can you generate an AI voice without signing up?

Yes, but rarely with a download. ReadSpeaker offers a demo that is “completely free with no signup required” and lets you download audio files, with all commercial use excluded. QuillBot also promises “no sign-up required” and offers a download button, without publishing any limit. Elsewhere, the no-account trial stops at the moment you try to export.

How many characters make one minute of audio?

About a thousand. That is the ratio ElevenLabs publishes, whose free plan pairs 10,000 credits with “~10” included minutes, given that one character equals one credit there. Fish Audio is a little more generous, at roughly 1,143 characters a minute, MiniMax a little tighter, at roughly 833.

How do you get the voice as an MP3?

Check the download line of the grid before signing up, because that is where most free offers stop. LuvVoice advertises MP3 from its free plan, Voicebooking caps you at three downloads, HeyGen exports with a watermark until you move to a paid plan.

Can you make a celebrity voice say whatever you want?

No. In Tennessee, the ELVIS Act of 2024 gives every individual a property right in their voice, explicitly including a simulation of it, and makes publishing an unauthorised voice actionable. Other states protect the same thing through their right of publicity, with different wording and different limits. This is general information, not legal advice.

Can a free AI voice be used on a monetised YouTube channel?

A monetised channel is commercial use. Of the eight free offers measured, only three allow it, and ReadSpeaker forbids it in writing. Check the “commercial use” line of the free column before you publish, not after.

Do you have to disclose that a voice was AI-generated?

As soon as the voice imitates a real person, yes, and for two reasons. Disclosure is what separates a synthetic performance from an attempt to pass it off as authentic, which is what right-of-publicity claims turn on. And if you distribute in the European Union, Article 50 of the EU AI Act, applicable since 2 August 2026, requires whoever publishes a deep fake to disclose that the content was artificially generated.

Why do non-English voices sound worse?

Voice catalogues are larger in English and multilingual models are often trained on English first. Two settings make up a lot of ground: spelling out numbers and cutting long sentences. Try several voices on the same sentence too, because the gap between two voices from the same vendor is wider than the gap between vendors.

What to remember

The AI voice market sells free without ever saying how much. Convert everything into the same unit, the minute of audio, and the ranking sorts itself out: ten minutes a month at the leader, seven at its direct rival, one at the tightest of the panel, zero at two services that rank in Europe. Then look at the line below, the licence line, because that is the one that decides whether the file is any use to you.

And keep in mind that a person’s voice is not raw material. Since 2024 it is named property in Tennessee law and a cloned voice in a robocall is illegal under FCC rules. Saying that it is AI is usually enough to stay on the right side.

For the rest of the production chain, we measured the real price of a second of AI video the same way, along with the quotas and rights of AI music generators and what actually blocks you on song generators with lyrics. For the editing that comes after the voice-over, see our pick of free AI video editing tools.

Sources. Pricing and product pages consulted on 11 August 2026: QuillBot, Voicebooking, ElevenLabs, Canva, ReadSpeaker, LuvVoice, MiniMax Audio, Fish Audio, TTSStudio, Speechify, HeyGen, Narakeet, CapCut. Tennessee Public Chapter 588 (House Bill 2091), Ensuring Likeness, Voice, and Image Security Act of 2024, effective 1 July 2024. FCC Declaratory Ruling of 8 February 2024, news release DOC-400393A1. Regulation (EU) 2024/1689, Articles 50 and 113, on EUR-Lex. Search results positions measured on 11 August 2026.