# Naadly > Naadly is a self-hosted text-to-speech, voice-cloning and dubbing service: 1,481 licensed voices across 11 catalogue languages, 23 Indian languages on a dedicated engine, and 35 translation languages. It runs open models on its own containers in Paris, so it meters minutes of audio rather than characters. ## Facts - Voices: 1,481, each labelled for gender, accent, pace and measured quality, each cleared for commercial use with its licence printed. - Engines: Piper, Kokoro, piper-plus, Indic-Mio; cloning uses OpenVoice V2; transcription uses faster-whisper. - Output formats: mp3, wav, ogg, flac, u-law, plus SRT subtitles timed from the render. - Free plan: 30 minutes of audio a month, no card, commercial use permitted. - Paid plans: $15/month for 6,000 minutes, $39/month for uncapped standard-language speech, $99/month for white-label resale. - Every render carries an inaudible AudioSeal watermark, a synthetic-media response header and a provenance record. - Voice cloning requires a spoken consent recording from the speaker; the consent record survives account deletion. - Customer scripts, uploads and reference audio are never used to train models. - Subscriptions are non-refundable outside the exceptions named at https://naadly.com/refunds. ## Primary pages - [Voices](https://naadly.com/voices): the full catalogue with filters and free previews. - [Pricing](https://naadly.com/pricing): plans, limits and a per-hour comparison with dated sources. - [API](https://naadly.com/guides/text-to-speech-api): HTTP reference, streaming and webhooks. - [Voice licences](https://naadly.com/licences): model, licence and attribution per voice. - [Legal](https://naadly.com/legal): privacy, terms, refunds, acceptable use, DPA, subprocessors, cookies, security. - [Resources](https://naadly.com/resources): guides, comparisons, use cases, answers and notes. ## Guides - [How to turn text into speech that does not sound like a robot](https://naadly.com/guides/text-to-speech): A practical text to speech guide: pick a voice, fix the pronunciation and pauses, render an MP3 or WAV, and know what you are allowed to do with the file. - [What an AI voice generator can and cannot do in 2026](https://naadly.com/guides/ai-voice-generator): An honest account of AI voice generation: what sounds human now, what still gives it away, what it costs, and where the law has caught up. - [Voice cloning, done legally and done well](https://naadly.com/guides/voice-cloning): How voice cloning works, what a usable reference recording sounds like, what consent you need on file, and where cloning is the wrong tool. - [Speech to speech: keep your performance, change the voice](https://naadly.com/guides/ai-voice-changer): How speech-to-speech voice conversion works, when it beats text to speech, and how to use an acted reference to get emotion out of a synthetic voice. - [Hindi and Indian-language text to speech that reads Devanagari properly](https://naadly.com/guides/hindi-text-to-speech): Hindi, Tamil, Telugu, Bengali, Marathi and 18 more: how Indian-language text to speech works on Naadly, what it costs, and where it still struggles. - [Japanese text to speech, and the kanji problem](https://naadly.com/guides/japanese-text-to-speech): How Japanese text to speech handles kanji readings and pitch accent, what Naadly uses, and how to check a render before it embarrasses you. - [A text-to-speech API you can read in one sitting](https://naadly.com/guides/text-to-speech-api): Authenticate, list voices, render speech, stream it, and get told when a long job finishes - with Python and JavaScript examples and the real limits. - [Dubbing a video with AI: transcribe, translate, re-voice](https://naadly.com/guides/ai-dubbing): How to dub video into another language with AI - transcription, translation, voice matching, subtitles and the lip-sync problem nobody solves. - [Narrating an audiobook with AI without it sounding like one](https://naadly.com/guides/audiobook-narration): How to produce an audiobook with AI narration: voice choice, chapter splitting, pronunciation lists, consistency and what the retailers require. - [AI voice-over for YouTube and short-form video](https://naadly.com/guides/youtube-voice-over): How to make AI voice-over for YouTube, Shorts, Reels and TikTok that fits the edit, keeps monetisation intact and does not sound like every other faceless channel. - [Voice-over for courses and training that survives version 7](https://naadly.com/guides/elearning-voice-over): How to narrate e-learning modules with AI so that updating slide 14 next quarter costs one line, not one studio day. - [AI voice for podcasts: intros, ads, corrections and full episodes](https://naadly.com/guides/podcast-voice-over): Where synthetic voice belongs in a podcast workflow - and where listeners will notice - plus transcripts, ad reads and fixing a line without a re-record. - [Transcription and subtitles that are actually usable](https://naadly.com/guides/transcription-and-subtitles): How to get an accurate transcript and clean SRT subtitles from audio or video, what accuracy to expect, and how to fix the parts machines always miss. - [Pronunciation, pauses and timing: the controls that matter](https://naadly.com/guides/pronunciation-and-timing): How to make a synthetic voice say your product name properly, pause where you want, and hit a length - with punctuation, a pronunciation dictionary and per-line speed. - [Bulk and long-form: hundreds of lines without babysitting](https://naadly.com/guides/bulk-and-long-form): How to render hundreds of lines or a whole book with a CSV import, queued jobs and webhooks - and what to check before starting a job you cannot watch. - [Give Claude, Cursor and ChatGPT a voice with MCP](https://naadly.com/guides/mcp-voice-for-claude-and-cursor): How to connect Naadly to an MCP client so an assistant can list voices and render speech itself - setup, what the tools do, and the limits. - [What you may legally do with AI-generated audio](https://naadly.com/guides/commercial-use-and-licensing): Who owns AI-generated audio, when open voice models require attribution, what a client can ask you for, and how Naadly's licensing works on every plan. - [Open-source text to speech: what it costs to run yourself](https://naadly.com/guides/open-source-text-to-speech): Piper, Kokoro, Whisper and friends: what open TTS models are good at, what running them actually costs, and when a hosted service is the cheaper answer. ## Comparisons - [Naadly vs ElevenLabs: where each one wins](https://naadly.com/compare/elevenlabs-alternative): An honest comparison of Naadly and ElevenLabs - expressiveness against price and per-minute metering - with what each is genuinely better at. - [Naadly vs Murf: studio workflow against render volume](https://naadly.com/compare/murf-alternative): Naadly and Murf compared on voices, editing workflow, collaboration, pricing model and what each is best suited to. - [Naadly vs Play.ht: API-first against catalogue-first](https://naadly.com/compare/play-ht-alternative): Naadly and Play.ht compared for developers and publishers: API shape, streaming, voice catalogue, licensing and price per hour of audio. - [Naadly vs Speechify: listening to text against producing audio](https://naadly.com/compare/speechify-alternative): Speechify is built for reading; Naadly is built for publishing. Which one you need, and why they are barely competitors. - [Naadly vs Resemble AI: cloning-first against catalogue-first](https://naadly.com/compare/resemble-ai-alternative): Resemble is built around custom cloned voices; Naadly is built around a licensed catalogue with cloning included. Which suits which job. - [Self-hosted TTS vs a hosted service, costed properly](https://naadly.com/compare/self-hosted-vs-hosted-tts): What it actually costs to run Piper, Kokoro and Whisper yourself versus paying for a hosted service - containers, engineering time and the things that break. ## Use cases - [For YouTube creators](https://naadly.com/use-cases/youtube-creators): Daily uploads, faceless channels and Shorts: how creators use Naadly for narration, hooks, subtitles and translated versions. - [For agencies and studios](https://naadly.com/use-cases/agencies): How agencies use Naadly for client voice-over: seats and roles, licence evidence for client legal teams, and reselling it under your own brand. - [For e-learning and training teams](https://naadly.com/use-cases/elearning): Module narration that can be updated one slide at a time, in several languages, with captions and answers for procurement. - [For publishers and long-form catalogues](https://naadly.com/use-cases/publishers): Turning a backlist into audio: chapter jobs, one consistent voice, pronunciation lists, and what retailers require. - [For developers putting voice inside a product](https://naadly.com/use-cases/app-developers): Adding speech to an app: synchronous renders, streaming, queued jobs with signed webhooks, per-key usage and the limits that matter. - [For phone systems and IVR](https://naadly.com/use-cases/phone-systems-and-ivr): Menu prompts, hold messages and out-of-hours announcements in u-law, regenerated whenever the opening hours change. - [For Indian regional content at volume](https://naadly.com/use-cases/indian-regional-content): Hindi, Tamil, Telugu, Bengali, Marathi and more for creators, edtech and D2C brands publishing across Indian languages. - [For accessible audio versions](https://naadly.com/use-cases/accessibility): Publishing an audio version of your material with captions, in several languages, licensed for distribution. ## Answers - [Is AI voice legal for commercial use?](https://naadly.com/answers/is-ai-voice-legal-for-commercial-use): Yes, provided the service licenses it and you are not imitating a real person. On Naadly, audio rendered on any plan - including free - may be used commercially; the open voice models behind it carry their own licences, mostly CC BY 4.0, and a few require a credit line, printed on each voice's page. Cloning an identifiable person additionally requires their documented consent. - [How much audio do you need to clone a voice?](https://naadly.com/answers/how-much-audio-to-clone-a-voice): About 30 seconds of clean, single-speaker speech. Recording quality matters far more than length: a quiet room, a consistent distance from the microphone and normal delivery beat ten minutes of noisy audio. Naadly also requires a spoken consent statement from the speaker before it will create the clone. - [Is there a genuinely free AI voice generator?](https://naadly.com/answers/is-there-a-free-ai-voice-generator): Yes, with limits that are stated rather than hidden. Naadly gives 30 minutes of audio a month without a card, and the audio may be used commercially. Free tools that advertise unlimited generation usually pay for it with an audible watermark, with rights to your input, or by closing. - [Does AI voice sound human yet?](https://naadly.com/answers/does-ai-voice-sound-human): For narration, usually - listeners do not notice a good synthetic narrator reading well-written copy. For performance it does not: sarcasm, grief, comic timing and two-person dialogue remain audibly synthetic, and the giveaway is rhythm rather than timbre. The reliable workaround is to act the line yourself and convert it to the voice you want. - [How do you make an AI voice sound natural?](https://naadly.com/answers/how-to-make-ai-voice-sound-natural): Fix the script before the settings: one idea per sentence, punctuation where a breath goes, numbers written the way they should be heard, and names in the pronunciation dictionary. Then choose a voice whose accent and pace match the material, slow it slightly, and re-render individual lines rather than accepting the first pass. - [What does an hour of AI voice-over cost?](https://naadly.com/answers/text-to-speech-price-per-hour): It depends whether you are billed for characters or minutes. Naadly meters minutes: $15 a month buys 6,000 minutes, and standard-language speech is uncapped on $39. Per-character services work out at roughly 55,000-60,000 characters per finished hour, which is the conversion to do before comparing anything. - [Can you monetise YouTube videos with an AI voice?](https://naadly.com/answers/can-youtube-monetise-ai-voice): Yes. The policies that get channels demonetised target mass-produced, repetitive, low-value content, not synthetic narration as such - a well-made video with an AI narrator is monetisable. Platforms do provide disclosure fields for realistic synthetic media and expect them to be used; check the current policy where you publish. - [What is the best AI voice for an audiobook?](https://naadly.com/answers/best-ai-voice-for-audiobooks): A deliberate, clearly-graded voice in your reader's accent, chosen by auditioning it on your hardest page rather than on a demo line - dialogue and a name-heavy paragraph. Whatever you pick, keep it for the entire book and the entire series: consistency matters more than the marginal naturalness of a second voice. - [Which Indian languages can AI text to speech read?](https://naadly.com/answers/which-indian-languages-are-supported): Naadly reads 23: Assamese, Bengali, Bodo, Dogri, Gujarati, Hindi, Indian English, Kannada, Kashmiri, Konkani, Maithili, Malayalam, Manipuri, Marathi, Nepali, Odia, Punjabi, Sanskrit, Santali, Sindhi, Tamil, Telugu and Urdu. The engine infers the language from the script you paste, so Devanagari Hindi with English brand names inside it reads correctly. - [Do AI voice services train on the text and audio you upload?](https://naadly.com/answers/is-my-data-used-to-train-models): Some do; Naadly does not. Scripts, uploads and reference recordings are used to do the job you asked for and nothing else, and no model here is trained on customer input. You can export the whole workspace as JSON or delete it - audio in storage included - from Settings, without emailing anybody. - [Are AI-generated voices watermarked?](https://naadly.com/answers/are-ai-voices-watermarked): On Naadly, yes: every render carries an inaudible AudioSeal watermark, is served with a header declaring it synthetic, and has a provenance record. It does not change what the audio sounds like, and it is not optional - it is what lets you answer a platform's disclosure field, a broadcaster or a procurement questionnaire. - [Do you need permission to clone someone's voice?](https://naadly.com/answers/do-i-need-consent-to-clone-a-voice): Yes. Cloning an identifiable person without documented consent exposes you to personality and publicity claims and, increasingly, to specific AI statutes - and it is prohibited here. Naadly requires the speaker to record a consent statement aloud before a clone is created, and keeps that record with a hash of the reference clip as evidence. - [Can AI dubbing match lip movement?](https://naadly.com/answers/can-ai-dub-with-lip-sync): Naadly does not attempt lip sync. Dubs are timed to subtitle cues, which is correct for narration, interviews, courses and screencasts, and visibly wrong for close-up dialogue. If your footage has faces filling the frame, budget for a lip-sync tool or re-shoot - a dub that nearly matches is worse than one that clearly does not. - [What is the difference between voice cloning and a voice changer?](https://naadly.com/answers/difference-between-voice-cloning-and-voice-changer): Cloning builds a new voice from a reference recording and then reads any text in it. A voice changer converts one specific recording into an existing voice, keeping your timing, emphasis and emotion. If you need a voice to read future scripts, clone; if you need this line delivered well, use the changer. - [How many AI voices does Naadly have?](https://naadly.com/answers/how-many-voices-does-naadly-have): 1,481 labelled voices across 11 catalogue languages, plus 23 Indian languages on a dedicated engine and 35 translation languages. Every voice is labelled for gender, accent, pace, tone and measured quality, and every one is cleared for commercial use with its licence and attribution line printed on its page. - [What audio formats can AI text to speech export?](https://naadly.com/answers/what-audio-formats-can-i-export): Naadly exports MP3 for video and the web, WAV for further editing, FLAC for archiving, OGG for the web and 8 kHz u-law for telephony, plus an SRT subtitle file timed from the render. MP3 for delivery, WAV or FLAC if anything downstream will process the audio again. - [Can you use AI voice-over in client work?](https://naadly.com/answers/can-i-use-ai-voice-for-client-work): Yes, and the practical requirement is evidence rather than permission: clients ask whose voice it is, whether it can run in a paid campaign, and whether it must be disclosed. Naadly prints each voice's model, licence and credit line, licenses commercial use on every plan, watermarks every render and keeps a provenance log. - [Is there a free text-to-speech API?](https://naadly.com/answers/is-there-a-text-to-speech-api-free-tier): Not on Naadly: API keys start on the $15 plan, because an open API on a free tier is a bill somebody else pays. The free tier is 30 minutes a month through the interface. Free API tiers elsewhere usually cap requests hard or reserve rights over your input. - [Is text to speech billed per character or per minute?](https://naadly.com/answers/do-you-charge-per-character-or-per-minute): Both models exist, and it changes how you work. Per-character billing makes every retry a small decision; Naadly meters minutes of audio - 30 free a month, 6,000 on $15, uncapped standard speech on $39 - because it runs its own models, so re-rendering a line eleven times costs almost nothing. - [How do you stop an AI voice mispronouncing a name?](https://naadly.com/answers/how-do-i-fix-a-mispronounced-name): Add the word to the workspace pronunciation dictionary, spelled the way it should sound, once. Every voice, project, retry and teammate on the account inherits the fix from then on. Do it before a long render, not after finding the same name wrong in nineteen chapters. - [Can AI voices express emotion?](https://naadly.com/answers/can-ai-voices-express-emotion): Not reliably from text alone - which is why so many demos use the same upbeat register. The dependable route is speech to speech: record yourself delivering the line with the feeling you want and convert it to the target voice, which keeps your timing and emphasis and changes only the timbre. - [Can you delete your data from an AI voice service?](https://naadly.com/answers/can-i-delete-my-account-and-data): On Naadly, an owner can export the whole workspace as JSON and delete it from Settings - accounts, projects, lines, jobs, cloned voices, usage and the audio in object storage, in the same request. Two things deliberately survive: cloning consent records, which protect the person who was cloned, and administrator audit entries. - [What happens to your audio if you cancel an AI voice subscription?](https://naadly.com/answers/what-happens-if-i-cancel): On Naadly, files you already rendered stay yours and stay usable - the licence does not expire with the subscription. The account drops to the free tier's limits, paid features stop, and there is no refund for the remainder of a period; subscriptions are cancelled rather than refunded, with the statutory exceptions listed in the refunds policy. ## Notes - [What rendering all 1,481 voices in production found](https://naadly.com/blog/what-a-1481-voice-sweep-found): A full catalogue sweep against production found 169 failing voices - and the cause was a memory quota, not the voices. What we changed. - [Why we meter minutes instead of characters](https://naadly.com/blog/why-we-meter-minutes-not-characters): Per-character pricing is a pass-through of somebody else's bill. Running our own models changes what the honest unit is - and makes retries free. - [Watermarking every render, and the bug that taught us how](https://naadly.com/blog/watermarking-every-render): Why every Naadly render carries an AudioSeal watermark and a synthetic-media header - and the thread-local state bug that broke it after the first request. - [Japanese text to speech needs a real front-end, not a bigger model](https://naadly.com/blog/japanese-needs-a-real-front-end): Kanji readings and pitch accent are dictionary problems. How we measured it, and why we refuse the other languages that model claims. - [A cold container deleted an entire language from the shop](https://naadly.com/blog/cold-container-deleted-a-language): A five-minute cold start made every Hindi and Japanese voice vanish and turned saved voices into server errors. The cache design that caused it. - [Consent records that outlive the account that made them](https://naadly.com/blog/consent-records-that-outlive-the-account): Deleting a workspace removes everything - except the cloning consent records and the admin audit log. Why those two exceptions are the right ones.