Open-weight model

Kokoro TTS Online

Speech studio

32 voices · 18 languages

5,000 free characters/day
141 / 1,000
Pick a voicePlay any of them
Free voices
Natural voices3 cr / 1k chars
Premium voices10 cr / 1k chars
Ultra voices20 cr / 1k chars

Try Kokoro-82M in your browser — no Python, no GPU, no model download. Type your text, pick a voice, and download the MP3. It is the same open-weight model that powers our free tier.

Apache-2.0 modelNo installFree to try

How it works

Three steps, about a minute

  1. 1

    Paste or write your text

    Drop in a script, a paragraph, or a whole article. Longer text is split into parts automatically and joined back into one file.

  2. 2

    Pick a voice and preview it

    Every voice previews before you spend anything, including the paid ones. Filter by language, then hear two or three before committing.

  3. 3

    Generate and download

    Audio comes back in seconds as MP3 or WAV, with a commercial licence on every tier — including the free one.

A small studio microphone on a minimal wooden desk in morning light

Examples

Hear it on real scripts

Every clip below was generated here on the free tier — no account, no credits. Play one to hear the voice on that kind of writing, or load the script into the generator and make your own version of it.

YouTube intro

Opens a video without the "hey guys, welcome back" throat-clearing.

Three years ago, this cost four thousand dollars and a studio. Today it costs nothing and takes about a minute. In this video I'll show you exactly how — no jargon, no upsell, and every step is one you can copy while you watch.

Read by Bella · Warm, natural narration

YouTube outro

Closes the loop and asks for one thing, not four.

That's the whole method. If it worked for you, the next video takes it further — same tools, bigger project. And if it didn't, tell me where it broke. I read every comment, and half of these videos exist because someone got stuck.

Read by Adam · Clear, confident

Audiobook passage

Narrative prose, to hear how a voice handles description and pace.

The house had been empty for eleven years, and it had the particular stillness of a place that has stopped expecting anyone. Dust lay on the banister in a way that suggested nobody had run a hand along it in a very long time. She did, once, on the way up, and felt slightly guilty about it.

Read by Bella · Warm, natural narration

Included

What you get

Free without an account
Thousands of characters a day as a guest, no sign-up and nothing to install. An account raises the limit; it is not a gate on the tool.
Commercial use included
Monetised videos, client work and ads are all fine on every tier. You own what you generate.
MP3 and WAV export
MP3 for editors and uploads, WAV where you need uncompressed audio. Speed is adjustable from 0.75× to 1.5× on engines that support it.
Long text, one file
Text past the per-request limit is split at sentence boundaries, generated in order, and stitched back into a single download.
Voices across 18 languages
Several languages are read by voices trained on that language rather than an English voice reading foreign words — the difference is audible immediately.
Credits that never expire
Premium and ultra voices are billed from one-time credit packs. No subscription and no monthly reset.

Why Kokoro got popular

Most open TTS models ask you to trade something away: quality, speed, or a license that quietly forbids commercial use. Kokoro-82M is unusual because it gives you all three — genuinely natural output, fast inference even on modest hardware, and a clean Apache-2.0 license.

At roughly 82 million parameters it is a fraction of the size of most competing models, which is what makes it cheap enough to run at scale. That combination is why it shows up as the default engine in so many self-hosted setups.

We run Kokoro too — that is the point

The free tier of this site is powered by Kokoro. So when you generate audio above, you are hearing the real model, not a demo with a different engine swapped in behind the scenes.

That also means you can use this page as a straightforward capability check: if Kokoro sounds good enough for your project here, it will sound the same when you self-host it. If it does not, you can switch to a premium engine in the voice selector and compare them side by side in the same session.

When to self-host vs. use it here

A rough guide:

  • Use it here when you want to evaluate the model, generate a handful of files, or avoid infrastructure entirely
  • Self-host when you have steady high volume, need offline operation, or want to fine-tune
  • Use a premium engine when you need strong emotional range, voice cloning, or broad language coverage beyond what Kokoro handles

Questions

Frequently asked