Kokoro-82M vs KittenTTS Nano

See how the two models differ in download size, language coverage, runtime needs, voice features, and output format. When both are available, use the same sentence to decide which voice works better for your project.

Kokoro-82MKittenTTS Nano
Parameters82M15M
Download88 MB CPU / 310 MB WebGPU57 MB
SpeedRuns around real time on a current CPU and becomes faster when WebGPU is available.Runs faster than real time on many CPU-only devices with a 57 MB download.
Voices288
Languages21
Voice cloningNoNo
StreamingYesNo
HardwareCPU ok, WebGPU fasterAny CPU (WASM)
LicenseApache-2.0Apache-2.0
Output24.0 kHz24.0 kHz

Which should you choose?

  • Kokoro-82M if its main focus matches your project: natural english speech with a practical browser download.
  • KittenTTS Nano if its main focus matches your project: a small english model for lightweight devices.
  • KittenTTS Nano if a smaller download is important; its listed download is 57 MB.
  • Kokoro-82M if you want audio to start playing before generation finishes.

Frequently asked questions

Which sounds better, Kokoro-82M or KittenTTS Nano?

There is no universal winner. Listen for pronunciation, pacing, and tone on the script you actually plan to use, then weigh that result against download size and device requirements. Both models are available here, so you can generate the same sentence with each.

Where do the speed and size details come from?

Model sizes and runtime requirements come from the maintained model records and linked upstream sources. Use the benchmark page to measure generation speed on your own device.