Kokoro TTS
Lightweight and expressive TTS model with 82M parameters for fast local inference.
About
Kokoro TTS is a lightweight text-to-speech model with only 82 million parameters that achieves high-quality expressive speech synthesis. Supports multiple voices and styles. Fast enough for real-time inference on consumer hardware. Apache 2.0 license.
Reviews (1)
Leave a Review
Decent tool but just doesn't have the best sounding voices
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Easy (2/5)
- License
- Apache-2.0
- Added
- Apr 3, 2026
Related Tools
Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.
Open-source TTS model by Resemble AI with emotion and accent control.
Expressive zero-shot TTS model by Resemble AI with emotion and accent control.
Singing voice conversion model based on VITS and SoftVC for voice-to-voice transfer.
Zero-shot TTS model with high naturalness and speaker similarity.
Instant voice cloning TTS by MyShell requiring only a short audio reference.