eSpeak NG
Compact open-source speech synthesizer supporting 100+ languages.
About
eSpeak NG is a compact open-source speech synthesizer that supports more than 100 languages and accents. Instead of neural networks it uses formant synthesis, which keeps the whole system down to a few megabytes and keeps speech intelligible even at very high reading speeds, at the cost of a distinctly robotic timbre. Written in C, it runs on Linux, Windows, Android, macOS, and BSD, and works as a command-line program, a shared library, or a Windows SAPI5 voice; it can also act as a front end to MBROLA diphone voices for more natural output. SSML and HTML markup are supported, and the C API stays compatible with the original eSpeak. Because it needs no GPU and almost no resources, it is a staple of screen readers such as NVDA, embedded devices, and accessibility projects, and its text-to-phoneme rules are often borrowed as a front end by neural TTS systems. Licensed under GPL 3.0.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Beginner (1/5)
- License
- GPL-3.0
- Added
- Apr 3, 2026
Related Tools
Lightweight and expressive TTS model with 82M parameters for fast local inference.
Conversational TTS model optimized for dialogue and chat applications.
Multilingual large voice generation model with full-stack inference, training, and deployment.
Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.
Emotion-controllable TTS engine by NetEase with 2000+ voices.
Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.