MeloTTS
High-quality multilingual TTS library by MyShell with fast CPU inference.
About
MeloTTS came out of a collaboration between MIT and MyShell.ai as a multilingual text-to-speech library with an unusual emphasis on CPU speed: it performs real-time inference without a GPU. Language coverage includes English with American, British, Indian, and Australian accents, plus Spanish, French, Japanese, Korean, and Chinese, including mixed Chinese and English sentences. The implementation builds on the VITS family of architectures (VITS, VITS2, and Bert-VITS2) and can be used through a Python API, a command line interface, a web UI, or hosted demos, with documentation for training voices on custom datasets. Everything ships under the MIT license, which allows free commercial and non-commercial use, a point that distinguishes it from several popular TTS projects with restrictive terms. It is a frequent choice for embedding speech output in applications, voice agents, and edge deployments where GPU hosting is impractical.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Easy (2/5)
- License
- MIT
- Added
- Apr 3, 2026
Related Tools
Lightweight and expressive TTS model with 82M parameters for fast local inference.
Conversational TTS model optimized for dialogue and chat applications.
Multilingual large voice generation model with full-stack inference, training, and deployment.
Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.
Emotion-controllable TTS engine by NetEase with 2000+ voices.
Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.