Chatterbox TTS (Resemble)
Open-source TTS model by Resemble AI with emotion and accent control.
About
Chatterbox from Resemble AI is an open-source text-to-speech model family aimed at expressive narration and low-latency voice agents. The Turbo variant uses a 350M parameter backbone and a distilled mel decoder that produces output in a single step rather than ten, and accepts paralinguistic tags such as [cough], [laugh], and [chuckle] inline. Released under the MIT license; a paid hosted service is available for higher-volume use.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Easy (2/5)
- License
- MIT
- Minimum VRAM
- 4 GB
- Added
- Apr 3, 2026
Related Tools
Lightweight and expressive TTS model with 82M parameters for fast local inference.
Conversational TTS model optimized for dialogue and chat applications.
Multilingual large voice generation model with full-stack inference, training, and deployment.
Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.
Emotion-controllable TTS engine by NetEase with 2000+ voices.
Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.