EmotiVoice

Emotion-controllable TTS engine by NetEase with 2000+ voices.

Open SourceSelf HostedOffline CapableGPU Required (4GB+ VRAM)
0.0 (0)

About

EmotiVoice is an open-source text-to-speech engine from NetEase Youdao whose distinguishing feature is emotion control: generated speech can sound happy, excited, sad, angry, or otherwise expressive based on a prompt. It ships more than 2000 preset voices covering English and Chinese, and supports voice cloning when trained on a user's own recordings. Style factors such as pitch, speed, and energy are controllable alongside emotion, following a transformer-based design inspired by PromptTTS. The project provides several ways to run it: an interactive Streamlit web demo, batch synthesis via scripts, and a FastAPI server exposing an OpenAI-compatible TTS HTTP API, with a Docker image for setup on NVIDIA GPUs. Released under the Apache 2.0 license, it is used by developers and content creators who need expressive narration for audiobooks, voiceovers, dubbing, and accessibility features without relying on a commercial TTS service.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Intermediate (3/5)
License
Apache-2.0
Minimum VRAM
4 GB
Added
Apr 3, 2026

Related Tools

Featured

Lightweight and expressive TTS model with 82M parameters for fast local inference.

Open SourceSelf HostedOffline
Easy
4.0 (1)

Conversational TTS model optimized for dialogue and chat applications.

Open SourceSelf HostedOfflineGPU 4GB+
Intermediate
0.0 (0)

Multilingual large voice generation model with full-stack inference, training, and deployment.

Open SourceSelf HostedOfflineGPU
Intermediate
0.0 (0)

Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.

Open SourceSelf HostedOfflineGPU 8GB+
Advanced
0.0 (0)

Compact open-source speech synthesizer supporting 100+ languages.

Open SourceSelf HostedOffline
Beginner
0.0 (0)
Featured

Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.

Open SourceSelf HostedOfflineGPU 4GB+
Intermediate
0.0 (0)
Browse all Text-to-Speech (TTS) tools