NeuTTS Air
Lightweight neural TTS model optimized for edge and mobile deployment.
About
NeuTTS Air is a small text-to-speech model from Neuphonic built for on-device deployment. It pairs a compact LLM backbone of roughly 360 million active parameters, about 550 million including embeddings, with NeuCodec, a 50 Hz single-codebook neural audio codec that keeps audio quality high at low bitrates. The model clones a voice from as little as three seconds of reference audio and generates natural-sounding speech in real time on mid-range consumer hardware, which makes it suitable for embedded voice agents, assistants, toys, and other applications that need local audio generation without calling cloud APIs. Input text is converted to phonemes before synthesis, and generated audio carries a watermark. It ships in PyTorch form plus GGUF quantizations at Q4 and Q8 for efficient local inference, and the model is released under the Apache 2.0 license.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Easy (2/5)
- Added
- Apr 3, 2026
Related Tools
Lightweight and expressive TTS model with 82M parameters for fast local inference.
Conversational TTS model optimized for dialogue and chat applications.
Multilingual large voice generation model with full-stack inference, training, and deployment.
Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.
Emotion-controllable TTS engine by NetEase with 2000+ voices.
Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.