FishAudio S1-mini
Compact variant of Fish Speech optimized for faster inference.
About
FishAudio S1-mini is a compact variant of the Fish Speech text-to-speech system tuned for faster inference and lower VRAM use while keeping usable voice quality on consumer hardware. Like the larger Fish Audio models it supports multilingual synthesis and zero-shot voice cloning from a reference clip. It suits latency-sensitive or resource-constrained deployments. Distributed as part of the open-source Fish Speech project.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Easy (2/5)
- License
- Apache-2.0
- Minimum VRAM
- 4 GB
- Added
- Apr 3, 2026
Related Tools
Lightweight and expressive TTS model with 82M parameters for fast local inference.
Conversational TTS model optimized for dialogue and chat applications.
Multilingual large voice generation model with full-stack inference, training, and deployment.
Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.
Emotion-controllable TTS engine by NetEase with 2000+ voices.
Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.