So-VITS-SVC
Singing voice conversion model based on VITS and SoftVC for voice-to-voice transfer.
About
So-VITS-SVC is a singing voice conversion model that combines a SoftVC content encoder with the VITS synthesis architecture to convert one singer's voice into another while preserving the performance. It is focused on singing voice conversion rather than text-to-speech and is widely used for voice cloning in music. The repository is community-maintained and archived. Distributed under its repository license.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Text-to-Speech (TTS)
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Advanced (4/5)
- Minimum VRAM
- 6 GB
- Added
- Apr 3, 2026
Related Tools
Lightweight and expressive TTS model with 82M parameters for fast local inference.
Conversational TTS model optimized for dialogue and chat applications.
Multilingual large voice generation model with full-stack inference, training, and deployment.
Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.
Emotion-controllable TTS engine by NetEase with 2000+ voices.
Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.