Tools/Text-to-Speech (TTS)/Sherpa-ONNX TTS

Sherpa-ONNX TTS

Cross-platform TTS using ONNX Runtime for on-device speech synthesis.

Open SourceSelf HostedOffline Capable
0.0 (0)

About

Text-to-speech is one part of sherpa-onnx, an on-device speech toolkit from the k2-fsa next-gen Kaldi project that runs entirely offline through ONNX Runtime. On the synthesis side it executes models from several engines, including VITS, Piper, Matcha, Kokoro, and MeloTTS, producing multilingual speech without any cloud service; the same framework also covers speech recognition, voice activity detection, speaker identification and diarization, keyword spotting, and speech enhancement. Portability is the project's defining trait: it runs on x86, 32-bit and 64-bit ARM, and RISC-V processors, on Linux, Windows, macOS, Android, iOS, HarmonyOS, and WebAssembly, and on boards like Raspberry Pi and NVIDIA Jetson, with NPU support for several vendors. Bindings exist for C++, C, Python, Go, C#, Java, Kotlin, JavaScript, Swift, Rust, Dart, and Object Pascal. Open source under Apache 2.0, it is a common choice for embedding offline speech synthesis in mobile apps, desktop software, and embedded devices where privacy or connectivity rules out hosted APIs.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Easy (2/5)
License
Apache-2.0
Added
Apr 3, 2026

Related Tools

Featured

Lightweight and expressive TTS model with 82M parameters for fast local inference.

Open SourceSelf HostedOffline
Easy
4.0 (1)

Conversational TTS model optimized for dialogue and chat applications.

Open SourceSelf HostedOfflineGPU 4GB+
Intermediate
0.0 (0)

Multilingual large voice generation model with full-stack inference, training, and deployment.

Open SourceSelf HostedOfflineGPU
Intermediate
0.0 (0)

Large-scale multilingual TTS model by Alibaba with zero-shot voice cloning.

Open SourceSelf HostedOfflineGPU 8GB+
Advanced
0.0 (0)

Emotion-controllable TTS engine by NetEase with 2000+ voices.

Open SourceSelf HostedOfflineGPU 4GB+
Intermediate
0.0 (0)
Featured

Transformer-based text-to-audio model by Suno that generates speech, music, and sound effects.

Open SourceSelf HostedOfflineGPU 4GB+
Intermediate
0.0 (0)
Browse all Text-to-Speech (TTS) tools