Bark (Audio Effects)
Transformer-based text-to-audio model by Suno supporting speech, music, and sound effects.
About
Bark by Suno is a transformer text-to-audio model that generates realistic multilingual speech along with music, background noise, and simple sound effects from a text prompt, and can produce nonverbal sounds like laughter and sighing. It is fully generative rather than a conventional text-to-speech system, and pretrained checkpoints are available for research and commercial use. Released under the MIT license.
Reviews (0)
Leave a Review
No reviews yet. Be the first to review!
Details
- Category
- Music & Audio Generation
- Price
- Free
- Platform
- Local/Desktop
- Difficulty
- Easy (2/5)
- License
- MIT
- Minimum VRAM
- 6 GB
- Added
- Apr 3, 2026
Related Tools
Audio generation framework by Meta including MusicGen for text-to-music.
Latent diffusion model for text-to-audio, music, and speech generation.
Audio super-resolution model for upsampling audio to higher sample rates.
State-of-the-art music source separation model by Meta for splitting tracks.
Fast music generation model producing full songs with lyrics in seconds.
PyTorch library for deep learning research on audio generation including MusicGen and AudioGen.