- 46kTTStext-to-speech
πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
46kMPL-2.0 - 40kChatTTStext-to-speech
A generative speech model for daily dialogue.
40kAGPL-3.0 - 39kBarktext-to-speech
π Text-Prompted Generative Audio Model
39kMIT - 37kOpenVoicetext-to-speech
Instant voice cloning by MIT and MyShell. Audio foundation model.
37kMIT - 37kRetrieval-based-Voice-Conversion-WebUItext-to-speech
Easily train a good VC model with voice data <= 10 mins!
37kMIT - 31kFish Speechtext-to-speech
SOTA Open Source TTS
31kNot specified - 28kSoftVC VITS Singing Voice Conversiontext-to-speech
SoftVC VITS Singing Voice Conversion
28kAGPL-3.0 - 24kAudioCrafttext-to-speech
Audiocraft is a library for audio processing and generation with deep learning. It features the state-of-the-art EnCodec audio compressor / tokenizer, along with MusicGen, a simple and controllable music generation LM with textual and melodic conditioning.
24kMIT - 11kPipertext-to-speech
A fast, local neural text to speech system
11kMIT - 11kMoshitext-to-speech
Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
11kApache-2.0 - SponsorReach 50,000+ buyers
Enterprise buyers looking for private AI solutions see your brand here.
- 8.1kkokorotext-to-speech
https://hf.co/hexgrad/Kokoro-82M
8.1kApache-2.0 - 5.6kParler-TTStext-to-speech
Inference and training library for high-quality TTS models.
5.6kApache-2.0
Stop paying for AI APIs. Everything here runs on your hardware.
Enterprise buyers looking for private AI solutions see your brand here.
No fluff. No spam. Join 12,000+ builders.