- 105kWhisperspeech-to-text
Robust Speech Recognition via Large-Scale Weak Supervision
105kMIT - 52kwhisper.cppspeech-to-text
Port of OpenAI's Whisper model in C/C++
52kMIT - 24kFaster Whisperspeech-to-text
Faster Whisper transcription with CTranslate2
24kMIT - 23kWhisperXspeech-to-text
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
23kBSD-2-Clause - 20kBuzzspeech-to-text
Buzz transcribes and translates audio offline on your personal computer. Powered by OpenAI's Whisper.
20kMIT - 18kNVIDIA NeMo Speechspeech-to-text
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
18kApache-2.0 - 15kVosk Speech Recognition Toolkitspeech-to-text
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
15kApache-2.0 - SponsorReach 50,000+ buyers
Enterprise buyers looking for private AI solutions see your brand here.
- 12kSpeechBrainspeech-to-text
A PyTorch-based Speech Toolkit
12kApache-2.0 - 9.6kSilero VADspeech-to-text
Silero VAD: pre-trained enterprise-grade Voice Activity Detector
9.6kMIT - 2.8kWhisper WebUIspeech-to-text
A Web UI for easy subtitle using whisper model.
2.8kApache-2.0
Stop paying for AI APIs. Everything here runs on your hardware.
Enterprise buyers looking for private AI solutions see your brand here.
No fluff. No spam. Join 12,000+ builders.