WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Buzz transcribes and translates audio offline on your personal computer. Powered by OpenAI's Whisper.
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on devices ranging from a Raspberry Pi 4 to high power GPU servers.