Open source repositories tagged with #speech-recognition, ranked by health score.
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
A private logbook with a staff of personal AI assistants. Agents read what you record and propose what to do next — you approve the changes. End-to-end encrypted sync between your own devices — servers only ever see ciphertext. Local AI optional.
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
Native UI for the Whispering Tiger project - https://github.com/Sharrnah/whispering (live transcription / translation)
Lightning-fast, free, local first voice dictation for macOS with on-device transcription
Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!
Open-source speech-to-text and dictation app for macOS, Windows, iPhone, iPad, and Android, with offline recognition and optional cleanup.
On-device Speech AI for Apple Silicon
Fully-offline transcription and translator w/ speech-to-text and text-to-speech supporting 76 models across 34 languages