#speech-to-text
Open source repositories tagged with #speech-to-text, ranked by health score.
VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Local speech-to-text for macOS on-device AI, fully private, optional cloud
Lightning-fast, free, local first voice dictation for macOS with on-device transcription
Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!
Open-source speech-to-text and dictation app for macOS, Windows, iPhone, iPad, and Android, with offline recognition and optional cleanup.
On-device Speech AI for Apple Silicon
Free, sub-second on-device AI dictation for macOS. Dual engines (Whisper + Parakeet) on Apple Silicon. Optional AI polish with our own on-device model (EG-1), Apple Intelligence (macOS 26), Ollama, or your own OpenAI/Gemini key. No account, no subscription, audio never leaves your Mac.
🎙️ Offline voice productivity for Windows - dictate text into any app and control a local AI assistant by voice. Manages notes, to-do lists, appointments & reminders. 100% local: faster-whisper + Ollama + SQLite. No cloud, no telemetry. Free and open source alternative to Wispr Flow