Port of OpenAI's Whisper model in C/C++
Top SPEECH-RECOGNITION GitHub Repositories & Tools (2026)
Discover the most starred and trending open source tools tagged with #speech-recognition.
π€ Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Frontier CoreML audio models in your apps β text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Native macOS transcription for selected-app audio and microphone input, using Appleβs on-device speech frameworks with durable timestamped Markdown output
Open-source, offline voice-to-text for macOS. Hold a hotkey, speak, text appears. Private on-device dictation with multiple speech engines.
AI speech toolkit for Apple Silicon β ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
On-device Speech AI for Apple Silicon