QuentinFuxa/WhisperLiveKit๐ฅ active
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
Optimized Whisper models for streaming and on-device use
$ git clone https://github.com/TheStageAI/TheWhisper.gitReal-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
On-device, real-time multimodal AI with features similar to GPT-Live
ARIA - AI Realtime Intelligent Audio | Universal real-time AI subtitles for Windows
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!
AI speech toolkit for Apple Silicon โ ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
Data from GitHub ยท snapshot Sep 24, 2026