Speech Recognition
QuentinFuxa/WhisperLiveKit๐ฅ active
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
PythonApache-2.0updated Sep 21, 2026
A 0.9B model for long-form transcription in 50+ languages with speaker diarization, timestamps, and acoustic event awareness
$ git clone https://github.com/OpenMOSS/MOSS-Transcribe-Diarize.gitReal-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!
turnkey self-hosted offline transcription and diarization service with llm summary
On-device streaming speech-to-text engine powered by deep learning
On-device speech-to-text engine powered by deep learning
Data from GitHub ยท snapshot Sep 24, 2026