Speech Recognition
QuentinFuxa/WhisperLiveKitπ₯ active
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
PythonApache-2.0updated Sep 21, 2026
simple delaysum, MVDR and CGMM-MVDR
$ git clone https://github.com/AkojimaSLP/Beamforming-for-speech-enhancement.gitReal-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
End-to-End Speech Processing Toolkit
Speech recognition module for Python, supporting several engines and APIs, online and offline.
A Deep-Learning-Based Chinese Speech Recognition System εΊδΊζ·±εΊ¦ε¦δΉ ηδΈζθ―ι³θ―ε«η³»η»
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows, Raspberry Pi, VisionFive2, LicheePi4A etc.
Data from GitHub Β· snapshot Sep 24, 2026