Speech Recognition
k2-fsa/sherpa๐ฅ active
Speech-to-text server framework with next-gen Kaldi
C++Apache-2.0updated Sep 24, 2026
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows, Raspberry Pi, VisionFive2, LicheePi4A etc.
$ git clone https://github.com/k2-fsa/sherpa-ncnn.gitSpeech-to-text server framework with next-gen Kaldi
An Optimized Speech-to-Text Pipeline for the Whisper Model Supporting Multiple Inference Engine
WebSocket, gRPC and WebRTC speech recognition server based on Vosk and Kaldi libraries
Whisper.net. Speech to text made simple using Whisper Models
Data from GitHub ยท snapshot Sep 24, 2026