Speech Recognition
shashikg/WhisperS2T
An Optimized Speech-to-Text Pipeline for the Whisper Model Supporting Multiple Inference Engine
Jupyter NotebookMITupdated Aug 27, 2024
Whisper.net. Speech to text made simple using Whisper Models
$ git clone https://github.com/sandrohanea/whisper.net.gitAn Optimized Speech-to-Text Pipeline for the Whisper Model Supporting Multiple Inference Engine
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows, Raspberry Pi, VisionFive2, LicheePi4A etc.
๐ A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
On-device voice activity detection (VAD) powered by deep learning
Data from GitHub ยท snapshot Sep 24, 2026