๐Ÿ† #974 overall#95 of 301 in Speech Recognition

alumae /kaldi-gstreamer-server

Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.

$ git clone https://github.com/alumae/kaldi-gstreamer-server.git
GitHub social preview for alumae/kaldi-gstreamer-server
Stars
1.1K
1,094
Forks
337
337
Language
Python
License
BSD-2-Clause
Created
Jan 6, 2014
12.7 years old
Last push
Jun 8, 2024

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

ggml-org/whisper.cpp๐Ÿ”ฅ active

Port of OpenAI's Whisper model in C/C++

C++MITupdated Sep 24, 2026
GitHub โ†—โ˜… 53.9Kโ‘‚ 6.2K
Speech Recognition

m-bain/whisperX

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

PythonBSD-2-Clauseupdated Aug 30, 2026
GitHub โ†—โ˜… 24.2Kโ‘‚ 2.4K
Speech Recognition

kaldi-asr/kaldi

kaldi-asr/kaldi is the official location of the Kaldi project.

ShellOtherupdated Sep 22, 2025
GitHub โ†—โ˜… 15.5Kโ‘‚ 5.4K
Speech Recognition

abus-aikorea/voice-pro

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

PythonGPL-3.0updated Jul 13, 2026
GitHub โ†—โ˜… 12.9Kโ‘‚ 1.9K
Speech Recognition

PaddlePaddle/PaddleSpeech

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

PythonApache-2.0updated Aug 12, 2026
GitHub โ†—โ˜… 12.7Kโ‘‚ 2K

Data from GitHub ยท snapshot Sep 24, 2026