๐Ÿ† #668 overall#52 of 301 in Speech Recognition

syhw /wer_are_we

Attempt at tracking states of the arts and recent results (bibliography) on speech recognition.

$ git clone https://github.com/syhw/wer_are_we.git
GitHub social preview for syhw/wer_are_we
Stars
1.9K
1,864
Forks
224
224
Language
Other
License
None
Created
Aug 6, 2015
11.1 years old
Last push
Jun 27, 2022

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

ggml-org/whisper.cpp๐Ÿ”ฅ active

Port of OpenAI's Whisper model in C/C++

C++MITupdated Sep 24, 2026
GitHub โ†—โ˜… 53.9Kโ‘‚ 6.2K
Speech Recognition

m-bain/whisperX

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

PythonBSD-2-Clauseupdated Aug 30, 2026
GitHub โ†—โ˜… 24.2Kโ‘‚ 2.4K
Speech Recognition

kaldi-asr/kaldi

kaldi-asr/kaldi is the official location of the Kaldi project.

ShellOtherupdated Sep 22, 2025
GitHub โ†—โ˜… 15.5Kโ‘‚ 5.4K
Speech Recognition

abus-aikorea/voice-pro

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

PythonGPL-3.0updated Jul 13, 2026
GitHub โ†—โ˜… 12.9Kโ‘‚ 1.9K
Speech Recognition

PaddlePaddle/PaddleSpeech

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

PythonApache-2.0updated Aug 12, 2026
GitHub โ†—โ˜… 12.7Kโ‘‚ 2K

Data from GitHub ยท snapshot Sep 24, 2026