πŸ† #2,234 overall#285 of 301 in Speech Recognition

smeetrs /deep_avsr

A PyTorch implementation of the Deep Audio-Visual Speech Recognition paper.

$ git clone https://github.com/smeetrs/deep_avsr.git
GitHub social preview for smeetrs/deep_avsr
Stars
244
244
Forks
42
42
Language
Python
License
MIT
Created
Dec 7, 2019
6.8 years old
Last push
Feb 15, 2024

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

QuentinFuxa/WhisperLiveKitπŸ”₯ active

Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.

PythonApache-2.0updated Sep 21, 2026
GitHub β†—β˜… 11.1Kβ‘‚ 1.1K
Speech Recognition

coqui-ai/STT

🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

C++MPL-2.0updated Mar 11, 2024
GitHub β†—β˜… 2.6Kβ‘‚ 295
Speech Recognition

TensorSpeech/TensorFlowASRπŸ”₯ active

:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords

PythonApache-2.0updated Sep 24, 2026
GitHub β†—β˜… 1Kβ‘‚ 236
Speech Recognition

Picovoice/cheetah

On-device streaming speech-to-text engine powered by deep learning

PythonApache-2.0updated Sep 9, 2026
GitHub β†—β˜… 671β‘‚ 77
Speech Recognition

Picovoice/leopard

On-device speech-to-text engine powered by deep learning

PythonApache-2.0updated Sep 9, 2026
GitHub β†—β˜… 485β‘‚ 28

Data from GitHub Β· snapshot Sep 24, 2026