๐Ÿ† #1,195 overall#121 of 301 in Speech Recognition

salute-developers /GigaAM

Foundational Model for Speech Recognition Tasks

$ git clone https://github.com/salute-developers/GigaAM.git
GitHub social preview for salute-developers/GigaAM
Stars
826
826
Forks
107
107
Language
Python
License
MIT
Created
Apr 1, 2024
2.5 years old
Last push
Aug 17, 2026

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

PaddlePaddle/PaddleSpeech

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

PythonApache-2.0updated Aug 12, 2026
GitHub โ†—โ˜… 12.7Kโ‘‚ 2K
Speech Recognition

m3hrdadfi/soxan

Wav2Vec for speech recognition, classification, and audio classification

Jupyter NotebookApache-2.0updated Apr 2, 2022
GitHub โ†—โ˜… 277โ‘‚ 37
Speech Recognition

ASR-project/Multilingual-PR

Phoneme Recognition using pre-trained models Wav2vec2, HuBERT and WavLM. Throughout this project, we compared specifically three different self-supervised models, Wav2vec (2019, 2020), HuBERT (2021) and WavLM (2022) pretrained on a corpus of English speech that we will use in various ways to perform phoneme recognition for different languages with a network trained with Connectionist Temporal Classification (CTC) algorithm.

Pythonno licenseupdated May 9, 2022
GitHub โ†—โ˜… 267โ‘‚ 24
Speech Recognition

ggml-org/whisper.cpp๐Ÿ”ฅ active

Port of OpenAI's Whisper model in C/C++

C++MITupdated Sep 24, 2026
GitHub โ†—โ˜… 53.9Kโ‘‚ 6.2K

Data from GitHub ยท snapshot Sep 24, 2026