πŸ† #2,104 overall#265 of 301 in Speech Recognition

m3hrdadfi /soxan

Wav2Vec for speech recognition, classification, and audio classification

$ git clone https://github.com/m3hrdadfi/soxan.git
GitHub social preview for m3hrdadfi/soxan
Stars
277
277
Forks
37
37
Language
Jupyter Notebook
License
Apache-2.0
Created
May 25, 2021
5.3 years old
Last push
Apr 2, 2022

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

QuentinFuxa/WhisperLiveKitπŸ”₯ active

Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.

PythonApache-2.0updated Sep 21, 2026
GitHub β†—β˜… 11.1Kβ‘‚ 1.1K
Speech Recognition

wenet-e2e/wenet

Production First and Production Ready End-to-End Speech Recognition Toolkit

PythonApache-2.0updated Sep 7, 2026
GitHub β†—β˜… 5.2Kβ‘‚ 1.2K
Speech Recognition

coqui-ai/STT

🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

C++MPL-2.0updated Mar 11, 2024
GitHub β†—β˜… 2.6Kβ‘‚ 295
Speech Recognition

OpenMOSS/MOSS-Transcribe-Diarize

A 0.9B model for long-form transcription in 50+ languages with speaker diarization, timestamps, and acoustic event awareness

PythonApache-2.0updated Sep 15, 2026
GitHub β†—β˜… 2.1Kβ‘‚ 127
Speech Recognition

FireRedTeam/FireRedASR

Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.

PythonApache-2.0updated Feb 25, 2026
GitHub β†—β˜… 2Kβ‘‚ 166

Data from GitHub Β· snapshot Sep 24, 2026