🏆 #1,065 overall#104 of 301 in Speech Recognition🔥 active this week

BinWang28 /audio-ai-hub

The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.

$ git clone https://github.com/BinWang28/audio-ai-hub.git
GitHub social preview for BinWang28/audio-ai-hub
Stars
958
958
Forks
54
54
Language
Python
License
None
Created
Jun 15, 2024
2.3 years old
Last push
Sep 21, 2026
🔥 this week

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

semperai/amica🔥 active

Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.

TypeScriptMITupdated Sep 22, 2026
GitHub ↗★ 1.6K⑂ 266
Speech Recognition

coqui-ai/open-speech-corpora

💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

OtherMITupdated Jun 6, 2024
GitHub ↗★ 1.4K⑂ 152
Speech Recognition

sgl-project/sglang-omni🔥 active

SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.

PythonApache-2.0updated Sep 24, 2026
GitHub ↗★ 1.3K⑂ 522
Speech Recognition

janvarev/Irene-Voice-Assistant

Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.

PythonOtherupdated Jul 26, 2026
GitHub ↗★ 1.2K⑂ 151
Speech Recognition

athena-team/athena

an open-source implementation of sequence-to-sequence based speech processing engine

C++Apache-2.0updated Dec 2, 2022
GitHub ↗★ 967⑂ 197

Data from GitHub · snapshot Sep 24, 2026