๐Ÿ† #166 overall#17 of 301 in Speech Recognition๐Ÿ”ฅ active this week

QwenAudio /SenseVoice

Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

$ git clone https://github.com/QwenAudio/SenseVoice.git
GitHub social preview for QwenAudio/SenseVoice
Stars
9.4K
9,380
Forks
829
829
Language
C
License
MIT
Created
Jul 3, 2024
2.2 years old
Last push
Sep 22, 2026
๐Ÿ”ฅ this week

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

QwenAudio/Fun-ASR

Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

CApache-2.0updated Sep 10, 2026
GitHub โ†—โ˜… 1.6Kโ‘‚ 152
Speech Recognition

FireRedTeam/FireRedASR2S

A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switching, and both speech and singing ASR. FireRedVAD supports speech/singing/music in 100+ langs. FireRedLID supports 100+ langs and 20+ zh dialects. FireRedPunc supports zh and en.

PythonApache-2.0updated Jun 2, 2026
GitHub โ†—โ˜… 694โ‘‚ 46
Speech Recognition

modelscope/FunASR๐Ÿ”ฅ active

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

PythonMITupdated Sep 24, 2026
GitHub โ†—โ˜… 20.5Kโ‘‚ 2K
Speech Recognition

modelscope/FunClip

FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

PythonMITupdated Sep 16, 2026
GitHub โ†—โ˜… 6.3Kโ‘‚ 753
Speech Recognition

TheDeathDragon/LiveTranslate

Real-time audio translation, captures system audio + mic, runs ASR (Whisper/SenseVoice), translates via LLM API with streaming display. Perfect for VTubers, livestreamers, and watching foreign content. Windows ๅฎžๆ—ถ้Ÿณ้ข‘็ฟป่ฏ‘๏ผŒASR ่ฏญ้Ÿณ่ฏ†ๅˆซๅŽ LLM ๆตๅผ็ฟป่ฏ‘ๆ˜พ็คบ๏ผŒ้€‚ๅˆ VTuberใ€ไธปๆ’ญๅ’Œๅค–่ฏญ่ง†้ข‘่ง‚็œ‹ใ€‚

PythonMITupdated Aug 17, 2026
GitHub โ†—โ˜… 692โ‘‚ 53
Speech Recognition

AudarAI/Audar-ASR-V1

Arabic-first generative speech recognition โ€” Audar-ASR-V1 (Flash + Turbo). #1 on the Open Universal Arabic ASR Leaderboard. Model cards, benchmarks & inference.

PythonApache-2.0updated Aug 22, 2026
GitHub โ†—โ˜… 505โ‘‚ 4

Data from GitHub ยท snapshot Sep 24, 2026