๐Ÿ† #779 overall#63 of 301 in Speech Recognition

QwenAudio /Fun-ASR

Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

$ git clone https://github.com/QwenAudio/Fun-ASR.git
GitHub social preview for QwenAudio/Fun-ASR
Stars
1.6K
1,553
Forks
152
152
Language
C
License
Apache-2.0
Created
Dec 15, 2025
0.8 years old
Last push
Sep 10, 2026

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

QwenAudio/SenseVoice๐Ÿ”ฅ active

Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

CMITupdated Sep 22, 2026
GitHub โ†—โ˜… 9.4Kโ‘‚ 829
Speech Recognition

AudarAI/Audar-ASR-V1

Arabic-first generative speech recognition โ€” Audar-ASR-V1 (Flash + Turbo). #1 on the Open Universal Arabic ASR Leaderboard. Model cards, benchmarks & inference.

PythonApache-2.0updated Aug 22, 2026
GitHub โ†—โ˜… 505โ‘‚ 4
Speech Recognition

modelscope/FunASR๐Ÿ”ฅ active

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

PythonMITupdated Sep 24, 2026
GitHub โ†—โ˜… 20.5Kโ‘‚ 2K
Speech Recognition

modelscope/FunClip

FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

PythonMITupdated Sep 16, 2026
GitHub โ†—โ˜… 6.3Kโ‘‚ 753
Speech Recognition

TheDeathDragon/LiveTranslate

Real-time audio translation, captures system audio + mic, runs ASR (Whisper/SenseVoice), translates via LLM API with streaming display. Perfect for VTubers, livestreamers, and watching foreign content. Windows ๅฎžๆ—ถ้Ÿณ้ข‘็ฟป่ฏ‘๏ผŒASR ่ฏญ้Ÿณ่ฏ†ๅˆซๅŽ LLM ๆตๅผ็ฟป่ฏ‘ๆ˜พ็คบ๏ผŒ้€‚ๅˆ VTuberใ€ไธปๆ’ญๅ’Œๅค–่ฏญ่ง†้ข‘่ง‚็œ‹ใ€‚

PythonMITupdated Aug 17, 2026
GitHub โ†—โ˜… 692โ‘‚ 53
Speech Recognition

crosswk/SayIt๐Ÿ”ฅ active

Open-source voice typing for Windows โ€” a Wispr Flow / Superwhisper alternative. Press a shortcut, speak, and AI-polished text lands at your cursor. Local models, your own API keys, or a self-hosted backend.

TypeScriptAGPL-3.0updated Sep 24, 2026
GitHub โ†—โ˜… 421โ‘‚ 63

Data from GitHub ยท snapshot Sep 24, 2026