๐Ÿ† #68 overall#5 of 301 in Speech Recognition๐Ÿ”ฅ active this week

modelscope /FunASR

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

$ git clone https://github.com/modelscope/FunASR.git
GitHub social preview for modelscope/FunASR
Stars
20.5K
20,493
Forks
2K
2,043
Language
Python
License
MIT
Created
Nov 24, 2022
3.8 years old
Last push
Sep 24, 2026
๐Ÿ”ฅ this week

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

modelscope/FunClip

FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

PythonMITupdated Sep 16, 2026
GitHub โ†—โ˜… 6.3Kโ‘‚ 753
Speech Recognition

speechbrain/speechbrain

A PyTorch-based Speech Toolkit

๐Ÿค– Language Models
PythonApache-2.0updated Aug 27, 2026
GitHub โ†—โ˜… 11.8Kโ‘‚ 1.7K
Speech Recognition

QwenAudio/SenseVoice๐Ÿ”ฅ active

Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

CMITupdated Sep 22, 2026
GitHub โ†—โ˜… 9.4Kโ‘‚ 829
Speech Recognition

QwenAudio/Fun-ASR

Fun-ASR speech recognition models, with native Hugging Face Transformers support for Fun-ASR-Nano and separate FunASR, vLLM and llama.cpp deployment paths.

CApache-2.0updated Sep 10, 2026
GitHub โ†—โ˜… 1.6Kโ‘‚ 152
Speech Recognition

nyrahealth/CrisperWhisper๐Ÿ”ฅ active

Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.

PythonOtherupdated Sep 22, 2026
GitHub โ†—โ˜… 1.4Kโ‘‚ 92
Speech Recognition

mravanelli/SincNet

SincNet is a neural architecture for efficiently processing raw audio samples.

PythonMITupdated Apr 28, 2021
GitHub โ†—โ˜… 1.2Kโ‘‚ 273

Data from GitHub ยท snapshot Sep 24, 2026