πŸ† #1,331 overall#146 of 301 in Speech Recognition

FireRedTeam /FireRedASR2S

A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switching, and both speech and singing ASR. FireRedVAD supports speech/singing/music in 100+ langs. FireRedLID supports 100+ langs and 20+ zh dialects. FireRedPunc supports zh and en.

$ git clone https://github.com/FireRedTeam/FireRedASR2S.git
GitHub social preview for FireRedTeam/FireRedASR2S
Stars
694
694
Forks
46
46
Language
Python
License
Apache-2.0
Created
Feb 12, 2026
0.6 years old
Last push
Jun 2, 2026

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

FireRedTeam/FireRedASR

Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.

PythonApache-2.0updated Feb 25, 2026
GitHub β†—β˜… 2Kβ‘‚ 166
Speech Recognition

QwenAudio/SenseVoiceπŸ”₯ active

Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

CMITupdated Sep 22, 2026
GitHub β†—β˜… 9.4Kβ‘‚ 829
Speech Recognition

modelscope/FunClip

FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

PythonMITupdated Sep 16, 2026
GitHub β†—β˜… 6.3Kβ‘‚ 753
Speech Recognition

wenet-e2e/wenet

Production First and Production Ready End-to-End Speech Recognition Toolkit

PythonApache-2.0updated Sep 7, 2026
GitHub β†—β˜… 5.2Kβ‘‚ 1.2K
Speech Recognition

coqui-ai/STT

🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

C++MPL-2.0updated Mar 11, 2024
GitHub β†—β˜… 2.6Kβ‘‚ 295

Data from GitHub Β· snapshot Sep 24, 2026