๐Ÿ† #2,032 overall#254 of 301 in Speech Recognition

Soul-AILab /SoulX-Transcriber

An end-to-end framework for multi-speaker transcription that jointly models who spoke, when, and what.

$ git clone https://github.com/Soul-AILab/SoulX-Transcriber.git
GitHub social preview for Soul-AILab/SoulX-Transcriber
Stars
297
297
Forks
16
16
Language
Python
License
Apache-2.0
Created
Jun 2, 2026
0.3 years old
Last push
Jun 22, 2026

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

FireRedTeam/FireRedASR

Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.

PythonApache-2.0updated Feb 25, 2026
GitHub โ†—โ˜… 2Kโ‘‚ 166
Speech Recognition

roothch/PreenCut

AI-Powered Video Retrieval & Clipping Tool

PythonMITupdated Aug 22, 2025
GitHub โ†—โ˜… 418โ‘‚ 70
Speech Recognition

m-bain/whisperX

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

PythonBSD-2-Clauseupdated Aug 30, 2026
GitHub โ†—โ˜… 24.2Kโ‘‚ 2.4K
Speech Recognition

PaddlePaddle/PaddleSpeech

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

PythonApache-2.0updated Aug 12, 2026
GitHub โ†—โ˜… 12.7Kโ‘‚ 2K
Speech Recognition

modelscope/FunClip

FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

PythonMITupdated Sep 16, 2026
GitHub โ†—โ˜… 6.3Kโ‘‚ 753

Data from GitHub ยท snapshot Sep 24, 2026