Speech Recognition
🏆 #1,744 overall#211 of 301 in Speech Recognition🔥 active this week
agan-j /xiaoniu
小牛视频翻译 是一款支持本地视频翻译、字幕翻译和 YouTube 视频翻译下载的 AI 工具,集成自动语音识别与多语言翻译功能,助力创作者高效完成视频翻译,应用于视频本地化与视频出海场景。
$ git clone https://github.com/agan-j/xiaoniu.gitStars
408
408
Forks
23
23
Language
Other
License
None
Created
Mar 1, 2024
2.6 years old
Last push
Sep 23, 2026
🔥 this week
Categories
More in Speech Recognition
Speech Recognition
deepgram/deepgram-js-sdk🔥 active
Official JavaScript SDK for Deepgram.
TypeScriptMITupdated Sep 23, 2026
Speech Recognition
tomchang25/whisper-auto-transcribe
Auto transcribe tool based on whisper
🤖 Language Models
PythonMITupdated May 5, 2026
Speech Recognition
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
PythonBSD-2-Clauseupdated Aug 30, 2026
Speech Recognition
abus-aikorea/voice-pro
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
PythonGPL-3.0updated Jul 13, 2026
Speech Recognition
Blaizzy/mlx-audio🔥 active
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
PythonMITupdated Sep 23, 2026
Data from GitHub · snapshot Sep 24, 2026