πŸ† #118 overall#11 of 301 in Speech Recognition

abus-aikorea /voice-pro

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

$ git clone https://github.com/abus-aikorea/voice-pro.git
GitHub social preview for abus-aikorea/voice-pro
Stars
12.9K
12,920
Forks
1.9K
1,868
Language
Python
License
GPL-3.0
Created
Jul 29, 2024
2.2 years old
Last push
Jul 13, 2026

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

Blaizzy/mlx-audioπŸ”₯ active

A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.

PythonMITupdated Sep 23, 2026
GitHub β†—β˜… 7.9Kβ‘‚ 725
Speech Recognition

Purfview/whisper-standalone-win

Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.

Otherno licenseupdated Nov 7, 2025
GitHub β†—β˜… 3.2Kβ‘‚ 166
Speech Recognition

coqui-ai/open-speech-corpora

πŸ’Ž A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

OtherMITupdated Jun 6, 2024
GitHub β†—β˜… 1.4Kβ‘‚ 152
Speech Recognition

pluja/whishper

Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!

SvelteAGPL-3.0updated Jul 31, 2026
GitHub β†—β˜… 3.1Kβ‘‚ 181

Data from GitHub Β· snapshot Sep 24, 2026