Blaizzy/mlx-audioπ₯ active
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
$ git clone https://github.com/abus-aikorea/voice-pro.gitA text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
π A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
πΊπ¦ Speech Recognition & Synthesis for Ukrainian
Auto transcribe tool based on whisper
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!
Data from GitHub Β· snapshot Sep 24, 2026