coqui-ai/open-speech-corpora
๐ A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.
$ git clone https://github.com/huawei-noah/Speech-Backbones.git๐ A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Kalliope is a framework that will help you to create your own personal assistant.
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
Data from GitHub ยท snapshot Sep 24, 2026