coqui-ai/open-speech-corpora
π A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
πΊπ¦ Speech Recognition & Synthesis for Ukrainian
$ git clone https://github.com/egorsmkv/speech-recognition-uk.gitπ A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
AI Vtuber for Streaming on Youtube/Twitch
The J.A.R.V.I.S. Speech API is designed to be simple and efficient, using the speech engines created by Google to provide functionality for parts of the API. Essentially, it is an API written in Java, including a recognizer, synthesizer, and a microphone capture utility. The project uses Google services for the synthesizer and recognizer. While this requires an Internet connection, it provides a complete, modern, and fully functional speech API in Java.
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Data from GitHub Β· snapshot Sep 24, 2026