🏆 #1,909 overall#235 of 301 in Speech Recognition

Renovamen /Speech-and-Text

Speech to text (PocketSphinx, Iflytex API, Baidu API) and text to speech (pyttsx3) | 语音转文字(PocketSphinx、百度 API、科大讯飞 API)和文字转语音(pyttsx3)

$ git clone https://github.com/Renovamen/Speech-and-Text.git
GitHub social preview for Renovamen/Speech-and-Text
Stars
341
341
Forks
77
77
Language
Python
License
None
Created
Apr 15, 2019
7.4 years old
Last push
Jun 3, 2019

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

abus-aikorea/voice-pro

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

PythonGPL-3.0updated Jul 13, 2026
GitHub ↗★ 12.9K⑂ 1.9K
Speech Recognition

Blaizzy/mlx-audio🔥 active

A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.

PythonMITupdated Sep 23, 2026
GitHub ↗★ 7.9K⑂ 725
Speech Recognition

coqui-ai/open-speech-corpora

💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

OtherMITupdated Jun 6, 2024
GitHub ↗★ 1.4K⑂ 152
Speech Recognition

VRCWizard/TTS-Voice-Wizard

Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)

C#MITupdated Aug 19, 2026
GitHub ↗★ 808⑂ 81
Speech Recognition

mmpneo/curses

Speech to Text and KB input captions for OBS, VRChat, Twitch chat and Discord

TypeScriptAGPL-3.0updated Jun 18, 2024
GitHub ↗★ 727⑂ 50
Speech Recognition

toverainc/willow-inference-server

Open source, local, and self-hosted highly optimized language inference server supporting ASR/STT, TTS, and LLM across WebRTC, REST, and WS

PythonApache-2.0updated Feb 12, 2026
GitHub ↗★ 511⑂ 60

Data from GitHub · snapshot Sep 24, 2026