gooofy/zamia-speech
Open tools and data for cloudless automatic speech recognition
Pronunciation lexicon covering both English and Chinese languages for Automatic Speech Recognition.
$ git clone https://github.com/speechio/BigCiDian.gitOpen tools and data for cloudless automatic speech recognition
ๅฐ็่ง้ข็ฟป่ฏ ๆฏไธๆฌพๆฏๆๆฌๅฐ่ง้ข็ฟป่ฏใๅญๅน็ฟป่ฏๅ YouTube ่ง้ข็ฟป่ฏไธ่ฝฝ็ AI ๅทฅๅ ท๏ผ้ๆ่ชๅจ่ฏญ้ณ่ฏๅซไธๅค่ฏญ่จ็ฟป่ฏๅ่ฝ๏ผๅฉๅๅไฝ่ ้ซๆๅฎๆ่ง้ข็ฟป่ฏ๏ผๅบ็จไบ่ง้ขๆฌๅฐๅไธ่ง้ขๅบๆตทๅบๆฏใ
ๆฉ่ณ - Real-time multilingual speech-to-text on CPU only. Live subtitles, browser dashboard, speaker labels, translation. No GPU, no cloud.
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.
Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
Data from GitHub ยท snapshot Sep 24, 2026