Speech Recognition
Uberi/speech_recognition
Speech recognition module for Python, supporting several engines and APIs, online and offline.
PythonBSD-3-Clauseupdated Sep 2, 2026
Word-accurate timestamps for Qur'anic audio.
$ git clone https://github.com/cpfair/quran-align.gitSpeech recognition module for Python, supporting several engines and APIs, online and offline.
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
A 0.9B model for long-form transcription in 50+ languages with speaker diarization, timestamps, and acoustic event awareness
Cross-Platform, GPU Accelerated Whisper ๐๏ธ
Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.
Tools for handling multimodal data in machine learning projects.
Data from GitHub ยท snapshot Sep 24, 2026