Speech Recognition
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
PythonBSD-2-Clauseupdated Aug 30, 2026
AI-Powered Video Retrieval & Clipping Tool
$ git clone https://github.com/roothch/PreenCut.gitWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
πΈSTT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering outstanding singing lyrics recognition capability.
Data from GitHub Β· snapshot Sep 24, 2026