Speech Recognition
jonatasgrosman/huggingsound
HuggingSound: A toolkit for speech-related tasks based on Hugging Face's tools
PythonMITupdated Sep 20, 2023
OpenAI Whisper ASR Webservice API
$ git clone https://github.com/ahmetoner/whisper-asr-webservice.gitHuggingSound: A toolkit for speech-related tasks based on Hugging Face's tools
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Ultra fast and portable Parakeet implementation for on-device inference in C++ using Axiom with MPS+Unified Memory
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
πΈSTT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
Data from GitHub Β· snapshot Sep 24, 2026