MahmoudAshraf97/whisper-diarization
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml. TTS, ASR/STT, VAD, voice conversion, speaker diarization, music generation. No Python dependency.
$ git clone https://github.com/kigner/audio.cpp-webui.gitAutomatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
HuggingSound: A toolkit for speech-related tasks based on Hugging Face's tools
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Ultra fast and portable Parakeet implementation for on-device inference in C++ using Axiom with MPS+Unified Memory
Data from GitHub ยท snapshot Sep 24, 2026