Speech Recognition
Picovoice/cheetah
On-device streaming speech-to-text engine powered by deep learning
PythonApache-2.0updated Sep 9, 2026
Obsidian plugin to create high-quality transcriptions from markdown linked audio files
$ git clone https://github.com/djmango/obsidian-transcription.gitOn-device streaming speech-to-text engine powered by deep learning
On-device speech-to-text engine powered by deep learning
Transcribe any video URL or audio file into plaintext. No GPU. No cloud. One command.
Local-first speech-to-text: no cloud, no telemetry, fail-closed by design. One CLI, seven model families, signed model catalog, OpenAI-compatible local API.
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
Data from GitHub ยท snapshot Sep 24, 2026