wenet-e2e/wenet
Production First and Production Ready End-to-End Speech Recognition Toolkit
Fine-tune and evaluate Whisper models for Automatic Speech Recognition (ASR) on custom datasets or datasets from huggingface.
$ git clone https://github.com/vasistalodagala/whisper-finetune.gitProduction First and Production Ready End-to-End Speech Recognition Toolkit
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inference and support Web deployment, Windows desktop deployment, and Android deployment
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
Data from GitHub ยท snapshot Sep 24, 2026