OpenBMB/UltraEval-Audio
Your faithful, impartial partner for audio evaluation โ know yourself, know your rivals. ็ๅฎ่ฏๆต๏ผ็ฅๅทฑ็ฅๅฝผใA unified benchmark framework for ASR/TTS/Audio Codec/audio LLM evaluation
Python module for evaluating ASR hypotheses (e.g. word error rate, word recognition rate).
$ git clone https://github.com/belambert/asr-evaluation.gitYour faithful, impartial partner for audio evaluation โ know yourself, know your rivals. ็ๅฎ่ฏๆต๏ผ็ฅๅทฑ็ฅๅฝผใA unified benchmark framework for ASR/TTS/Audio Codec/audio LLM evaluation
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
Production First and Production Ready End-to-End Speech Recognition Toolkit
Data from GitHub ยท snapshot Sep 24, 2026