Speech Recognition
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
PythonBSD-2-Clauseupdated Aug 30, 2026
A list of publically available audio data that anyone can download for ASR or other speech activities
$ git clone https://github.com/robmsmt/ASR-Audio-Data-Links.gitWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型
Pytorch实现的流式与非流式的自动语音识别框架,同时兼容在线和离线识别,目前支持Conformer、Squeezeformer、DeepSpeech2模型,支持多种数据增强方法。
HuggingSound: A toolkit for speech-related tasks based on Hugging Face's tools
Data from GitHub · snapshot Sep 24, 2026