Speech Recognition
kaldi-asr/kaldi
kaldi-asr/kaldi is the official location of the Kaldi project.
ShellOtherupdated Sep 22, 2025
UniSpeech - Large Scale Self-Supervised Learning for Speech
$ git clone https://github.com/microsoft/UniSpeech.gitkaldi-asr/kaldi is the official location of the Kaldi project.
Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.
๐ A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
Tools for handling multimodal data in machine learning projects.
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
Allosaurus is a pretrained universal phone recognizer for more than 2000 languages
Data from GitHub ยท snapshot Sep 24, 2026