Speech Recognition
lhotse-speech/lhotseπ₯ active
Tools for handling multimodal data in machine learning projects.
PythonApache-2.0updated Sep 22, 2026
A Implementation of SpecAugment with Tensorflow & Pytorch, introduced by Google Brain
$ git clone https://github.com/DemisEom/SpecAugment.gitTools for handling multimodal data in machine learning projects.
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
A Deep-Learning-Based Chinese Speech Recognition System εΊδΊζ·±εΊ¦ε¦δΉ ηδΈζθ―ι³θ―ε«η³»η»
Machine Learning and Agentic AI Resources, Practice and Research
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
Data from GitHub Β· snapshot Sep 24, 2026