Speech Recognition
speechbrain/speechbrain
A PyTorch-based Speech Toolkit
๐ค Language Models
PythonApache-2.0updated Aug 27, 2026
Code, Dataset, and Pretrained Models for Audio and Speech Large Language Model "Listen, Think, and Understand".
$ git clone https://github.com/YuanGongND/ltu.gitA PyTorch-based Speech Toolkit
SALMONN family: A suite of advanced multi-modal LLMs
SincNet is a neural architecture for efficiently processing raw audio samples.
Tools for handling multimodal data in machine learning projects.
Code and Pretrained Models for Interspeech 2023 Paper "Whisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong Audio Event Taggers"
Data from GitHub ยท snapshot Sep 24, 2026