๐Ÿ† #1,701 overall#203 of 301 in Speech Recognition

YuanGongND /whisper-at

Code and Pretrained Models for Interspeech 2023 Paper "Whisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong Audio Event Taggers"

$ git clone https://github.com/YuanGongND/whisper-at.git
GitHub social preview for YuanGongND/whisper-at
Stars
426
426
Forks
36
36
Language
Python
License
BSD-2-Clause
Created
Jul 6, 2023
3.2 years old
Last push
Feb 21, 2024

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

YuanGongND/ltu

Code, Dataset, and Pretrained Models for Audio and Speech Large Language Model "Listen, Think, and Understand".

๐Ÿค– Language Models
Pythonno licenseupdated Apr 24, 2024
GitHub โ†—โ˜… 479โ‘‚ 41
Speech Recognition

pszemraj/vid2cleantxt

Python API & command-line tool to easily transcribe speech-based video files into clean text

๐Ÿง  Natural Language Processing
Jupyter NotebookApache-2.0updated Oct 29, 2024
GitHub โ†—โ˜… 229โ‘‚ 28
Speech Recognition

speechbrain/speechbrain

A PyTorch-based Speech Toolkit

๐Ÿค– Language Models
PythonApache-2.0updated Aug 27, 2026
GitHub โ†—โ˜… 11.8Kโ‘‚ 1.7K
Speech Recognition

Uberi/speech_recognition

Speech recognition module for Python, supporting several engines and APIs, online and offline.

PythonBSD-3-Clauseupdated Sep 2, 2026
GitHub โ†—โ˜… 9Kโ‘‚ 2.4K
Speech Recognition

Blaizzy/mlx-audio๐Ÿ”ฅ active

A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.

PythonMITupdated Sep 23, 2026
GitHub โ†—โ˜… 7.9Kโ‘‚ 725
Speech Recognition

huggingface/distil-whisper

Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.

PythonMITupdated Jan 8, 2025
GitHub โ†—โ˜… 4.1Kโ‘‚ 356

Data from GitHub ยท snapshot Sep 24, 2026