๐Ÿ† #785 overall#64 of 301 in Speech Recognition

bytedance /SALMONN

SALMONN family: A suite of advanced multi-modal LLMs

$ git clone https://github.com/bytedance/SALMONN.git
GitHub social preview for bytedance/SALMONN
Stars
1.5K
1,536
Forks
125
125
Language
Other
License
Apache-2.0
Created
Aug 11, 2023
3.1 years old
Last push
Aug 24, 2026

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

YuanGongND/ltu

Code, Dataset, and Pretrained Models for Audio and Speech Large Language Model "Listen, Think, and Understand".

๐Ÿค– Language Models
Pythonno licenseupdated Apr 24, 2024
GitHub โ†—โ˜… 479โ‘‚ 41
Speech Recognition

speechbrain/speechbrain

A PyTorch-based Speech Toolkit

๐Ÿค– Language Models
PythonApache-2.0updated Aug 27, 2026
GitHub โ†—โ˜… 11.8Kโ‘‚ 1.7K
Speech Recognition

mravanelli/SincNet

SincNet is a neural architecture for efficiently processing raw audio samples.

PythonMITupdated Apr 28, 2021
GitHub โ†—โ˜… 1.2Kโ‘‚ 273
Speech Recognition

YuanGongND/whisper-at

Code and Pretrained Models for Interspeech 2023 Paper "Whisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong Audio Event Taggers"

PythonBSD-2-Clauseupdated Feb 21, 2024
GitHub โ†—โ˜… 426โ‘‚ 36
Speech Recognition

pszemraj/vid2cleantxt

Python API & command-line tool to easily transcribe speech-based video files into clean text

๐Ÿง  Natural Language Processing
Jupyter NotebookApache-2.0updated Oct 29, 2024
GitHub โ†—โ˜… 229โ‘‚ 28
Speech Recognition

NVIDIA/DeepLearningExamples

State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.

๐Ÿค– Language Models๐Ÿง  Natural Language Processing
Jupyter Notebookno licenseupdated Aug 12, 2024
GitHub โ†—โ˜… 14.9Kโ‘‚ 3.4K

Data from GitHub ยท snapshot Sep 24, 2026