๐Ÿ† #1,536 overall#177 of 301 in Speech Recognition

DmitryRyumin /ICASSP-2023-24-Papers

ICASSP 2023-2024 Papers: A complete collection of influential and exciting research papers from the ICASSP 2023-24 conferences. Explore the latest advancements in acoustics, speech and signal processing. Code included. Star the repository to support the advancement of audio and signal processing!

$ git clone https://github.com/DmitryRyumin/ICASSP-2023-24-Papers.git
GitHub social preview for DmitryRyumin/ICASSP-2023-24-Papers
Stars
527
527
Forks
23
23
Language
Python
License
MIT
Created
Aug 1, 2023
3.1 years old
Last push
May 5, 2025

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

m-bain/whisperX

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

PythonBSD-2-Clauseupdated Aug 30, 2026
GitHub โ†—โ˜… 24.2Kโ‘‚ 2.4K
Speech Recognition

modelscope/FunASR๐Ÿ”ฅ active

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

PythonMITupdated Sep 24, 2026
GitHub โ†—โ˜… 20.5Kโ‘‚ 2K
Speech Recognition

alphacep/vosk-api

Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

Jupyter NotebookApache-2.0updated Aug 9, 2026
GitHub โ†—โ˜… 15.1Kโ‘‚ 1.8K
Speech Recognition

PaddlePaddle/PaddleSpeech

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

PythonApache-2.0updated Aug 12, 2026
GitHub โ†—โ˜… 12.7Kโ‘‚ 2K
Speech Recognition

speechbrain/speechbrain

A PyTorch-based Speech Toolkit

๐Ÿค– Language Models
PythonApache-2.0updated Aug 27, 2026
GitHub โ†—โ˜… 11.8Kโ‘‚ 1.7K
Speech Recognition

QwenAudio/SenseVoice๐Ÿ”ฅ active

Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.

CMITupdated Sep 22, 2026
GitHub โ†—โ˜… 9.4Kโ‘‚ 829

Data from GitHub ยท snapshot Sep 24, 2026