🏆 #1,632 overall#190 of 301 in Speech Recognition

double22a /speech_dataset

The dataset of Speech Recognition

$ git clone https://github.com/double22a/speech_dataset.git
GitHub social preview for double22a/speech_dataset
Stars
469
469
Forks
81
81
Language
Other
License
Apache-2.0
Created
Apr 7, 2021
5.5 years old
Last push
Jan 4, 2026

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

coqui-ai/STT

🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

C++MPL-2.0updated Mar 11, 2024
GitHub ↗★ 2.6K⑂ 295
Speech Recognition

nyrahealth/CrisperWhisper🔥 active

Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.

PythonOtherupdated Sep 22, 2026
GitHub ↗★ 1.4K⑂ 92
Speech Recognition

lhotse-speech/lhotse🔥 active

Tools for handling multimodal data in machine learning projects.

PythonApache-2.0updated Sep 22, 2026
GitHub ↗★ 1.2K⑂ 281
Speech Recognition

yeyupiaoling/PPASR

基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型

PythonApache-2.0updated Dec 17, 2025
GitHub ↗★ 869⑂ 129

Data from GitHub · snapshot Sep 24, 2026