Speech Recognition
jonatasgrosman/huggingsound
HuggingSound: A toolkit for speech-related tasks based on Hugging Face's tools
PythonMITupdated Sep 20, 2023
The dataset of Speech Recognition
$ git clone https://github.com/double22a/speech_dataset.gitHuggingSound: A toolkit for speech-related tasks based on Hugging Face's tools
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.
Tools for handling multimodal data in machine learning projects.
基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型
Data from GitHub · snapshot Sep 24, 2026