๐Ÿ† #2,291 overall#291 of 301 in Speech Recognition

pszemraj /vid2cleantxt

Python API & command-line tool to easily transcribe speech-based video files into clean text

$ git clone https://github.com/pszemraj/vid2cleantxt.git
GitHub social preview for pszemraj/vid2cleantxt
Stars
229
229
Forks
28
28
Language
Jupyter Notebook
License
Apache-2.0
Created
Mar 9, 2021
5.5 years old
Last push
Oct 29, 2024

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

lhotse-speech/lhotse๐Ÿ”ฅ active

Tools for handling multimodal data in machine learning projects.

PythonApache-2.0updated Sep 22, 2026
GitHub โ†—โ˜… 1.2Kโ‘‚ 281
Speech Recognition

Uberi/speech_recognition

Speech recognition module for Python, supporting several engines and APIs, online and offline.

PythonBSD-3-Clauseupdated Sep 2, 2026
GitHub โ†—โ˜… 9Kโ‘‚ 2.4K
Speech Recognition

julius-speech/julius

Open-Source Large Vocabulary Continuous Speech Recognition Engine

CBSD-3-Clauseupdated Jun 16, 2025
GitHub โ†—โ˜… 1.9Kโ‘‚ 303
Speech Recognition

nyrahealth/CrisperWhisper๐Ÿ”ฅ active

Controllable Transcription. Verbatim ( every, filler, pause, stutter, vocal sound) , or intended ( what the speaker meant to say, optimized for readability) with word-level timestamps.

PythonOtherupdated Sep 22, 2026
GitHub โ†—โ˜… 1.4Kโ‘‚ 92
Speech Recognition

pykaldi/pykaldi

A Python wrapper for Kaldi

๐Ÿค– Language Models
PythonApache-2.0updated Nov 30, 2025
GitHub โ†—โ˜… 1Kโ‘‚ 247
Speech Recognition

DemisEom/SpecAugment

A Implementation of SpecAugment with Tensorflow & Pytorch, introduced by Google Brain

PythonApache-2.0updated Apr 5, 2022
GitHub โ†—โ˜… 655โ‘‚ 133

Data from GitHub ยท snapshot Sep 24, 2026