πŸ† #1,659 overall#195 of 301 in Speech Recognition

inclusionAI /Ming-UniAudio

Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation

$ git clone https://github.com/inclusionAI/Ming-UniAudio.git
GitHub social preview for inclusionAI/Ming-UniAudio
Stars
457
457
Forks
30
30
Language
Python
License
MIT
Created
Sep 29, 2025
1.0 years old
Last push
Nov 27, 2025

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

lkuza2/java-speech-api

The J.A.R.V.I.S. Speech API is designed to be simple and efficient, using the speech engines created by Google to provide functionality for parts of the API. Essentially, it is an API written in Java, including a recognizer, synthesizer, and a microphone capture utility. The project uses Google services for the synthesizer and recognizer. While this requires an Internet connection, it provides a complete, modern, and fully functional speech API in Java.

JavaGPL-3.0updated May 2, 2019
GitHub β†—β˜… 544β‘‚ 289
Speech Recognition

echogarden-project/echogarden

Cross-platform speech toolset, used from the command-line or as a Node.js library. Includes a variety of engines for speech synthesis, speech recognition, forced alignment, speech translation, voice isolation, language detection and more.

TypeScriptno licenseupdated Sep 8, 2026
GitHub β†—β˜… 451β‘‚ 44
Speech Recognition

VITA-MLLM/Freeze-Omni

✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM

πŸ€– Language Models
PythonOtherupdated May 27, 2025
GitHub β†—β˜… 398β‘‚ 31
Speech Recognition

m-bain/whisperX

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

PythonBSD-2-Clauseupdated Aug 30, 2026
GitHub β†—β˜… 24.2Kβ‘‚ 2.4K
Speech Recognition

kaldi-asr/kaldi

kaldi-asr/kaldi is the official location of the Kaldi project.

ShellOtherupdated Sep 22, 2025
GitHub β†—β˜… 15.5Kβ‘‚ 5.4K

Data from GitHub Β· snapshot Sep 24, 2026