πŸ† #186 overall#20 of 301 in Speech RecognitionπŸ”₯ active this week

Blaizzy /mlx-audio

A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.

$ git clone https://github.com/Blaizzy/mlx-audio.git
GitHub social preview for Blaizzy/mlx-audio
Stars
7.9K
7,942
Forks
725
725
Language
Python
License
MIT
Created
Nov 27, 2024
1.8 years old
Last push
Sep 23, 2026
πŸ”₯ this week

Categories

GitHub topics

More in Speech Recognition

Speech Recognition

abus-aikorea/voice-pro

Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

PythonGPL-3.0updated Jul 13, 2026
GitHub β†—β˜… 12.9Kβ‘‚ 1.9K
Speech Recognition

coqui-ai/open-speech-corpora

πŸ’Ž A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

OtherMITupdated Jun 6, 2024
GitHub β†—β˜… 1.4Kβ‘‚ 152
Speech Recognition

TheStageAI/TheWhisperπŸ”₯ active

Optimized Whisper models for streaming and on-device use

PythonMITupdated Sep 23, 2026
GitHub β†—β˜… 897β‘‚ 56
Speech Recognition

davidmartinrius/speech-dataset-generator

πŸ”Š Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. 🎧πŸ‘₯πŸ“Š Advanced audio processing.

PythonMITupdated Jun 10, 2024
GitHub β†—β˜… 263β‘‚ 27
Speech Recognition

fikrikarim/parlor

On-device, real-time multimodal AI with features similar to GPT-Live

PythonApache-2.0updated Aug 3, 2026
GitHub β†—β˜… 2.1Kβ‘‚ 272

Data from GitHub Β· snapshot Sep 24, 2026