TheStageAI/TheWhisper๐ฅ active
Optimized Whisper models for streaming and on-device use
On-device, real-time multimodal AI with features similar to GPT-Live
$ git clone https://github.com/fikrikarim/parlor.gitOptimized Whisper models for streaming and on-device use
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning โ natively on MLX. Unsloth-compatible API.
AI speech toolkit for Apple Silicon โ ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
A high-performance API server that provides OpenAI-compatible endpoints for MLX models. Developed using Python and powered by the FastAPI framework, it provides an efficient, scalable, and user-friendly solution for running MLX-based vision and language models locally with an OpenAI-compatible interface.
Data from GitHub ยท snapshot Sep 24, 2026