Language Models
huggingface/speech-to-speechπ₯ active
Build voice agents with open-source models
PythonApache-2.0updated Sep 23, 2026
An Open-Sourced LLM-empowered Foundation TTS System
$ git clone https://github.com/FireRedTeam/FireRedTTS.gitBuild voice agents with open-source models
VibeVoiceFusion is a full-stack, multi-speaker voice generation web system featuring LoRA fine-tuning, batch generation, and VRAM optimization. Based on Microsoft's VibeVoice (AR + diffusion architecture)
PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo https://lifeiteng.github.io/valle/index.html
A multilingual model for long-form, multi-speaker dialogue synthesis with flexible speaker control and zero-shot voice cloning
A Python wrapper for Kaldi
β¨β¨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
Data from GitHub Β· snapshot Sep 24, 2026