tts

56 projects share this GitHub topic

tts — unsloth ★75.4kttsGPT-SoVITS — ★59.9kLocalAI — ★48.8kChatTTS — ★39.8kOpenVoice — ★37kmastra — ★27.6kPixelle-Video — ★25.7kSpeech — ★18.4kedge-tts — ★11.5kKrillinAI — ★11.3kvoltagent — ★10.5kBert-VITS2 — ★8.8kVision-Agents — ★8.1kStyleTTS2 — ★6.3kOrpheus-TTS — ★6.3kabogen — ★5.8kdramaclaw — ★4.8kStreamer-Sales — ★3.8kTTS-WebUI — ★3.2kawesome-speech-recognition-speech-synthesis-papers — ★3.1kvall-e — ★3kStable-Diffusion — ★2.8kmarytts — ★2.6kReadAny — ★2.3kvall-e — ★2.2kopenai-edge-tts — ★2.1kElatoAI — ★1.9kDot — ★1.9kbailing — ★1.8kuzu — ★1.7kvox-director — ★1.7kava-whatsapp-agent-course — ★1.7kGenie-TTS — ★1.7kamica — ★1.6kBrowserAI — ★1.4kSpeech-AI-Forge — ★1.4ksoprano — ★1.2kXZVoice — ★1.2kIrene-Voice-Assistant — ★1.1kepub2tts — ★956MimikaStudio — ★730GPT-SoVITS★ 59.9kLocalAI★ 48.8kChatTTS★ 39.8kOpenVoice★ 37kmastra★ 27.6kPixelle-Video★ 25.7kSpeech★ 18.4kedge-tts★ 11.5kKrillinAI★ 11.3kvoltagent★ 10.5kBert-VITS2★ 8.8kVision-Agents★ 8.1kStyleTTS2★ 6.3kOrpheus-TTS★ 6.3kabogen★ 5.8kdramaclaw★ 4.8kStreamer-Sales★ 3.8kTTS-WebUI★ 3.2kawesome-speech-recogniti…★ 3.1kvall-e★ 3kStable-Diffusion★ 2.8kmarytts★ 2.6kReadAny★ 2.3kvall-e★ 2.2kopenai-edge-tts★ 2.1kElatoAI★ 1.9kDot★ 1.9kbailing★ 1.8kuzu★ 1.7kvox-director★ 1.7kava-whatsapp-agent-cours…★ 1.7kGenie-TTS★ 1.7kamica★ 1.6kBrowserAI★ 1.4kSpeech-AI-Forge★ 1.4ksoprano★ 1.2kXZVoice★ 1.2kIrene-Voice-Assistant★ 1.1kepub2tts★ 956MimikaStudio★ 730

Lines connect members that are measurably related to each other. Dot size reflects stars.

🧬 Members
unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, Kimi K3, MiniMax-H3, Gemma…
★ 75.4k
GPT-SoVITS
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
★ 59.9k
LocalAI
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No…
★ 48.8k
ChatTTS
A generative speech model for daily dialogue.
★ 39.8k
OpenVoice
Instant voice cloning by MIT and MyShell. Audio foundation model.
★ 37k
mastra
Mastra is the modern TypeScript framework for AI-powered applications and agents.
★ 27.6k
Pixelle-Video
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
★ 25.7k
Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models,…
★ 18.4k
edge-tts
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or…
★ 11.5k
KrillinAI
AI video translation & dubbing tool for humans and AI Agents, powered by LLMs. Full pipeline: download,…
★ 11.3k
voltagent
AI Agent Engineering Platform built on an Open Source TypeScript AI Agent Framework
★ 10.5k
Bert-VITS2
vits2 backbone with multilingual-bert
★ 8.8k
Vision-Agents
Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses…
★ 8.1k
StyleTTS2
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large…
★ 6.3k
Orpheus-TTS
Towards Human-Sounding Speech
★ 6.3k
abogen
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
★ 5.8k
dramaclaw
A general-purpose AIGC video engine: script to finished film in one pipeline — dramas, ads, product videos,…
★ 4.8k
Streamer-Sales
★ 3.8k
TTS-WebUI
A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS,…
★ 3.2k
awesome-speech-recognition-speech-synthesis-papers
Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language…
★ 3.1k
vall-e
An unofficial PyTorch implementation of the audio LM VALL-E
★ 3k
Stable-Diffusion
FLUX, Stable Diffusion, SDXL, SD3, LoRA, Fine Tuning, DreamBooth, Training, Automatic1111, Forge WebUI,…
★ 2.8k
marytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
★ 2.6k
ReadAny
AI-powered cross-platform e-book reader with semantic search, RAG chat, local vector store, notes, TTS, and…
★ 2.3k
vall-e
PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo…
★ 2.2k
openai-edge-tts
Free, high-quality text-to-speech API endpoint to replace OpenAI, Azure, or ElevenLabs
★ 2.1k
ElatoAI
Realtime Voice AI with 100+ Models on Arduino ESP32 with Secure Websockets and Edge Functions for AI…
★ 1.9k
Dot
Text-To-Speech, RAG, and LLMs. All local!
★ 1.9k
bailing
百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek…
★ 1.8k
uzu
A high-performance inference engine for AI models
★ 1.7k
vox-director
Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas…
★ 1.7k
ava-whatsapp-agent-course
Meet Ava, the WhatsApp Agent
★ 1.7k
Genie-TTS
GPT-SoVITS ONNX Inference Engine & Model Converter
★ 1.7k
amica
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and…
★ 1.6k
BrowserAI
Run local LLMs like llama, deepseek-distill, kokoro and more inside your browser
★ 1.4k
Speech-AI-Forge
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a…
★ 1.4k
soprano
Soprano: Instant, Ultra-Realistic Text-to-Speech
★ 1.2k
XZVoice
Free and open source text-to-speech software
★ 1.2k
Irene-Voice-Assistant
Ирина - русский голосовой ассистент для работы оффлайн.…
★ 1.1k
epub2tts
Turn an epub or text file into an audiobook
★ 956
MimikaStudio
MimikaStudio - A local-first application for macOS (Apple Silicon) + Agentic MCP Support
★ 730
UnrealGenAISupport
Unreal Engine plugin for LLM/GenAI models & MCP UE5 server. OpenAI GPT-5, Deepseek R1, Claude Opus/Sonnet,…
★ 638
tiktok-voice
Simple Python script to interact with the TikTok TTS API
★ 607
vits2_pytorch
unofficial vits2-TTS implementation in pytorch
★ 548
orpheus-tts-local
Run Orpheus 3B Locally With LM Studio
★ 544
Orkas-VideoStudio
Turn your coding agent into a video studio: describe a video in plain language, and your agent writes the…
★ 525
ComfyUI-OmniVoice-TTS
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and…
★ 509
dectalk
Modern builds for the 90s/00s DECtalk text-to-speech application.
★ 451
skills
Collections of skills for building with ElevenLabs
★ 430
floe-guard
The spend meter, cost/margin, & budget gate for AI voice agents. Meters STT + TTS + LLM + telephony per call,…
★ 410
Stream-Omni
Stream-Omni is a GPT-4o-like language-vision-speech chatbot that simultaneously supports interaction across…
★ 392
VoxNovel
VoxNovel: generate audiobooks giving each character a different voice actor.
★ 374
easevoice-trainer
EaseVoice Trainer is a simple and user-friendly voice cloning and speech model trainer.
★ 352
gemini-youtube-automation
A fully autonomous AI Agent/Python pipeline that utilizes Large Language Models (LLMs) like Gemini to…
★ 334
voiceai
Set of 📝 with 🔗 to help those building Voice AI agents 🎙️🤖
★ 306
Qwen3-TTS-EasyFinetuning
Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.
★ 121
🔗 Related families

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.