text-to-speech
Tracked open-source repos tagged text-to-speech, sorted by stars.
- #91
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools
★ 1,187+0Star change over the last 7 days - #92
A practical lab for building, testing, and evaluating apps with Apple's Foundation Models framework.
★ 1,178+0Star change over the last 7 days - #93★ 1,173+0Star change over the last 7 days
- #94
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
★ 1,161+0Star change over the last 7 days - #95
The Self-Coding System for Your App — Alan AI SDK for Cordova
★ 1,132+0Star change over the last 7 days - #96
SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.
★ 1,118+56Star change over the last 7 days - #97
Uncensored local AI studio for Windows, Linux, and macOS. Zero-setup GUI for Image Generation, GGUF LLMs, Text to Speech & Speech to Text
★ 1,118+13Star change over the last 7 days - #98
Open-source voice-AI SDK. The Vapi/Retell alternative for builders who want to own the stack. Give your AI agent a phone number in 4 lines — Python and TypeScript, MIT licensed, Twilio, Telnyx, and Plivo.
★ 1,051+7Star change over the last 7 days - #99
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
★ 1,014+1Star change over the last 7 days - #100
AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.
★ 1,000+10Star change over the last 7 days - #101
A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech
★ 973+1Star change over the last 7 days - #102★ 972+4Star change over the last 7 days
- #103
Botium Speech Processing
★ 943+0Star change over the last 7 days - #104
Make Azure natural TTS voices accessible to any SAPI 5-compatible application.
★ 934+7Star change over the last 7 days - #105
Vonage REST API client for PHP. API support for SMS, Voice, Text-to-Speech, Numbers, Verify (2FA) and more.
★ 931-1Star change over the last 7 days - #106
An Open-Sourced LLM-empowered Foundation TTS System
★ 919-1Star change over the last 7 days - #107
Captains log and 3d star map for Elite Dangerous
★ 900+4Star change over the last 7 days - #108★ 884-1Star change over the last 7 days
- #109★ 868+2Star change over the last 7 days
- #110
Terminal eBook Reader with Audiobook-Quality Text-to-Speech — Supports EPUB, PDF, DOCX, HTML, RTF, TXT, and MD.
★ 805+1Star change over the last 7 days - #111★ 805-1Star change over the last 7 days
- #112
A lightweight, offline Android Text-to-Speech (TTS) engine enabling seamless system-wide voice cloning and high-fidelity text reading. / 运行在安卓本地的轻量级文字转语音 (TTS) 引擎,支持离线发音人提取、零门槛音色克隆与双擎系统级全局听书。
★ 804+5Star change over the last 7 days - #113
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)
★ 804+5Star change over the last 7 days - #114
Open-source, local-first video editor where creators and AI agents edit the same real timeline.
★ 798+82Star change over the last 7 days - #115
Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine
★ 794+6Star change over the last 7 days - #116
无需 ROOT 的开源 Android 屏幕实时翻译工具,适合游戏、视觉小说和漫画。支持端侧与云端 OCR、离线 LLM、多种翻译服务和文字朗读(TTS),译文可直接显示在画面上。Open-source no-root Android real-time screen translator for games, visual novels, and manga. Supports on-device and cloud OCR, offline LLMs, multiple translation services, on-screen translations, and text-to-speech (TTS).
★ 794+73Star change over the last 7 days - #117
Local voice chatbot for engaging conversations, powered by Ollama, Hugging Face Transformers, and Coqui TTS Toolkit
★ 788+1Star change over the last 7 days - #118
A modular Swift SDK for audio processing with MLX on Apple Silicon
★ 774+5Star change over the last 7 days - #119
Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime on a 4090. Apache 2.0.
★ 759+7Star change over the last 7 days - #120
Augmentative and Alternative Communication (AAC) system with text-to-speech for the browser
★ 756+9Star change over the last 7 days