tts
Tracked open-source repos tagged tts, sorted by stars.
- #61
AI-powered cross-platform e-book reader with semantic search, RAG chat, local vector store, notes, TTS, and WebDAV sync.
★ 2,305+24Star change over the last 7 days - #62★ 2,296+7Star change over the last 7 days
- #63
PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo https://lifeiteng.github.io/valle/index.html
★ 2,215+0Star change over the last 7 days - #64
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
★ 2,214+115Star change over the last 7 days - #65
Controllable and fast Text-to-Speech for over 7000 languages!
★ 2,207-1Star change over the last 7 days - #66
Free, high-quality text-to-speech API endpoint to replace OpenAI, Azure, or ElevenLabs
★ 2,075+5Star change over the last 7 days - #67
EPUB to audiobook converter, optimized for Audiobookshelf, WebUI included
★ 2,064+1Star change over the last 7 days - #68
Realtime Voice AI with 100+ Models on Arduino ESP32 with Secure Websockets and Edge Functions for AI Companions, and Devices
★ 1,938+5Star change over the last 7 days - #69
Управление Яндекс.Станцией и другими устройствами умного дома с Алисой из Home Assistant
★ 1,917-1Star change over the last 7 days - #70★ 1,912+1Star change over the last 7 days
- #71
Workflow-to-APP、ScreenShare&FloatingVideo、GPT & 3D、SpeechRecognition&TTS
★ 1,861-2Star change over the last 7 days - #72★ 1,836+1Star change over the last 7 days
- #73
A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents.
★ 1,835+14Star change over the last 7 days - #74★ 1,760+3Star change over the last 7 days
- #75
百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断
★ 1,759+2Star change over the last 7 days - #76★ 1,749+14Star change over the last 7 days
- #77
Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to create a new dubbed & translated audio track where the speech is synced using the subtitle's timings.
★ 1,746-1Star change over the last 7 days - #78
Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas Cloud + ffmpeg. An agent skill.
★ 1,698+67Star change over the last 7 days - #79★ 1,687+4Star change over the last 7 days
- #80
Meet Ava, the WhatsApp Agent
★ 1,675+2Star change over the last 7 days - #81
Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch
★ 1,646+1Star change over the last 7 days - #82
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
★ 1,622+8Star change over the last 7 days - #83
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
★ 1,591-2Star change over the last 7 days - #84
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows.
★ 1,555+4Star change over the last 7 days - #85★ 1,548+60Star change over the last 7 days
- #86
Dicio assistant app for Android
★ 1,463+3Star change over the last 7 days - #87
Run local LLMs like llama, deepseek-distill, kokoro and more inside your browser
★ 1,449+1Star change over the last 7 days - #88
A Python/Pytorch app for easily synthesising human voices
★ 1,440+0Star change over the last 7 days - #89★ 1,437+0Star change over the last 7 days
- #90
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.
★ 1,427+1Star change over the last 7 days