text-to-speech
Tracked open-source repos tagged text-to-speech, sorted by stars.
- #121
Flutter Text to Speech package
★ 753+0Star change over the last 7 days - #122
🔊 Kokoro Web: Free AI text-to-speech, online or self-hosted, OpenAI compatible!
★ 733+2Star change over the last 7 days - #123
Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channel integration, agent state management, and observability.
★ 726+4Star change over the last 7 days - #124★ 718-1Star change over the last 7 days
- #125
Fully automated video maker using motion graphics and text-to-speech synthesis to turn newsletters into daily YouTube videos.
★ 703+0Star change over the last 7 days - #126
Implementation of Voicebox, new SOTA Text-to-speech network from MetaAI, in Pytorch
★ 703+0Star change over the last 7 days - #127
Offline Speech Recognition with OpenAI Whisper and TensorFlow Lite for Android
★ 689+1Star change over the last 7 days - #128
Local, OpenAI-compatible text-to-speech (TTS) API using Chatterbox, enabling users to generate voice cloned speech anywhere the OpenAI API is used (e.g. Open WebUI, AnythingLLM, etc.)
★ 677+1Star change over the last 7 days - #129
📚 A customizable dictionary extension that supports double-click lookups in 20+ languages, 1000+ dictionaries, text-to-speech, translation and Anki integration.
★ 671+1Star change over the last 7 days - #130
LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesis
★ 659+1Star change over the last 7 days - #131
Implementation of F5-TTS in MLX
★ 644+0Star change over the last 7 days - #132★ 626+2Star change over the last 7 days
- #133
Turn PDFs and EPUBs into audiobooks; subtitles or videos into dubbed videos (including translation), and more. For free. Pandrator uses local models, including voice-cloning (instant, RVC-enhanced, XTTS fine-tuning) and LLM processing. It aspires to be a user-friendly app with a GUI, an installer and all-in-one packages.
★ 622+3Star change over the last 7 days - #134
Simple Python script to interact with the TikTok TTS API
★ 611+0Star change over the last 7 days - #135
ComfyUI custom node for the VibeVoice TTS. Expressive, long-form, multi-speaker conversational audio
★ 596-1Star change over the last 7 days - #136
Modified version of Chatterbox that accepts text files as input and no character restrictions. I use it to make audiobooks, especially for my kids.
★ 575-1Star change over the last 7 days - #137
The Self-Coding System for Your App — Alan AI SDK for React Native
★ 575+0Star change over the last 7 days - #138
Run Qwen3-TTS text-to-speech locally on Mac (M1/M2/M3/M4). Voice cloning, voice design, custom voices. 100% offline using MLX.
★ 561+2Star change over the last 7 days - #139
🔥🔥🔥 A curated list of papers on LLMs-based multimodal generation (image, video, 3D and audio).
★ 552+0Star change over the last 7 days - #140
Run Orpheus 3B Locally With LM Studio
★ 549+0Star change over the last 7 days - #141
unofficial vits2-TTS implementation in pytorch
★ 549+1Star change over the last 7 days - #142
AI Plugin is a powerful extension for the Payload CMS, integrating advanced AI capabilities to enhance content creation and management.
★ 548-1Star change over the last 7 days - #143
Lightning-fast, free, local first voice dictation for macOS with on-device transcription
★ 547+13Star change over the last 7 days - #144
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
★ 547+6Star change over the last 7 days - #145★ 525+2Star change over the last 7 days
- #146
Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in Pytorch
★ 517+0Star change over the last 7 days - #147
A Vietnamese Voice Cloning Text-to-Speech Model ✨
★ 516-1Star change over the last 7 days - #148
Open source, local, and self-hosted highly optimized language inference server supporting ASR/STT, TTS, and LLM across WebRTC, REST, and WS
★ 512+0Star change over the last 7 days - #149
ComfyUI node for highly expressive speech and realistic zero-shot voice cloning
★ 509+0Star change over the last 7 days - #150
An open-source read-along document reader server with high-quality TTS options, synchronized highlighting, and audiobook export for EPUB, PDF, DOCX, TXT, and MD.
★ 505+4Star change over the last 7 days