speech-synthesis
Tracked open-source repos tagged speech-synthesis, sorted by stars.
- #31
An open-source ChatGPT app with a voice
★ 2,350+0Star change over the last 7 days - #32
Controllable and fast Text-to-Speech for over 7000 languages!
★ 2,207+0Star change over the last 7 days - #33★ 1,836+0Star change over the last 7 days
- #34
Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch
★ 1,646+0Star change over the last 7 days - #35
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
★ 1,622+0Star change over the last 7 days - #36
Custom nodes that extend the capabilities of Comfyui
★ 1,525+0Star change over the last 7 days - #37★ 1,498+0Star change over the last 7 days
- #38★ 1,449+0Star change over the last 7 days
- #39
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.
★ 1,427+0Star change over the last 7 days - #40
💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
★ 1,399+0Star change over the last 7 days - #41
General Speech Restoration
★ 1,375+0Star change over the last 7 days - #42★ 1,339+0Star change over the last 7 days
- #43
StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.
★ 1,288+0Star change over the last 7 days - #44
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
★ 1,228+0Star change over the last 7 days - #45★ 1,227+0Star change over the last 7 days
- #46
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
★ 1,153+0Star change over the last 7 days - #47
AI Vtuber for Streaming on Youtube/Twitch
★ 1,116+3Star change over the last 7 days - #48★ 1,100+0Star change over the last 7 days
- #49
YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone
★ 1,054+1Star change over the last 7 days - #50
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
★ 1,014+1Star change over the last 7 days - #51
AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.
★ 1,000+10Star change over the last 7 days - #52★ 972+4Star change over the last 7 days
- #53
Make Azure natural TTS voices accessible to any SAPI 5-compatible application.
★ 934+7Star change over the last 7 days - #54
An Open-Sourced LLM-empowered Foundation TTS System
★ 919-1Star change over the last 7 days - #55
A talking LLM that runs on your own computer without needing the internet.
★ 891+1Star change over the last 7 days - #56★ 884-1Star change over the last 7 days
- #57
Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine
★ 794+6Star change over the last 7 days - #58
Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime on a 4090. Apache 2.0.
★ 759+7Star change over the last 7 days - #59
Thorsten-Voice: A free to use, offline working, high quality german TTS voice should be available for every project without any license struggling.
★ 729+0Star change over the last 7 days - #60
INTERSPEECH 2023-2024 Papers: A complete collection of influential and exciting research papers from the INTERSPEECH 2023-24 conference. Explore the latest advances in speech and language processing. Code included. Star the repository to support the advancement of speech technology!
★ 684+0Star change over the last 7 days