speech-synthesis
Tracked open-source repos tagged speech-synthesis, sorted by stars.
Related topics
Topics that frequently appear alongside speech-synthesis on the same repo.
Recent risers
Repos created in the last 90 days, tagged speech-synthesis.
No new repos tagged with this topic in the last 90 days.
- #1
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
★ 45,985+18Star change over the last 7 days - #2
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
★ 36,595+345Star change over the last 7 days - #3
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
★ 18,379+21Star change over the last 7 days - #4★ 17,476+1Star change over the last 7 days
- #5
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
★ 14,846+4Star change over the last 7 days - #6
Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.
★ 13,757+20Star change over the last 7 days - #7
Build voice agents with open-source models
★ 13,015+82Star change over the last 7 days - #8
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
★ 12,730+54Star change over the last 7 days - #9
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.
★ 12,676+5Star change over the last 7 days - #10
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
★ 11,847+29Star change over the last 7 days - #11
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
★ 10,276+1Star change over the last 7 days - #12★ 9,949+3Star change over the last 7 days
- #13
so-vits-svc fork with realtime support, improved interface and more features.
★ 9,329+3Star change over the last 7 days - #14
EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
★ 8,522-1Star change over the last 7 days - #15
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
★ 7,826+23Star change over the last 7 days - #16
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
★ 6,802+13Star change over the last 7 days - #17
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
★ 6,342+3Star change over the last 7 days - #18
Silero Models: pre-trained text-to-speech models made embarrassingly simple
★ 6,084-1Star change over the last 7 days - #19★ 5,831+45Star change over the last 7 days
- #20
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
★ 4,855+3Star change over the last 7 days - #21
An Open Source text-to-speech system built by inverting Whisper.
★ 4,645+3Star change over the last 7 days - #22★ 4,258+18Star change over the last 7 days
- #23
Foundational model for human-like, expressive TTS
★ 4,203-1Star change over the last 7 days - #24
Converts text to speech in realtime
★ 4,021+2Star change over the last 7 days - #25
:stuck_out_tongue_closed_eyes: TensorFlowTTS: Real-Time State-of-the-art Speech Synthesis for Tensorflow 2 (supported including English, French, Korean, Chinese, German and Easy to adapt for other languages)
★ 3,996+0Star change over the last 7 days - #26★ 2,864+0Star change over the last 7 days
- #27★ 2,818+1Star change over the last 7 days
- #28
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
★ 2,583+0Star change over the last 7 days - #29
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng Việt • TTS tiếng Việt
★ 2,468+21Star change over the last 7 days - #30
HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis
★ 2,367+0Star change over the last 7 days