voice-conversion
Tracked open-source repos tagged voice-conversion, sorted by stars.
Related topics
Topics that frequently appear alongside voice-conversion on the same repo.
Recent risers
Repos created in the last 90 days, tagged voice-conversion.
- #1
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
★ 2,214 - #2★ 220
- #1
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
★ 45,985+18Star change over the last 7 days - #2
Easily train a good VC model with voice data <= 10 mins!
★ 38,020+114Star change over the last 7 days - #3
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
★ 12,730+54Star change over the last 7 days - #4
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
★ 10,276+1Star change over the last 7 days - #5★ 9,949+3Star change over the last 7 days
- #6
so-vits-svc fork with realtime support, improved interface and more features.
★ 9,329+3Star change over the last 7 days - #7★ 5,831+45Star change over the last 7 days
- #8
A simple, high-quality voice conversion tool focused on ease of use and performance.
★ 3,674+14Star change over the last 7 days - #9
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
★ 3,061+164Star change over the last 7 days - #10
🔊 A comprehensive list of open-source datasets for voice and sound computing (95+ datasets).
★ 2,223+1Star change over the last 7 days - #11
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
★ 2,214+115Star change over the last 7 days - #12★ 1,425+3Star change over the last 7 days
- #13
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools
★ 1,187+7Star change over the last 7 days - #14★ 1,100+0Star change over the last 7 days
- #15
YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone
★ 1,053+0Star change over the last 7 days - #16★ 970+2Star change over the last 7 days
- #17★ 763+1Star change over the last 7 days
- #18★ 716+0Star change over the last 7 days
- #19
Unsupervised Speech Decomposition Via Triple Information Bottleneck
★ 698+0Star change over the last 7 days - #20★ 524+1Star change over the last 7 days
- #21
StarGANv2-VC: A Diverse, Unsupervised, Non-parallel Framework for Natural-Sounding Voice Conversion
★ 522+0Star change over the last 7 days - #22
An Open-Source Project to Unify Audio Processing and Generation
★ 512+1Star change over the last 7 days