voice-cloning
Tracked open-source repos tagged voice-cloning, sorted by stars.
- #31
Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime on a 4090. Apache 2.0.
★ 759+7Star change over the last 7 days - #32
MimikaStudio - A local-first application for macOS (Apple Silicon) + Agentic MCP Support
★ 735+4Star change over the last 7 days - #33
Local, OpenAI-compatible text-to-speech (TTS) API using Chatterbox, enabling users to generate voice cloned speech anywhere the OpenAI API is used (e.g. Open WebUI, AnythingLLM, etc.)
★ 677+1Star change over the last 7 days - #34
Turn PDFs and EPUBs into audiobooks; subtitles or videos into dubbed videos (including translation), and more. For free. Pandrator uses local models, including voice-cloning (instant, RVC-enhanced, XTTS fine-tuning) and LLM processing. It aspires to be a user-friendly app with a GUI, an installer and all-in-one packages.
★ 620+3Star change over the last 7 days - #35
Legacy YouDub AI video translation and voice-cloning pipeline. Active development continues in YouDub WebUI.
★ 615-1Star change over the last 7 days - #36
ComfyUI custom node for the VibeVoice TTS. Expressive, long-form, multi-speaker conversational audio
★ 596-1Star change over the last 7 days - #37
Modified version of Chatterbox that accepts text files as input and no character restrictions. I use it to make audiobooks, especially for my kids.
★ 575-1Star change over the last 7 days - #38
Run Qwen3-TTS text-to-speech locally on Mac (M1/M2/M3/M4). Voice cloning, voice design, custom voices. 100% offline using MLX.
★ 561+2Star change over the last 7 days - #39
Worlds first open-source real-time end-to-end spoken dialogue model with personalized voice cloning.
★ 550+0Star change over the last 7 days - #40
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
★ 545+5Star change over the last 7 days - #41
ComfyUI node for highly expressive speech and realistic zero-shot voice cloning
★ 509-1Star change over the last 7 days