Skip to main content
buildradar
Sign in
Topic · voice-cloning

voice-cloning

Tracked open-source repos tagged voice-cloning, sorted by stars.

Repos
41
Total stars
355,522
Avg. stars
8,671
Share
0.02%

Topics that frequently appear alongside voice-cloning on the same repo.

Recent risers

Repos created in the last 90 days, tagged voice-cloning.

  • audio.cpp@0xShug0

    An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.

    2,214
  • GPT-SoVITS@RVC-Boss

    1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

    61,476+138Star change over the last 7 days
  • Clone a voice in 5 seconds to generate arbitrary speech in real-time

    60,117-4Star change over the last 7 days
  • TTS@coqui-ai

    🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

    45,985+18Star change over the last 7 days
  • VoxCPM@OpenBMB

    VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

    36,595+345Star change over the last 7 days
  • CosyVoice@QwenAudio

    Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

    23,426+359Star change over the last 7 days
  • ebook2audiobook@DrewThomasson

    Generate audiobooks from e-books, voice cloning & 1158+ languages!

    20,097+34Star change over the last 7 days
  • VideoLingo@Huanshere

    Netflix-level subtitle cutting, translation, alignment, and even dubbing - one-click fully automated AI video subtitle team | Netflix级字幕切割、翻译、对齐、甚至加上配音,一键全自动视频搬运AI字幕组

    18,339+49Star change over the last 7 days
  • VoiceStudio@debpalash

    VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

    14,835+2,880Star change over the last 7 days
  • voice-pro@abus-aikorea

    Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.

    12,730+54Star change over the last 7 days
  • PaddleSpeech@PaddlePaddle

    Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.

    12,676+5Star change over the last 7 days
  • YuE@multimodal-art-projection

    YuE: Open Full-song Music Generation Foundation Model, something similar to Suno.ai but open

    6,416+8Star change over the last 7 days
  • YouDub-webui@liuzhao1225

    Open-source AI video localization and dubbing for YouTube/Bilibili: speech recognition, subtitle translation, voice cloning, audio mixing and rendering. 开源 AI 视频翻译配音工具。

    5,404+25Star change over the last 7 days
  • MOSS-TTS@OpenMOSS

    An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS

    4,058+18Star change over the last 7 days
  • Applio@IAHispano

    A simple, high-quality voice conversion tool focused on ease of use and performance.

    3,674+14Star change over the last 7 days
  • MARS5-TTS@Camb-ai

    MARS5 speech model (TTS) from CAMB.AI

    2,818+1Star change over the last 7 days
  • audio.cpp@0xShug0

    An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.

    2,214+115Star change over the last 7 days
  • Genie-TTS@High-Logic

    GPT-SoVITS ONNX Inference Engine & Model Converter

    1,760+3Star change over the last 7 days
  • MiniMax-MCP@MiniMax-AI

    Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.

    1,573+2Star change over the last 7 days
  • VibeVoice-ComfyUI@Enemyx-net

    A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows.

    1,555+4Star change over the last 7 days
  • Voice-Cloning-App@voice-cloning-app

    A Python/Pytorch app for easily synthesising human voices

    1,440+0Star change over the last 7 days
  • Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.

    1,427+1Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • Twocast@panyanyany

    AI Podcast Generator for bilingual episodes, Multi Languages, Alternative to NotebookLLM;真人对话AI播客生成器,多语言,多音色

    1,289-1Star change over the last 7 days
  • audio-webui@gitmylo

    A webui for different audio related Neural Networks

    1,246+0Star change over the last 7 days
  • Irodori-TTS@Aratako

    A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control

    1,228+4Star change over the last 7 days
  • TTS-Audio-Suite@diodiogod

    A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools

    1,187+7Star change over the last 7 days
  • mlx-serve@ddalcu

    Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling.

    1,115+245Star change over the last 7 days
  • alexandria-audiobook@Finrandojin

    AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.

    999+17Star change over the last 7 days
  • Step-Audio-EditX@stepfun-ai

    A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech

    972+1Star change over the last 7 days
  • CloneTTS@sipeter

    A lightweight, offline Android Text-to-Speech (TTS) engine enabling seamless system-wide voice cloning and high-fidelity text reading. / 运行在安卓本地的轻量级文字转语音 (TTS) 引擎,支持离线发音人提取、零门槛音色克隆与双擎系统级全局听书。

    803+2Star change over the last 7 days
← Back to topics