Skip to main content
buildradar
Sign in
Topic · tts

tts

Tracked open-source repos tagged tts, sorted by stars.

189 repos
  • Speech-AI-Forge@lenML

    🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.

    1,418+0Star change over the last 7 days
  • SoniTranslate@R3gm

    Synchronized Translation for Videos. Video dubbing

    1,413+2Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • voicefixer@haoheliu

    General Speech Restoration

    1,375+0Star change over the last 7 days
  • edge-TTS-record@LuckyHookin

    一个可以录制 Microsoft Edge 浏览器的语音合成(TTS)语音并输出为 .wav 音频的(windows平台)工具。

    1,370+1Star change over the last 7 days
  • Matcha-TTS@shivammehta25

    [ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching

    1,353+3Star change over the last 7 days
  • voicemode@mbailey

    Natural voice conversations with Claude Code

    1,347+6Star change over the last 7 days
  • 小智ESP32的Java企业级管理平台,提供设备监控、音色定制、角色切换和对话记录管理的前后端及服务端一体化解决方案

    1,337+1Star change over the last 7 days
  • MouseTooltipTranslator@ttop32

    Mouseover Translate Any Language At Once - Chrome Extension: PDF Translator, EBOOK, EPUB, OCR, TTS, NETFLIX, YOUTUBE DUAL SUBTITLES, GOOGLE DOCS, AI, VIEWER, GMAIL, WRITING, IMAGE, DUAL SUBS, MANGA, HOVER, DICTIONARY, WEBTOON, EDGE, JAPANESE, ENGLISH

    1,305+4Star change over the last 7 days
  • VideoChat@Henry-23

    实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning, with initial package delay as low as 3s.

    1,303-1Star change over the last 7 days
  • StreamSpeech@ictnlp

    StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.

    1,288+0Star change over the last 7 days
  • my-translator@phuc-nt

    Real-time speech translation — macOS & Windows, free TTS, no server, your API keys only

    1,278+2Star change over the last 7 days
  • audio-webui@gitmylo

    A webui for different audio related Neural Networks

    1,246+0Star change over the last 7 days
  • Irodori-TTS@Aratako

    A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control

    1,228+4Star change over the last 7 days
  • vits_chinese@PlayVoice

    Best practice TTS based on BERT and VITS with some Natural Speech Features Of Microsoft; Support ONNX streaming out!

    1,226+0Star change over the last 7 days
  • ekho@hgneng

    Chinese text-to-speech engine

    1,211+0Star change over the last 7 days
  • TTS-Audio-Suite@diodiogod

    A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools

    1,187+7Star change over the last 7 days
  • sokuji@kizuna-ai-lab

    Real-time two-way speech translation for bilingual meetings — auto-detects the spoken language and translates both directions, cloud or fully offline on-device. Desktop (Windows · macOS · Linux) + browser extension (Chrome · Edge) for Zoom, Meet, Teams & any app.

    1,186+40Star change over the last 7 days
  • speech-swift@soniqo

    AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML

    1,161+4Star change over the last 7 days
  • Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.

    1,153+0Star change over the last 7 days
  • DragonianVoice@PriesiaMioShirakana

    多个SVC/TTS的C++推理库

    1,128-1Star change over the last 7 days
  • AI-Waifu-Vtuber@ardha27

    AI Vtuber for Streaming on Youtube/Twitch

    1,115+1Star change over the last 7 days
  • sglang-omni@sgl-project

    SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.

    1,105+130Star change over the last 7 days
  • nobodywho@nobodywho-ooo

    NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.

    1,094+13Star change over the last 7 days
  • GLM-TTS@zai-org

    GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning

    1,063+7Star change over the last 7 days
  • With one command, create a natural-sounding audiobook from a variety of input formats (epub, mobi, txt, PDF, HTML and more!)

    1,058+1Star change over the last 7 days
  • ms-ra-forwarder@wxxxcxx

    免费的在线文本转语音API

    1,057+3Star change over the last 7 days
  • YourTTS@Edresson

    YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone

    1,053+0Star change over the last 7 days
  • vits-simple-api@Artrajz

    A simple VITS HTTP API, developed by extending Moegoe with additional features.

    1,051-1Star change over the last 7 days
  • Cognitive-Speech-TTS@Azure-Samples

    Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.

    1,014+3Star change over the last 7 days
← Back to topics