Skip to main content
buildradar
Sign in
Topic · text-to-speech

text-to-speech

Tracked open-source repos tagged text-to-speech, sorted by stars.

150 repos
  • TTS-Audio-Suite@diodiogod

    A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio EditX, IndexTTS-2, Chatterbox (classic and multilingual), F5-TTS, Higgs Audio 2, 3, and VibeVoice with unlimited text length, SRT timing, Character support, and many audio tools

    1,187+0Star change over the last 7 days
  • A practical lab for building, testing, and evaluating apps with Apple's Foundation Models framework.

    1,178+0Star change over the last 7 days
  • dia2@nari-labs

    TTS model capable of streaming conversational audio in realtime.

    1,173+0Star change over the last 7 days
  • speech-swift@soniqo

    AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML

    1,161+0Star change over the last 7 days
  • The Self-Coding System for Your App — Alan AI SDK for Cordova

    1,132+0Star change over the last 7 days
  • sglang-omni@sgl-project

    SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.

    1,118+56Star change over the last 7 days
  • Uncensored-Local-Studio@techjarves

    Uncensored local AI studio for Windows, Linux, and macOS. Zero-setup GUI for Image Generation, GGUF LLMs, Text to Speech & Speech to Text

    1,118+13Star change over the last 7 days
  • Patter@PatterAI

    Open-source voice-AI SDK. The Vapi/Retell alternative for builders who want to own the stack. Give your AI agent a phone number in 4 lines — Python and TypeScript, MIT licensed, Twilio, Telnyx, and Plivo.

    1,051+7Star change over the last 7 days
  • Cognitive-Speech-TTS@Azure-Samples

    Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.

    1,014+1Star change over the last 7 days
  • alexandria-audiobook@Finrandojin

    AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.

    1,000+10Star change over the last 7 days
  • Step-Audio-EditX@stepfun-ai

    A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech

    973+1Star change over the last 7 days
  • NISQA@gabrielmittag

    NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment

    972+4Star change over the last 7 days
  • botium-speech-processing@codeforequity-at

    Botium Speech Processing

    943+0Star change over the last 7 days
  • Make Azure natural TTS voices accessible to any SAPI 5-compatible application.

    934+7Star change over the last 7 days
  • Vonage REST API client for PHP. API support for SMS, Voice, Text-to-Speech, Numbers, Verify (2FA) and more.

    931-1Star change over the last 7 days
  • FireRedTTS@FireRedTeam

    An Open-Sourced LLM-empowered Foundation TTS System

    919-1Star change over the last 7 days
  • EDDiscovery@EDDiscovery

    Captains log and 3d star map for Elite Dangerous

    900+4Star change over the last 7 days
  • diffwave@lmnt-com

    DiffWave is a fast, high-quality neural vocoder and waveform synthesizer.

    884-1Star change over the last 7 days
  • bark.cpp@PABannier

    Suno AI's Bark model in C/C++ for fast text-to-speech generation

    868+2Star change over the last 7 days
  • lue@paulilaaso

    Terminal eBook Reader with Audiobook-Quality Text-to-Speech — Supports EPUB, PDF, DOCX, HTML, RTF, TXT, and MD.

    805+1Star change over the last 7 days
  • lobe-tts@lobehub

    🎤 Lobe TTS - A high-quality & reliable TTS/STT library for Server and Browser

    805-1Star change over the last 7 days
  • CloneTTS@sipeter

    A lightweight, offline Android Text-to-Speech (TTS) engine enabling seamless system-wide voice cloning and high-fidelity text reading. / 运行在安卓本地的轻量级文字转语音 (TTS) 引擎,支持离线发音人提取、零门槛音色克隆与双擎系统级全局听书。

    804+5Star change over the last 7 days
  • Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)

    804+5Star change over the last 7 days
  • ai-video-editor@MartinDelophy

    Open-source, local-first video editor where creators and AI agents edit the same real timeline.

    798+82Star change over the last 7 days
  • Confucius4-TTS@netease-youdao

    Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine

    794+6Star change over the last 7 days
  • 无需 ROOT 的开源 Android 屏幕实时翻译工具,适合游戏、视觉小说和漫画。支持端侧与云端 OCR、离线 LLM、多种翻译服务和文字朗读(TTS),译文可直接显示在画面上。Open-source no-root Android real-time screen translator for games, visual novels, and manga. Supports on-device and cloud OCR, offline LLMs, multiple translation services, on-screen translations, and text-to-speech (TTS).

    794+73Star change over the last 7 days
  • june@mezbaul-h

    Local voice chatbot for engaging conversations, powered by Ollama, Hugging Face Transformers, and Coqui TTS Toolkit

    788+1Star change over the last 7 days
  • mlx-audio-swift@Blaizzy

    A modular Swift SDK for audio processing with MLX on Apple Silicon

    774+5Star change over the last 7 days
  • vui@fluxions-ai

    Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime on a 4090. Apache 2.0.

    759+7Star change over the last 7 days
  • cboard@cboard-org

    Augmentative and Alternative Communication (AAC) system with text-to-speech for the browser

    756+9Star change over the last 7 days
← Back to topics