Skip to main content
buildradar
Sign in
Topic · tts

tts

Tracked open-source repos tagged tts, sorted by stars.

189 repos
  • ReadAny@codedogQBY

    AI-powered cross-platform e-book reader with semantic search, RAG chat, local vector store, notes, TTS, and WebDAV sync.

    2,305+24Star change over the last 7 days
  • easyVoice@cosin2077

    开源文本转语音工具,支持超长文本,多角色配音

    2,296+7Star change over the last 7 days
  • vall-e@lifeiteng

    PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo https://lifeiteng.github.io/valle/index.html

    2,215+0Star change over the last 7 days
  • audio.cpp@0xShug0

    An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.

    2,214+115Star change over the last 7 days
  • IMS-Toucan@DigitalPhonetics

    Controllable and fast Text-to-Speech for over 7000 languages!

    2,207-1Star change over the last 7 days
  • openai-edge-tts@travisvn

    Free, high-quality text-to-speech API endpoint to replace OpenAI, Azure, or ElevenLabs

    2,075+5Star change over the last 7 days
  • EPUB to audiobook converter, optimized for Audiobookshelf, WebUI included

    2,064+1Star change over the last 7 days
  • ElatoAI@akdeb

    Realtime Voice AI with 100+ Models on Arduino ESP32 with Secure Websockets and Edge Functions for AI Companions, and Devices

    1,938+5Star change over the last 7 days
  • YandexStation@AlexxIT

    Управление Яндекс.Станцией и другими устройствами умного дома с Алисой из Home Assistant

    1,917-1Star change over the last 7 days
  • Dot@alexpinel

    Text-To-Speech, RAG, and LLMs. All local!

    1,912+1Star change over the last 7 days
  • comfyui-mixlab-nodes@MixLabPro

    Workflow-to-APP、ScreenShare&FloatingVideo、GPT & 3D、SpeechRecognition&TTS

    1,861-2Star change over the last 7 days
  • RHVoice@RHVoice

    a free and open source speech synthesizer for Russian and other languages

    1,836+1Star change over the last 7 days
  • kokoro-tts@nazdridoy

    A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents.

    1,835+14Star change over the last 7 days
  • Genie-TTS@High-Logic

    GPT-SoVITS ONNX Inference Engine & Model Converter

    1,760+3Star change over the last 7 days
  • bailing@wwbin2017

    百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R1等优秀大模型,接入openClaw,真正的个人语音助手,时延低至800ms,Mac等低配置也可运行,支持打断

    1,759+2Star change over the last 7 days
  • VideoClaw@HITsz-TMG

    🚀 AI 全自动化视频生成员工 | Your First AIGC Coworker. Chat an Idea. Get a Film. 🦞

    1,749+14Star change over the last 7 days
  • Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to create a new dubbed & translated audio track where the speech is synced using the subtitle's timings.

    1,746-1Star change over the last 7 days
  • vox-director@Alisa0808

    Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas Cloud + ffmpeg. An agent skill.

    1,698+67Star change over the last 7 days
  • uzu@trymirai

    A high-performance inference engine for AI models

    1,687+4Star change over the last 7 days
  • Meet Ava, the WhatsApp Agent

    1,675+2Star change over the last 7 days
  • ParallelWaveGAN@kan-bayashi

    Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch

    1,646+1Star change over the last 7 days
  • dsnote@mkiol

    Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.

    1,622+8Star change over the last 7 days
  • amica@semperai

    Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.

    1,591-2Star change over the last 7 days
  • VibeVoice-ComfyUI@Enemyx-net

    A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows.

    1,555+4Star change over the last 7 days
  • soprano@ekwek1

    Soprano: Instant, Ultra-Realistic Text-to-Speech

    1,548+60Star change over the last 7 days
  • dicio-android@DicioTeam

    Dicio assistant app for Android

    1,463+3Star change over the last 7 days
  • BrowserAI@sauravpanda

    Run local LLMs like llama, deepseek-distill, kokoro and more inside your browser

    1,449+1Star change over the last 7 days
  • Voice-Cloning-App@voice-cloning-app

    A Python/Pytorch app for easily synthesising human voices

    1,440+0Star change over the last 7 days
  • OuteTTS@edwko

    Interface for OuteTTS models.

    1,437+0Star change over the last 7 days
  • Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.

    1,427+1Star change over the last 7 days
← Back to topics