Skip to main content
buildradar
Sign in
Topic · text-to-speech

text-to-speech

Tracked open-source repos tagged text-to-speech, sorted by stars.

150 repos
  • RHVoice@RHVoice

    a free and open source speech synthesizer for Russian and other languages

    1,836+0Star change over the last 7 days
  • The Self-Coding System for Your App — Alan AI SDK for Android

    1,807+0Star change over the last 7 days
  • Genie-TTS@High-Logic

    GPT-SoVITS ONNX Inference Engine & Model Converter

    1,760+0Star change over the last 7 days
  • The Self-Coding System for Your App — Alan AI SDK for Flutter

    1,756+0Star change over the last 7 days
  • Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to create a new dubbed & translated audio track where the speech is synced using the subtitle's timings.

    1,746+0Star change over the last 7 days
  • read-aloud@ken107

    An awesome browser extension that reads aloud webpage content with one click

    1,737+0Star change over the last 7 days
  • react-native-executorch@software-mansion

    Declarative way to run AI models in React Native on device, powered by ExecuTorch.

    1,707+0Star change over the last 7 days
  • alan-sdk-ionic@alan-ai

    The Self-Coding System for Your App — Alan AI SDK for Ionic

    1,649+0Star change over the last 7 days
  • ParallelWaveGAN@kan-bayashi

    Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch

    1,646+0Star change over the last 7 days
  • dsnote@mkiol

    Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.

    1,622+0Star change over the last 7 days
  • video-podcast-maker@Agents365-ai

    Topic → 4K narrated video for coding agents. v4.0: all TTS via the ttsCN engine component (11 platforms incl. MiniMax voice clone, native word-level subtitle sync), manifest-based Asset Engine, Remotion composition, cost-gated AI generation

    1,599+0Star change over the last 7 days
  • MiniMax-MCP@MiniMax-AI

    Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.

    1,573+0Star change over the last 7 days
  • vllm-mlx@waybarrios

    High-performance OpenAI and Anthropic compatible LLM inference server for Apple Silicon. Native MLX, continuous batching, multimodal models, MCP tool calling, and Claude Code support.

    1,557+0Star change over the last 7 days
  • VibeVoice-ComfyUI@Enemyx-net

    A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your ComfyUI workflows.

    1,555+0Star change over the last 7 days
  • soprano@ekwek1

    Soprano: Instant, Ultra-Realistic Text-to-Speech

    1,548+0Star change over the last 7 days
  • RCLI@RunanywhereAI

    Talk to your Mac, query your docs, no cloud required. On-device voice AI + RAG

    1,542+0Star change over the last 7 days
  • TalkingHead@met4citizen

    Talking Head (3D): A JavaScript class for real-time lip-sync using full-body 3D avatars.

    1,522+0Star change over the last 7 days
  • Voice-Cloning-App@voice-cloning-app

    A Python/Pytorch app for easily synthesising human voices

    1,440+0Star change over the last 7 days
  • OuteTTS@edwko

    Interface for OuteTTS models.

    1,437+0Star change over the last 7 days
  • Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.

    1,427+0Star change over the last 7 days
  • Speech-AI-Forge@lenML

    🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.

    1,418+0Star change over the last 7 days
  • SoniTranslate@R3gm

    Synchronized Translation for Videos. Video dubbing

    1,413+0Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • mlx-tune@ARahim3

    Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.

    1,398+0Star change over the last 7 days
  • Matcha-TTS@shivammehta25

    [ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching

    1,353+0Star change over the last 7 days
  • WavTokenizer@jishengpeng

    [ICLR 2025] SOTA discrete acoustic codec models with 40/75 tokens per second for audio language modeling

    1,317+0Star change over the last 7 days
  • airunner@Capsize-Games

    Offline inference engine for art, real-time voice conversations, LLM powered chatbots and automated workflows

    1,314+0Star change over the last 7 days
  • StreamSpeech@ictnlp

    StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.

    1,288+0Star change over the last 7 days
  • audio-webui@gitmylo

    A webui for different audio related Neural Networks

    1,246+0Star change over the last 7 days
  • Irodori-TTS@Aratako

    A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control

    1,228+0Star change over the last 7 days
← Back to topics