Skip to main content
buildradar
Sign in
Topic · speech-synthesis

speech-synthesis

Tracked open-source repos tagged speech-synthesis, sorted by stars.

66 repos
  • chat-with-gpt@cogentapps

    An open-source ChatGPT app with a voice

    2,350+0Star change over the last 7 days
  • IMS-Toucan@DigitalPhonetics

    Controllable and fast Text-to-Speech for over 7000 languages!

    2,207+0Star change over the last 7 days
  • RHVoice@RHVoice

    a free and open source speech synthesizer for Russian and other languages

    1,836+0Star change over the last 7 days
  • ParallelWaveGAN@kan-bayashi

    Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch

    1,646+0Star change over the last 7 days
  • dsnote@mkiol

    Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.

    1,622+0Star change over the last 7 days
  • Custom nodes that extend the capabilities of Comfyui

    1,525+0Star change over the last 7 days
  • SAM@s-macke

    Software Automatic Mouth - Tiny Speech Synthesizer

    1,498+0Star change over the last 7 days
  • SpeechT5@microsoft

    Unified-Modal Speech-Text Pre-Training for Spoken Language Processing

    1,449+0Star change over the last 7 days
  • Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm), and CPU.

    1,427+0Star change over the last 7 days
  • 💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies

    1,399+0Star change over the last 7 days
  • voicefixer@haoheliu

    General Speech Restoration

    1,375+0Star change over the last 7 days
  • World@mmorise

    A high-quality speech analysis, manipulation and synthesis system

    1,339+0Star change over the last 7 days
  • StreamSpeech@ictnlp

    StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.

    1,288+0Star change over the last 7 days
  • Irodori-TTS@Aratako

    A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control

    1,228+0Star change over the last 7 days
  • BigVGAN@NVIDIA

    Official PyTorch implementation of BigVGAN (ICLR 2023)

    1,227+0Star change over the last 7 days
  • Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.

    1,153+0Star change over the last 7 days
  • AI-Waifu-Vtuber@ardha27

    AI Vtuber for Streaming on Youtube/Twitch

    1,116+3Star change over the last 7 days
  • autovc@auspicious3000

    AutoVC: Zero-Shot Voice Style Transfer with Only Autoencoder Loss

    1,100+0Star change over the last 7 days
  • YourTTS@Edresson

    YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone

    1,054+1Star change over the last 7 days
  • Cognitive-Speech-TTS@Azure-Samples

    Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.

    1,014+1Star change over the last 7 days
  • alexandria-audiobook@Finrandojin

    AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.

    1,000+10Star change over the last 7 days
  • NISQA@gabrielmittag

    NISQA - Non-Intrusive Speech Quality and TTS Naturalness Assessment

    972+4Star change over the last 7 days
  • Make Azure natural TTS voices accessible to any SAPI 5-compatible application.

    934+7Star change over the last 7 days
  • FireRedTTS@FireRedTeam

    An Open-Sourced LLM-empowered Foundation TTS System

    919-1Star change over the last 7 days
  • A talking LLM that runs on your own computer without needing the internet.

    891+1Star change over the last 7 days
  • diffwave@lmnt-com

    DiffWave is a fast, high-quality neural vocoder and waveform synthesizer.

    884-1Star change over the last 7 days
  • Confucius4-TTS@netease-youdao

    Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine

    794+6Star change over the last 7 days
  • vui@fluxions-ai

    Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime on a 4090. Apache 2.0.

    759+7Star change over the last 7 days
  • Thorsten-Voice@thorstenMueller

    Thorsten-Voice: A free to use, offline working, high quality german TTS voice should be available for every project without any license struggling.

    729+0Star change over the last 7 days
  • INTERSPEECH 2023-2024 Papers: A complete collection of influential and exciting research papers from the INTERSPEECH 2023-24 conference. Explore the latest advances in speech and language processing. Code included. Star the repository to support the advancement of speech technology!

    684+0Star change over the last 7 days
← Back to topics