Skip to main content
buildradar
Sign in
Topic · speech

speech

Tracked open-source repos tagged speech, sorted by stars.

74 repos
  • TTS-papers@coqui-ai

    🐸 collection of TTS papers

    734+1Star change over the last 7 days
  • MASR@yeyupiaoling

    Pytorch实现的流式与非流式的自动语音识别框架,同时兼容在线和离线识别,目前支持Conformer、Squeezeformer、DeepSpeech2模型,支持多种数据增强方法。

    727+0Star change over the last 7 days
  • FreeVC@OlaWod

    FreeVC: Towards High-Quality Text-Free One-Shot Voice Conversion

    716+0Star change over the last 7 days
  • openspeech@openspeech-team

    Open-Source Toolkit for End-to-End Speech Recognition leveraging PyTorch-Lightning and Hydra.

    714+0Star change over the last 7 days
  • BabelDuck@Orenoid

    Beginner-friendly AI conversation practice application

    693+1Star change over the last 7 days
  • chatterbox-tts-api@travisvn

    Local, OpenAI-compatible text-to-speech (TTS) API using Chatterbox, enabling users to generate voice cloned speech anywhere the OpenAI API is used (e.g. Open WebUI, AnythingLLM, etc.)

    677+5Star change over the last 7 days
  • MOSS-Audio@OpenMOSS

    An open-source model for understanding speech, environmental sounds, and music through captioning, question answering, and reasoning

    656+5Star change over the last 7 days
  • sonus@evancohen

    :speech_balloon: /so.nus/ STT (speech to text) for Node with offline hotword detection

    638+0Star change over the last 7 days
  • 语音方向实验室/公司/资源/实习等,欢迎推荐或自荐

    610+1Star change over the last 7 days
  • A list of publicly available room impulse response datasets and scripts to download them.

    609+2Star change over the last 7 days
  • Leaderboard@SpeechColab

    SpeechIO Leaderboard: a large, robust, comprehensive, benchmarking platform for Automatic Speech Recognition.

    553+3Star change over the last 7 days
  • tacotron@google

    Audio samples accompanying publications related to Tacotron, an end-to-end speech synthesis model.

    539+0Star change over the last 7 days
  • knn-vc@bshall

    Voice Conversion With Just Nearest Neighbors

    524+1Star change over the last 7 days
  • StarGANv2-VC@yl4579

    StarGANv2-VC: A Diverse, Unsupervised, Non-parallel Framework for Natural-Sounding Voice Conversion

    522+0Star change over the last 7 days
← Back to topics