Skip to main content
buildradar
Sign in
Owner · kyutai-labs

kyutai-labs

kyutai-labs's tracked open-source repos, sorted by stars.

5 repos
  • moshi@kyutai-labs

    Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.

    11,002+0Star change over the last 7 days
  • pocket-tts@kyutai-labs

    A TTS that fits in your CPU (and pocket)

    9,335+0Star change over the last 7 days
  • delayed-streams-modeling@kyutai-labs

    Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.

    3,021+0Star change over the last 7 days
  • hibiki@kyutai-labs

    Hibiki is a model for streaming speech translation (also known as simultaneous translation). Unlike offline translation—where one waits for the end of the source utterance to start translating--- Hibiki adapts its flow to accumulate just enough context to produce a correct translation in real-time, chunk by chunk.

    1,512+0Star change over the last 7 days
  • unmute@kyutai-labs

    Make text LLMs listen and speak

    1,508+0Star change over the last 7 days
← Back to owner ranking