Owner · kyutai-labs
kyutai-labs
kyutai-labs's tracked open-source repos, sorted by stars.
5 repos
- #1
Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
★ 11,002+0Star change over the last 7 days - #2
A TTS that fits in your CPU (and pocket)
★ 9,335+0Star change over the last 7 days - #3
Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.
★ 3,021+0Star change over the last 7 days - #4
Hibiki is a model for streaming speech translation (also known as simultaneous translation). Unlike offline translation—where one waits for the end of the source utterance to start translating--- Hibiki adapts its flow to accumulate just enough context to produce a correct translation in real-time, chunk by chunk.
★ 1,512+0Star change over the last 7 days - #5★ 1,508+0Star change over the last 7 days