Popular repositories Loading
-
-
delayed-streams-modeling
delayed-streams-modeling PublicKyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.
-
Repositories
Showing 10 of 29 repositories
- ARC-Encoder Public
- dactory Public
- kairos Public
- moshi Public
Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
- tts_longeval Public
- moshi-rag Public
MoshiRAG is a compact full-duplex speech language model augmented with asynchronous knowledge retrieval to improve factuality without sacrificing real-time interactivity.
Top languages
Loading…
Most used topics
Loading…