kyutai-labs/moshi
steadyMoshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
Python
View on GitHub
Stars
11,138
Forks
1,034
Open issues
73
24h
+14
+0.1%
7d
+44
+0.4%
Refresh
2h
Star history (7 days)
Last checked
42m ago
Last pushed
09 Sep 2026
Next check
just now