🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
Training

Moshi

Free while signed in. Answers cite the passages they came from.

First page
Moshi
The curator’s take

introduces a speech-text foundation model and full-duplex spoken dialogue framework; they present several components of the systems; Helium is a 7B parameter text LLM; Mimi is a semantic-acoustic neural audio code with state-of-the-art performance on audio quality; a hierarchical multi-stream architecture that can generate arbitrary conversation in a speech-to-speech manner.

Every Monday
Get next week’s papers.

The same picks and the same summaries, in your inbox. Free, and 176 issues deep.

Subscribe on Substack