Moshi
Extracting target signal
Moshi
Extracting target signal
Moshi
Arcmira media summary
Browse Moshi reviews, demos & launch coverage — 8 indexed from Alex Volkov from ThursdAI & AI Engineer, updated May 2026.
The first full-duplex speech-to-speech model for conversation developed by the speaker's team.
Mobile coding application mentioned by Max.
if you guys remember Moshi so they just seem to have released it seems like a 750 million parameter audio model.
Released the QT TTS pipeline for text-to-speech generation.
A 7 billion parameter dual-channel audio model.
Arcmira tracks 8 indexed media appearances or mentions for Moshi, tied to source videos, channels, and transcript-derived context.
Arcmira uses indexed YouTube videos and transcripts. Representative source evidence on this page includes "Voice AI: when is the "Her" moment? — Neil Zeghidour, Gradium AI" with transcript-derived context and links when available.
Moshi is connected to turn detection, malleable software, code in Arcmira's media graph.
8
Mentions
89.4K
Views
The trendline is visible, but the dated evidence behind Moshi is in the premium layer.

“The first full-duplex speech-to-speech model for conversation developed by the speaker's team.”

“Mobile coding application mentioned by Max.”

“if you guys remember Moshi so they just seem to have released it seems like a 750 million parameter audio model.”

“Released the QT TTS pipeline for text-to-speech generation.”

“A 7 billion parameter dual-channel audio model.”