Post-Cutoff.com
  1. Home
  2. Models
  3. Kyutai Moshi / Hibiki-Zero (full-duplex speech models)

Kyutai Moshi / Hibiki-Zero (full-duplex speech models)

Kyutaicurrentaudio/speechMoshiopen weights

Moshi (announced July 2024, weights + paper Sept 2024) is widely cited as the first real-time full-duplex open spoken dialogue model; NVIDIA PersonaPlex-7B (Jan 2026) is fine-tuned from Moshiko weights. Variants: moshiko (male)/moshika (female) in PyTorch bf16/int8, MLX int4/int8/bf16, Rust/Candle. Code MIT/Apache, weights CC-BY-4.0.

Input
audio
Output
audio, text
License
cc-by-4.0
Verified
2026-09-29

How to call it

ProviderModel idEndpoint / URLDocs
Hugging Facekyutai/moshiko-pytorch-bf16github.com/kyutai-labs/moshi—
Hugging Facekyutai/hibiki-zero-3b-pytorch-bf16huggingface.co/kyutai/hibiki-zero-3b-pytorch-bf16—
Web demo—moshi.chat—

Notable capabilities (3)

Other Kyutai models

Kyutai Pocket TTS · Kyutai TTS 1.6B / Kyutai STT + Unmute