Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. Meta launches Muse Voice Transcribe, its first real-time…

Meta launches Muse Voice Transcribe, its first real-time speech model on the Meta Model API

★★★after cutoffmodel-releaseMetaconfidence: high

On 2026-09-03 Meta Superintelligence Labs released Muse Voice Transcribe (muse-voice-transcribe-1.0), a streaming and file speech-to-text model on the Meta Model API at $0.18/hour that Meta says ranks #1 on the Artificial Analysis streaming STT leaderboard, with built-in diarization for 20+ speakers.

Key facts

What happened

Meta added its first audio model to the Meta Model API alongside Muse Spark, Muse Image and Muse Glimmer: a streaming ASR model aimed at developers building voice agents (typically chained STT -> Muse Spark -> third-party TTS).

Why it matters

It extends Meta's paid-API push beyond text and images into speech, landing the same day as Microsoft's MAI-Transcribe-2 amid a September 2026 price war in speech-to-text. Leaderboard claims are Meta's.

Changelog

  • 2026-09-29: created

Models

Related events

  1. Meta releases Muse Spark 1.1 and opens the Meta Model API public preview ★★★
  2. Meta Connect 2026: VR Glasses, Ray-Ban Meta Gen 3, camera-free audio glasses and Muse everywhere ★★★★
  3. Microsoft launches MAI-Transcribe-2, claiming the most accurate and cheapest speech recognition at $0.10/hour ★★★

Sources (3)

id: 2026-09-03-meta-muse-voice-transcribe · updated 2026-09-29 · open in the interactive timeline