StepFun
StepFun @StepFun_ai · x · 2026-09-15 · ★★★ · archived
Cited as a source by: stepaudio-3-realtime
Summary
Archived text
Introducing StepAudio 3, our new family of 5 audio models for real-time voice, speech recognition, speech generation, audio generation and music.
Realtime ranks #1 on Artificial Analysis for both Conversational Dynamics (98.9%) and Speech Reasoning (99.7%). ASR reaches 1.7% WER, matching the best result on the leaderboard.
Build voice agents that handle interruptions, reason while speaking, and call tools. Transcribe speech, generate expressive voices, and create full audio scenes and music.
Available now: Voice AI Lab: https://audio.stepfun.ai/ Blog: https://static.stepfun.com/blog/stepaudio3/
Media: https://pbs.twimg.com/media/HSRmzufaEAABOzS.jpg?name=orig
views 35383 · likes 409 · reposts 34 · replies 28 (at fetch time)
Archived 2026-09-29 via fxtwitter (unofficial).
Archived text
Introducing StepAudio 3, our new family of 5 audio models for real-time voice, speech recognition, speech generation, audio generation and music.
Realtime ranks #1 on Artificial Analysis for both Conversational Dynamics (98.9%) and Speech Reasoning (99.7%). ASR reaches 1.7% WER, matching the best result on the leaderboard.
Build voice agents that handle interruptions, reason while speaking, and call tools. Transcribe speech, generate expressive voices, and create full audio scenes and music.
Available now: Voice AI Lab: https://audio.stepfun.ai/ Blog: https://static.stepfun.com/blog/stepaudio3/
Media: https://pbs.twimg.com/media/HSRmzufaEAABOzS.jpg?name=orig
views 35383 · likes 409 · reposts 34 · replies 28 (at fetch time)
Archived 2026-09-29 via fxtwitter (unofficial).
All posts · id: x-stepfun_ai-2099916376274313630