Speechify Simba 3.2
Exact API model id string not verified (docs page 'SpeechifyAI Build TTS Models: Simba 3.2, 3.0, Multilingual, and English'). AA measured ~30.2 chars/s generation speed (the-decoder, Jul 2026). Quotes: Luke Oliff, Tyler Weitzman in the press release.
- Input
- text, audio
- Output
- audio
- License
- proprietary
- Pricing
- per 1m characters: $10 (USD per 1M characters (entry tier); $6 at scale tier) source
- Verified
- 2026-09-29
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| SpeechifyAI API | — | https://api.speechify.ai/v1/audio/speech (also /v1/audio/stream) | docs |
| Web | — | speechify.ai/models | — |
Notable capabilities (2)
- Briefly #1 on Artificial Analysis Speech Arena at a low price: Press release 2026-07-07 claimed #1 on the AA TTS leaderboard; a week later Qwen-Audio-3.0-TTS-Plus overtook it (1,236 vs 1,234 Elo). On 2026-09-29 it was #7 (Elo 1239). Speechify called it the cheapest model in the top ten ($10/$6 per 1M chars). source
- Streaming-native, low TTFB: Streaming-native Simba 3 model; <100 ms first byte claimed; emotional control, SSML prosody, instant voice cloning; 30+ locales with mixed-language input. Recommended model for English integrations. source
Sources: https://speechify.ai/blog/simba-3-2-and-the-models-endpoint , https://docs.speechify.ai/build/guides/concepts/models , PRWeb release (2026-07-07), https://artificialanalysis.ai/text-to-speech/leaderboard