BreezeBlue releases Breeze TTS 2, the new top open-weights text-to-speech model
On 2026-08-25 BreezeBlue published weights and inference code for Breeze TTS 2, a 3B text-to-speech model with voice cloning, voice design and voice direction and under-40 ms time-to-first-audio on an H100. It became the highest-rated open-weights model on the Artificial Analysis Speech Arena (~1,206-1,215 Elo, about 90 points above Fish Audio S2 Pro), though its weights are licensed for research/non-commercial use only.
Key facts
- 3B params; cloning, text-described voice design, voice direction, vocal events in one checkpoint
- TTFA <40 ms on H100 (fast path), streaming RTF 0.32; needs 12-24 GB VRAM
- Artificial Analysis: #1 open weights, ~#6 overall at launch; open-weights top 5 in late Sept 2026: Breeze TTS 2, Fish Audio S2 Pro, Step Audio EditX, Voxtral TTS, Kokoro 82M
- Weights: BreezeBlue Research and Non-Commercial License; code Apache-2.0; commercial use via breezeblue.ai subscription
- Model card lists English + Chinese; AA post cites 50 languages (unresolved)
What happened
BreezeBlue, a lab little known before this release, opened the weights of Breeze TTS 2 on Hugging Face and GitHub. A single 3B checkpoint does zero-shot cloning, voice design from a prompt, emotional/tonal direction, and real-time bilingual streaming.
Why it matters
It pushed the open-weights ceiling in TTS about 90 Elo higher, narrowing the gap to closed leaders (Eleven v4, Cartesia Sonic-3.6). "Open" here is weights-available but non-commercial, like Fish Audio S2 Pro and Higgs TTS 3. For commercially free options, MIT/Apache models such as Chatterbox and Kokoro remain the choice. Confidence is medium: the organisation is new, and its language coverage is reported inconsistently.
Changelog
- 2026-09-29: created
Models
- Breeze TTS 2 BreezeBlue · current
Related posts (1)
- Artificial Analysis Artificial Analysis @ArtificialAnlys · x · 2026-08-25
Cited as a source by: 2026-08-25-breeze-tts-2, breeze-tts-2
Related events
- Fish Audio open-sources S2: expressive 80+ language TTS with inline emotion tags ★★★
- Mistral releases Voxtral TTS, an open-weight 4B text-to-speech model with 3-second voice cloning ★★★
- ElevenLabs launches Eleven v4 and Eleven v4 Turbo, #1 on Artificial Analysis TTS arena ★★★★
Sources (4)
- codeHugging Face: BreezeBlue/Breeze-TTS-2
- codeGitHub: breezeblue-ai/breeze-tts
- discussionArtificial Analysis on X: Breeze TTS 2 leads open-weights TTS
- officialArtificial Analysis open-weights TTS leaderboard
id: 2026-08-25-breeze-tts-2 · updated 2026-09-29 · open in the interactive timeline