Breeze TTS 2 brings open-weight voice generation closer to the best proprietary models
BreezeBlue has released Breeze TTS 2, a 3-billion-parameter text-to-speech model designed for real-time voice applications. Available as downloadable weights, it combines voice design, voice cloning, performance direction and streaming generation in a single system.
The model is particularly relevant for voice agents, game characters, interactive storytelling and other applications that require more than conventional text-to-speech. A voice can be created from a natural-language description, then directed through instructions controlling its emotion, tone, pace and delivery. Breeze TTS 2 can also reproduce a voice from a reference recording and its exact transcript, while supporting vocal events such as laughter, sighs and coughs.
In the video below, we present the model, its main capabilities, its position among current TTS systems and the important limitations of its local license.
Breeze TTS 2 ranks sixth in the Artificial Analysis Speech Arena
At the time of writing, Breeze TTS 2 holds sixth place in the Artificial Analysis Provider Voice leaderboard, with an Elo score of approximately 1,215. It is also the highest-ranked open-weight model in the benchmark, ahead of Fish Audio S2 Pro, Step Audio EditX, Voxtral TTS and Kokoro.

The ranking is based on blind listening comparisons rather than benchmarks published solely by the model developer. It does not make Breeze TTS 2 the best option for every project, but it places the model unusually close to leading proprietary services. This makes it an interesting option for developers who want to experiment locally with expressive voices, character creation and low-latency speech generation.
BreezeBlue reports a time to first audio below 40 milliseconds and a real-time factor of 0.32 on an Nvidia H100 using its optimized fast path. These figures describe a specific high-end configuration and should not be treated as expected performance on every local GPU.
The local model is restricted to non-commercial use
Breeze TTS 2 should be described as an open-weight model rather than fully open source. Its PyTorch inference code is published under the Apache 2.0 license, but the downloaded model weights are governed by the BreezeBlue Research and Non-Commercial License.
This distinction also applies to generated content. According to the Breeze TTS 2 model page, the weights, derivative models and outputs produced through a self-hosted instance are restricted to research and non-commercial use. Commercial use requires written authorization from RESONIA, INC. or access through a commercial BreezeBlue service under the applicable terms.
As a result, the local checkpoint should not be used by default for monetized videos, paid client work, commercial games or revenue-generating applications. Downloadable weights do not automatically provide commercial rights.
A strong model, but not yet the fully open TTS alternative
There is currently no clearly permissive open-source TTS model for commercial local use that matches Breeze TTS 2 across this entire combination of voice design, voice cloning, natural-language direction, expressive events, streaming and competitive listening quality.
That gap remains surprising. Advanced open-weight generation models such as MiniMax H3 show that complex multimodal systems can run locally and provide a pathway to commercial use under defined conditions. H3 does, however, impose licensing conditions and requires formal authorization for local deployment in regions including the European Union.
TTS development may appear more accessible than training a large multimodal video model, yet local speech synthesis still lacks a permissively licensed system that clearly rivals ElevenLabs across quality, control and features—or consistently comes close to the complete service.
Breeze TTS 2 therefore represents meaningful progress and a compelling model for research, testing and non-commercial local projects. Its technical capabilities are open for developers to explore, but its license prevents it from becoming the unrestricted commercial alternative many creators are still waiting for.
Your comments enrich our articles, so don’t hesitate to share your thoughts! Sharing on social media helps us a lot. Thank you for your support!
