Write the performance
Everything stays in one flow. Build the voice, render it, then listen below.
Define who is speaking
Start from a recipe or write a precise voice description from scratch.
Generation settings
The defaults are deliberate. Open the controls only when you need them.
Take B keeps the script and voice brief, then changes only the seed.
Age, register, pacing, emotional state, and recording context are more useful than a long adjective list.
Whisper gives you a fast first pass. Review names, punctuation, vocal cues, and every spoken word before cloning.
Enable take B to hear another stochastic interpretation without changing the underlying voice brief.
| Mode | Language | Text to speak | Voice description | Direction strength |
|---|
Breeze TTS 2 is an open-weight bilingual speech model from BreezeBlue. Reference transcription uses openai/whisper-large-v3-turbo.
- Model: BreezeBlue/Breeze-TTS-2
- Inference code: breezeblue-ai/breeze-tts
- Generation API:
/generate· transcription API:/transcribe_reference - Pinned model revision:
c1c8ca18b70b· Whisper:41f01f3fe87f
The Breeze inference source is Apache-2.0. Breeze model weights, derivative models, and self-hosted outputs are licensed for research and non-commercial use only. Whisper is MIT licensed. Commercial Breeze use requires written authorization from RESONIA, INC. Only upload voices you have the right and consent to use. Do not use generated speech for impersonation, deception, fraud, harassment, or authentication bypass.