Speech generation · Opt-in

Breeze-TTS-2

A 3B model run by audio.cpp: English and Chinese cloning and voice design, with no Python.

  • BreezeBlueBy
  • 2Languages
  • 24 kHzOutput
  • ~4.73 GBDownload

Can do

  • Voice cloningYes. YesSpeaks in a voice copied from a short reference clip.
  • Voice designYes. YesBuilds a new voice from a written description.
  • Emotion controlNo. NoTakes graded emotion from a clip, a vector or text.
  • Preset voicesNo. NoShips fixed voices; no reference clip needed.

Languages

2languages
  • Chinesezh
  • Englishen

Runs on

  • audio.cpp server

    isolated
    audio.cppCUDAVulkanApple SiliconCPUROCm · build

    Prebuilts: Windows CPU, Vulkan and CUDA; Linux CPU and Vulkan (CUDA or ROCm from a self-build); macOS arm64 Metal.

    Engine doc

The Q8_0 GGUF is about 4.73 GiB; allow about 6 GB of dedicated VRAM.

Hardware

EngineCUDAROCmIntel ArcApple SiliconVulkanNPUCPU
audio.cpp serverSupportedNeeds a buildNot supportedSupportedSupportedNot supportedSupported

Supported Needs a build Experimental

Licence

Code
Apache-2.0 (audio.cpp code)
Weights
BreezeBlue Research and Non-Commercial License: research and non-commercial use only
  • Self-hosted outputs inherit the non-commercial restriction.

VoiceStudio is AGPL-3.0 and doesn't relicense model weights. Licence guide

Related

From the VoiceStudio app at 0834c8b. 27 models listed.

Quick answers