Model reference · open weights

higgs-tts-3

higgs-tts-3 is an open-weight audio or speech model from bosonai, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Licence fee required Audio bosonai 1 variants 201k downloads/mo
Request a licence + hosting quote All served models Not on the shared API today — deployed on request.

About

What higgs-tts-3 is

Higgs TTS 3 Higgs TTS 3 is built for voice chat: it speaks, not just reads. It turns model responses into expressive conversational speech across 100+ languages, with zero-shot voice cloning and inline control over emotion, style, prosody, pauses, and sound effects. [!TIP] Released for research and non-commercial use under the Boson Higgs TTS 3 Research and Non-Commercial License. Production, hosted APIs, embedding in a product/service, or reselling the model requires a separate commercial license. Prohibited: voice cloning without consent, impersonation, fraud, election deception, biometric surveillance, or any unlawful use. [!TIP] Free for digital creators — including monetized content. Under the license's Creator Use Grant, creators may use Higgs TTS 3 to make and monetize podcasts, videos, and social posts for free. The one requirement is to credit Boson AI's Higgs Audio — either in the audio or prominently in the accompanying text (e.g., the video description or show notes). Suggested credit: "This audio was created with Boson AI's Higgs Audio — https://www.boson.ai/higgs-audio". See the Creator Use section below. Higgs autoregressive decoder consumes interleaved text and audio tokens. Audio is encoded by the Higgs Tokenizer into 8 codebooks at 25 fps, staggered via a delay pattern, then mapped to backbone hidden states through a multi-codebook fused embedding. Output codes pass through a multi-codebook fused head, are de-delayed, and decoded back to waveform. Supported Languages The model reaches single-digit WER/CER on 102 languages, which split into two tiers. WER/CER under 5 — polished, production-quality (85) 🇿🇦 Afrikaans · 🇸🇦🇪🇬 Arabic · 🇦🇲 Armenian · 🇮🇳 Assamese · 🇪🇸 Asturian · 🇦🇿 Azerbaijani · 🇷🇺 Bashkir · 🇪🇸 Basque · 🇧🇾 Belarusian · 🇧🇩🇮🇳 Bengali · 🇧🇦 Bosnian · 🇧🇬 Bulgarian · 🇪🇸 Catalan · 🇵🇭 Cebuano · 🇮🇶 Central Kurdish · 🇨🇳 Chinese · 🇭🇷 Croatian · 🇨🇿 Czech · 🇩🇰 Danish · 🇳🇱🇧🇪 Dutch · 🇷🇺 Eastern Mari · 🇺🇸🇬🇧🇦🇺 English · 🌐 Esperanto · 🇪🇪 Estonian · 🇫🇮 Finnish · 🇫🇷🇨🇦 French · 🇪🇸 Galician · 🇬🇪 Georgian · 🇩🇪🇦🇹 German · 🇬🇷 Greek · 🇮🇳 Gujarati · 🇭🇹 Haitian Creole · 🇳🇬 Hausa · 🇮🇱 Hebrew · 🇮🇳 Hindi · 🇭🇺 Hungarian · 🇮🇩 Indonesian · 🇮🇹 Italian · 🇯🇵 Japanese · 🇮🇩

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makerbosonai
TypeAudio & music
Parameters (lead)4.7B
Variants1
Runs withtransformers
Released2026-06-04
Popularity201k downloads / month
Likes735
LicenceCommercial licence needed

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
higgs-tts-3-4b4.7BBF16~10.7 GBWeights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys higgs-tts-3 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (higgs-tts-3 below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="higgs-tts-3" -F file=@audio.mp3

Details

Languages, data & research

Languages

af ar as ast az ba be bg bn bs ca ceb ckb cs

Tags

transformers safetensors higgs_multimodal_qwen3 text-generation text-to-speech speech-generation voice-agent expressive-speech controllable-tts multilingual-tts af ar as ast

Licence

Commercial licence needed

The weights are open but its licence needs a commercial agreement for business use. AxForge can arrange that licence and host the model for you — you pay AxForge, we settle with the model’s maker. Ask us for a quote. Read the licence ↗

Sources

Weights & code

Want higgs-tts-3 on EU-owned hardware?

Request a licence + hosting quote See what’s served now

Explore

More audio & music

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms