Model reference · open weights
higgs-tts-3 is an open-weight audio or speech model from bosonai, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.
About
Higgs TTS 3 Higgs TTS 3 is built for voice chat: it speaks, not just reads. It turns model responses into expressive conversational speech across 100+ languages, with zero-shot voice cloning and inline control over emotion, style, prosody, pauses, and sound effects. [!TIP] Released for research and non-commercial use under the Boson Higgs TTS 3 Research and Non-Commercial License. Production, hosted APIs, embedding in a product/service, or reselling the model requires a separate commercial license. Prohibited: voice cloning without consent, impersonation, fraud, election deception, biometric surveillance, or any unlawful use. [!TIP] Free for digital creators — including monetized content. Under the license's Creator Use Grant, creators may use Higgs TTS 3 to make and monetize podcasts, videos, and social posts for free. The one requirement is to credit Boson AI's Higgs Audio — either in the audio or prominently in the accompanying text (e.g., the video description or show notes). Suggested credit: "This audio was created with Boson AI's Higgs Audio — https://www.boson.ai/higgs-audio". See the Creator Use section below. Higgs autoregressive decoder consumes interleaved text and audio tokens. Audio is encoded by the Higgs Tokenizer into 8 codebooks at 25 fps, staggered via a delay pattern, then mapped to backbone hidden states through a multi-codebook fused embedding. Output codes pass through a multi-codebook fused head, are de-delayed, and decoded back to waveform. Supported Languages The model reaches single-digit WER/CER on 102 languages, which split into two tiers. WER/CER under 5 — polished, production-quality (85) 🇿🇦 Afrikaans · 🇸🇦🇪🇬 Arabic · 🇦🇲 Armenian · 🇮🇳 Assamese · 🇪🇸 Asturian · 🇦🇿 Azerbaijani · 🇷🇺 Bashkir · 🇪🇸 Basque · 🇧🇾 Belarusian · 🇧🇩🇮🇳 Bengali · 🇧🇦 Bosnian · 🇧🇬 Bulgarian · 🇪🇸 Catalan · 🇵🇭 Cebuano · 🇮🇶 Central Kurdish · 🇨🇳 Chinese · 🇭🇷 Croatian · 🇨🇿 Czech · 🇩🇰 Danish · 🇳🇱🇧🇪 Dutch · 🇷🇺 Eastern Mari · 🇺🇸🇬🇧🇦🇺 English · 🌐 Esperanto · 🇪🇪 Estonian · 🇫🇮 Finnish · 🇫🇷🇨🇦 French · 🇪🇸 Galician · 🇬🇪 Georgian · 🇩🇪🇦🇹 German · 🇬🇷 Greek · 🇮🇳 Gujarati · 🇭🇹 Haitian Creole · 🇳🇬 Hausa · 🇮🇱 Hebrew · 🇮🇳 Hindi · 🇭🇺 Hungarian · 🇮🇩 Indonesian · 🇮🇹 Italian · 🇯🇵 Japanese · 🇮🇩
Summarised from the published model card. Read the full card on the HuggingFace links below.
Specifications
| Maker | bosonai |
|---|---|
| Type | Audio & music |
| Parameters (lead) | 4.7B |
| Variants | 1 |
| Runs with | transformers |
| Released | 2026-06-04 |
| Popularity | 201k downloads / month |
| Likes | 735 |
| Licence | Commercial licence needed |
How it works
Variants
Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.
| Variant | Params | Precision | VRAM | Fits 16 GB | Weights |
|---|---|---|---|---|---|
| higgs-tts-3-4b | 4.7B | BF16 | ~10.7 GB | ✓ | Weights ↗ |
Using it via the API
Once AxForge deploys higgs-tts-3 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (higgs-tts-3 below is illustrative; you get the exact model name on deployment.)
$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \ -H "Authorization: Bearer $AXFORGE_API_KEY" \ -F model="higgs-tts-3" -F file=@audio.mp3
Details
Languages
Tags
Licence
The weights are open but its licence needs a commercial agreement for business use. AxForge can arrange that licence and host the model for you — you pay AxForge, we settle with the model’s maker. Ask us for a quote. Read the licence ↗