Model reference · open weights
chatterbox is an open-weight audio or speech model from ResembleAI, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.
About
Latest Release: Chatterbox Multilingual V3 Chatterbox Multilingual V3 is the latest general-purpose multilingual TTS model in the Chatterbox family. It keeps the same 0.5B model size while improving speaker similarity, reducing hallucinations, and producing more natural, conversational speech across languages. V3 is designed for broad language coverage like V2, but with stronger stability and more expressive generation. It is the recommended multilingual model for users who want one voice cloning model that works across many languages. Try it in the Chatterbox Multilingual TTS V3 Space. Alongside V3, we are releasing the Single Language Pack: dedicated finetunes for priority languages where tighter quality control, stronger language-specific behavior, and more specialized speech generation are valuable. - Broad Multilingual Coverage: Designed as the main general-purpose multilingual Chatterbox model, supporting wide language coverage similar to V2. - Single Language Pack: Dedicated single-language models provide stronger specialization and quality control where language- and regional-dialect-specific performance matters most. - More Consistent Speaker Similarity: Improves voice identity and accent preservation across languages, making cross-language voice cloning more stable and reliable. - Reduced Hallucination: V3 is optimized to reduce unwanted continuation, repetition, and off-prompt speech, especially in cases where earlier multilingual models were less stable. Model Zoo Choose the right Chatterbox model for your application. 09/04 🔥 Introducing Chatterbox Multilingual in 23 Languages! We're excited to introduce Chatterbox and Chatterbox Multilingual, Resemble AI's production-grade open source TTS models. Chatterbox Multilingual supports Arabic, Danish, German, Greek, English, Spanish, Finnish, French, Hebrew, Hindi, Italian, Japanese, Korean, Malay, Dutch, Norwegian, Polish, Portuguese, Russian, Swedish, Swahili, Turkish, Chinese out of the box. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations. Whether you're working on memes, videos, games,
Summarised from the published model card. Read the full card on the HuggingFace links below.
Specifications
| Maker | ResembleAI |
|---|---|
| Type | Audio & music |
| Variants | 1 |
| Runs with | chatterbox |
| Released | 2025-04-24 |
| Popularity | 1.8M downloads / month |
| Likes | 1,774 |
| Licence | Open weights |
How it works
Variants
Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.
| Variant | Params | Precision | VRAM | Fits 16 GB | Weights |
|---|---|---|---|---|---|
| chatterbox | — | BF16 | — | — | Weights ↗ |
Using it via the API
Once AxForge deploys chatterbox for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (chatterbox below is illustrative; you get the exact model name on deployment.)
$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \ -H "Authorization: Bearer $AXFORGE_API_KEY" \ -F model="chatterbox" -F file=@audio.mp3
Details
Languages
Tags
Licence
Open weights under mit — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗