Model reference · open weights

parler-tts-mini-multilingual

parler-tts-mini-multilingual is an open-weight audio or speech model from parler-tts, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Audio parler-tts 1 variants 125k downloads/mo
Request this model on EU hardware All served models Not on the shared API today — deployed on request.

About

What parler-tts-mini-multilingual is

Parler-TTS Mini Multilingual v1.1 Parler-TTS Mini Multilingual v1.1 is a multilingual extension of Parler-TTS Mini. 🚨 As compared to Mini Multilingual v1, this version was trained with some consistent speaker names and with better format for descriptions. 🚨 It is a fine-tuned version, trained on a cleaned version of CML-TTS and on the non-English version of Multilingual LibriSpeech. In all, this represents some 9,200 hours of non-English data. To retain English capabilities, we also added back the LibriTTS-R English dataset, some 580h of high-quality English data. Parler-TTS Mini Multilingual can speak in 8 European languages: English, French, Spanish, Portuguese, Polish, German, Italian and Dutch. Thanks to its better prompt tokenizer, it can easily be extended to other languages. This tokenizer has a larger vocabulary and handles byte fallback, which simplifies multilingual training. 🚨 This work is the result of a collaboration between the HuggingFace audio team and the Quantum Squadra team. The AI4Bharat team also provided advice and assistance in improving tokenization. 🚨 📖 Quick Index 👨‍💻 Installation 🎲 Using a random voice 🎯 Using a specific speaker Motivation Optimizing inference 🛠️ Usage 🚨Unlike previous versions of Parler-TTS, here we use two tokenizers - one for the prompt and one for the description.🚨 👨‍💻 Installation Using Parler-TTS is as simple as "bonjour". Simply install the library once: 🎲 Random voice Parler-TTS Mini Multilingual has been trained to generate speech with features that can be controlled with a simple text prompt, for example: 🎯 Using a specific speaker To ensure speaker consistency across generations, this checkpoint was also trained on 16 speakers, characterized by name (e.g. Daniel, Christine, Richard, Nicole, ...). To take advantage of this, simply adapt your text description to specify which speaker to use: Daniel's voice is monotone yet slightly fast in delivery, with a very close recording that almost has no background noise. You can choose a speaker from this list: Tips: We've set up an inference guide to make generation faster. Think SDPA, torch.compile, batching and streaming! Include the term "very clear audio" to gener

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makerparler-tts
TypeAudio & music
Parameters (lead)938M
Variants1
Runs withtransformers
Released2024-11-22
Popularity125k downloads / month
Likes58
LicenceOpen weights

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
parler-tts-mini-multilingual-v1.1938MBF16~2.2 GBWeights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys parler-tts-mini-multilingual for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (parler-tts-mini-multilingual below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="parler-tts-mini-multilingual" -F file=@audio.mp3

Details

Languages, data & research

Languages

en fr es pt pl de nl it

Trained / evaluated on

facebook/multilingual_librispeech parler-tts/libritts_r_filtered parler-tts/libritts-r-filtered-speaker-descriptions parler-tts/mls_eng parler-tts/mls-eng-speaker-descriptions ylacombe/mls-annotated ylacombe/cml-tts-filtered-annotated PHBJT/cml-tts-filtered PHBJT/mls-annotated PHBJT/cml-tts-filtered-annotated

Tags

transformers safetensors parler_tts text-generation text-to-speech annotation en fr es pt pl de nl it

Papers

Licence

Open weights

Open weights under apache-2.0 — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗

Sources

Weights & code

Want parler-tts-mini-multilingual on EU-owned hardware?

Request this model on EU hardware See what’s served now

Explore

More audio & music

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms