Model reference · open weights

Higgs-Audio-Studio

Higgs-Audio-Studio is an open-weight audio or speech model from drbaph, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Licence fee required Audio drbaph 1 variants 200k downloads/mo
Request a licence + hosting quote All served models Not on the shared API today — deployed on request.

About

What Higgs-Audio-Studio is

Higgs Audio v3 Studio Runtime Files This repository hosts the downloadable runtime files for Higgs Audio v3 Studio, a Windows desktop app for local Higgs Audio v3 TTS, voice cloning, speech continuation, and multi-speaker generation. GitHub app repository: https://github.com/Saganaki22/Higgs-Audio-v3-Studio This repository is not the original upstream model release. It provides GGUF model builds, the Windows CUDA engine DLL package, checksums, and a manifest used by the desktop app downloader. What This Is Higgs Audio v3 Studio is a Rust/Tauri desktop application that runs a ported native C++/CUDA implementation of Higgs Audio v3 locally. The app provides: - Local TTS generation - Voice clone workflow - Continue speech workflow - Multi-speaker workflow - Speaker Gallery for reusable speaker identities - WAV/MP3 export - Local API server - API streaming endpoint - Whisper-assisted reference transcript workflow - Model/engine download UI - Hardware telemetry and VRAM diagnostics - Engine dependency diagnostics for missing CUDA/MSVC runtime DLLs App Download Use the desktop app from GitHub releases: https://github.com/Saganaki22/Higgs-Audio-v3-Studio/releases Source code: https://github.com/Saganaki22/Higgs-Audio-v3-Studio Runtime Repository Layout The app expects this Hugging Face repository layout: Available Downloads Recommended VRAM Engine Requirements The prebuilt engine package is intended for: - Windows x64 - NVIDIA RTX 30xx, 40xx, or 50xx GPU - CUDA 13 compatible NVIDIA driver - Higgs Audio v3 Studio 0.2.31 or newer recommended The engines/ folder contains the app engine DLL plus the CUDA/MSVC runtime DLLs the current Windows engine build needs. Important: - nvcuda.dll is not included and should not be uploaded here. - nvcuda.dll comes from the NVIDIA display driver. - Users still need a working NVIDIA driver installed. - Users should not need the full CUDA Toolkit installed if they use the app's Download Engine DLLs button. The app can use either: 1. DLLs already installed on the user's system, such as CUDA/MSVC runtime DLLs found through system paths. 2. DLLs downloaded from this repository into the app's writable engine folder. Engine Dependency Diagnost

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makerdrbaph
TypeAudio & music
Variants1
Released2026-07-02
Popularity200k downloads / month
Likes6
LicenceCommercial licence needed

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
Higgs-Audio-v3-StudioBF16Weights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys higgs-audio-studio for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (higgs-audio-studio below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="higgs-audio-studio" -F file=@audio.mp3

Details

Languages, data & research

Languages

af ar as ast az ba be bg bn bs ca ceb ckb cs

Tags

gguf text-to-speech speech-generation voice-cloning expressive-speech controllable-tts multilingual-tts cpp windows whisper.cpp cuda higgs-audio-v3 af ar

Licence

Commercial licence needed

The weights are open but its licence needs a commercial agreement for business use. AxForge can arrange that licence and host the model for you — you pay AxForge, we settle with the model’s maker. Ask us for a quote. Read the licence ↗

Sources

Weights & code

Want Higgs-Audio-Studio on EU-owned hardware?

Request a licence + hosting quote See what’s served now

Explore

More audio & music

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms