Model reference · open weights
Fun-ASR-Nano is an open-weight audio or speech model from FunAudioLLM. AxForge deploys and operates it for you on dedicated EU-owned hardware — with the licence handled where one is required.
Available as managed deployment — configured and operated for you on dedicated EU hardware, quoted per deployment.
What it is
| Released by | FunAudioLLM |
|---|---|
| Type | Audio & music |
| Task | Speech→text |
| Runs with | gguf |
| Released | 2026-06-20 |
| Popularity | 3k downloads / month |
| Licence | Open weights |
About
GGUF build of Fun-ASR-Nano (SenseVoice SAN-M encoder + adaptor + Qwen3-0.6B LLM decoder) for the zero-Python, CPU/edge FunASR llama.cpp runtime — the accuracy leader (LLM decoder), single C++ binary.
The Fun-ASR-Nano LLM (Qwen3-0.6B) ships in three tiers — all within 0.1% CER (184-file micro-CER). Pair any with funasr-encoder-f16.gguf (470 MB).
| LLM file | size | CER ↓ | speed |
|---|---|---|---|
qwen3-0.6b-q4km.gguf | 484 MB | 8.35% | 6.1× |
qwen3-0.6b-q5km.gguf | 551 MB | 8.25% | 5.7× |
qwen3-0.6b-q8_0.gguf | 805 MB | 8.30% | 6.0× |
Recommended: q4_K_M (smallest) or q5_K_M (best).
These are GGUF weights for the FunASR llama.cpp runtime — a whisper.cpp-style, single self-contained binary for CPU / edge. Grab a prebuilt binary, then fetch this model and run:
runtime-llamacpp-v*)bash download-funasr-model.sh nano ./gguf
llama-funasr-cli --enc ./gguf/funasr-encoder-f16.gguf -m ./gguf/qwen3-0.6b-q8_0.gguf --vad ./gguf/fsmn-vad.gguf -a audio.wav
| file | size | notes |
|---|---|---|
funasr-encoder-f16.gguf | 470 MB | audio encoder + adaptor (f16) |
qwen3-0.6b-q8_0.gguf | 805 MB | LLM decoder, recommended (Q8_0) |
qwen3-0.6b-q4km.gguf | 484 MB | LLM decoder, smaller (Q4_K_M) |
llama-funasr-cli --enc funasr-encoder-f16.gguf -m qwen3-0.6b-q8_0.gguf -a audio.wav --vad fsmn-vad.gguf
On CPU: 8.30 % CER on the 184-clip Mandarin benchmark (vs whisper.cpp 22–31 %).
From the published model card. Full card on the HuggingFace links in the sidebar.
How it works
Using it via the API
Once AxForge deploys fun-asr-nano for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (fun-asr-nano below is illustrative; you get the exact model name on deployment.)
$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \ -H "Authorization: Bearer $AXFORGE_API_KEY" \ -F model="fun-asr-nano" -F file=@audio.mp3
Create an account — your API key is available in the console. 3M free tokens every 30 days with every new account.