Model reference · open weights
Llama-4-Maverick-128E is an open-weight language model from meta-llama, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.
Specifications
| Maker | meta-llama |
|---|---|
| Type | Language models |
| Parameters (lead) | 401.6B |
| Variants | 2 |
| Runs with | transformers |
| Based on | meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 |
| Released | 2025-04-01 |
| Popularity | 78k downloads / month |
| Likes | 505 |
| Licence | Commercial licence needed |
How it works
Variants
Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.
Using it via the API
Once AxForge deploys llama-4-maverick-128e for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (llama-4-maverick-128e below is illustrative; you get the exact model name on deployment.)
$ curl -sS https://api.axforge.ai/v1/chat/completions \
-H "Authorization: Bearer $AXFORGE_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"llama-4-maverick-128e","messages":[{"role":"user","content":"Hello"}]}'
Licence
The weights are open but its licence needs a commercial agreement for business use. AxForge can arrange that licence and host the model for you — you pay AxForge, we settle with the model’s maker. Ask us for a quote. Read the licence ↗
Sources