Model reference · open weights

Olmo-3-1125

Olmo-3-1125 is an open-weight language model from allenai, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

LLMs allenai 1 variants 43k downloads/mo
Request this model on EU hardware All served models Not on the shared API today — deployed on request.

About

What Olmo-3-1125 is

Model Details Model Card for Olmo 3 32B We introduce Olmo 3, a new family of 7B and 32B models. This suite includes Base, Instruct, and Think variants. The Base models were trained using a staged training approach. Olmo is a series of Open language models designed to enable the science of language models. These models are trained on the Dolma 3 dataset. We are releasing all code, checkpoints, and associated training details. The core models released in this batch include the following: Installation Olmo 3 is supported in transformers v4.57.0 or higher: Inference You can use OLMo with the standard HuggingFace transformers library: For faster performance, you can quantize the model using the following method: The quantized model is more sensitive to data types and CUDA operations. To avoid potential issues, it's recommended to pass the inputs directly to CUDA using: We have released checkpoints for these models. For pretraining, the naming convention is stage1-stepXXX. The conventions for midtraining and long context are stage2-ingredientY-stepXXX and stage3-stepXXX, respectively. To load a specific model revision with HuggingFace, simply add the argument revision: Or, you can access all the revisions for the models via the following code snippet: Fine-tuning Model fine-tuning can be done from the final checkpoint (the main revision of this model) or many intermediate checkpoints. Two recipes for tuning are available. 1. Fine-tune with the OLMo-core repository: You can override most configuration options from the command-line. For example, to override the learning rate you could launch the script like this: For more documentation, see the GitHub readme. Model Description - Developed by: Allen Institute for AI (Ai2) - Model type: a Transformer style autoregressive language model. - Language(s) (NLP): English - License: The code and model are released under Apache 2.0. - Contact: Technical inquiries: olmo@allenai.org. Press: press@allenai.org - Date cutoff: Dec 2024 Model Sources - Project Page: https://allenai.org/olmo - Repositories: - Core repo (training, inference, fine-tuning etc.): https://github.com/allenai/OLMo-core - Evaluation code: https://github.com/alle

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makerallenai
TypeLanguage models
Parameters (lead)32.2B
Context64k tokens
Variants1
Runs withtransformers
Released2025-11-04
Popularity43k downloads / month
Likes127
LicenceOpen weights

How it works

How language models work

Your prompttext / messagesTransformerattention over tokensNext-token loopgenerate + streamResponsetext · tool callsA language model reads your tokens and predicts the next one, again and again, streaming the reply back.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
Olmo-3-1125-32B32.2BBF16~74.1 GBWeights ↗

Benchmarks

Reported results

As published on the model card — the maker's own numbers, not measured by AxForge.

TaskDatasetMetricScore
text-generationBenchmarksOlmo 3-Eval Math61.6
text-generationBenchmarksBigCodeBench43.9
text-generationBenchmarksHumanEval66.5
text-generationBenchmarksDeepSeek LeetCode1.9
text-generationBenchmarksDS 100029.7
text-generationBenchmarksMBPP60.2
text-generationBenchmarksMultiPL HumanEval35.9
text-generationBenchmarksMultiPL MBPPP41.8
text-generationBenchmarksOlmo 3-Eval Code40
text-generationBenchmarksARC MC94.7
text-generationBenchmarksMMLU STEM70.8
text-generationBenchmarksMedMCQA MC57.6
text-generationBenchmarksMedQA MC53.8
text-generationBenchmarksSciQ MC95.5
text-generationBenchmarksOlmo 3-Eval MC_STEM74.5
text-generationBenchmarksMMLU Humanities78.3
text-generationBenchmarksMMLU Social Sci.83.9
text-generationBenchmarksMMLU Other75.1
text-generationBenchmarksCSQA MC82.3
text-generationBenchmarksPIQA MC85.6
text-generationBenchmarksSocialIQA MC83.9
text-generationBenchmarksCoQA Gen2MC MC96.4
text-generationBenchmarksDROP Gen2MC MC87.2
text-generationBenchmarksJeopardy Gen2MC MC92.3

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys olmo-3-1125 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (olmo-3-1125 below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/chat/completions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"olmo-3-1125","messages":[{"role":"user","content":"Hello"}]}'

Details

Languages, data & research

Languages

en

Trained / evaluated on

allenai/dolma3_mix-5.5T-1125

Tags

transformers safetensors olmo3 text-generation en dataset:allenai/dolma3_mix-5.5T-1125 model-index endpoints_compatible

Licence

Open weights

Open weights under apache-2.0 — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗

Sources

Weights & code

Want Olmo-3-1125 on EU-owned hardware?

Request this model on EU hardware See what’s served now

Explore

More language models

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms