Model reference · open weights

HARC-Llama-3.1

HARC-Llama-3.1 is an open-weight language model from microsoft, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

LLMs microsoft 1 variants 531 downloads/mo
Request this model on EU hardware All served models Not on the shared API today — deployed on request.

About

What HARC-Llama-3.1 is

HARC — Llama-3.1-8B-Instruct HARC safety-alignment LoRA merged into meta-llama/Llama-3.1-8B-Instruct (full standalone model). Part of the HARC release; see paper arXiv:2607.00572.

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makermicrosoft
TypeLanguage models
Parameters (lead)8.0B
Variants1
Runs withtransformers
Based onmeta-llama/Llama-3.1-8B-Instruct
Released2026-07-02
Popularity531 downloads / month
Likes1
LicenceOpen, with conditions

How it works

How language models work

Your prompttext / messagesTransformerattention over tokensNext-token loopgenerate + streamResponsetext · tool callsA language model reads your tokens and predicts the next one, again and again, streaming the reply back.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
HARC-Llama-3.1-8B-Instruct8.0BBF16~18.5 GBWeights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys harc-llama-3-1 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (harc-llama-3-1 below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/chat/completions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"harc-llama-3-1","messages":[{"role":"user","content":"Hello"}]}'

Details

Languages, data & research

Tags

transformers safetensors llama text-generation safety alignment conversational text-generation-inference endpoints_compatible

Papers

Licence

Open, with conditions

Open weights under llama3.1, which carries conditions (e.g. attribution or an acceptable-use clause). Worth a read before production use — we can walk you through it. Read the licence ↗

Sources

Weights & code

Want HARC-Llama-3.1 on EU-owned hardware?

Request this model on EU hardware See what’s served now

Explore

More language models

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms