Model reference · open weights

truthfulqa-truth-judge-llama2

truthfulqa-truth-judge-llama2 is an open-weight language model from allenai, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

LLMs allenai 1 variants 27k downloads/mo
Request this model on EU hardware All served models Not on the shared API today — deployed on request.

About

What truthfulqa-truth-judge-llama2 is

This model is built based on LLaMa2 7B in replacement of the truthfulness/informativeness judge models that were originally introduced in the TruthfulQA paper. That model is based on OpenAI's Curie engine using their finetuning API. However, as of February 08, 2024, OpenAI has taken down its Curie engine, and thus, we cannot use it for TruthfulQA evaluation anymore. So, we decided to train the judge models using an open model (i.e., LLaMa), which can make the evaluation more accessible and reproducible. Released Models We released two models for the truthfulness and informativeness evaluation, respectively. Truthfulness Judge Informativenss Judge Training Details The training code and validation results of these models can be found here Usage These models are only intended for the TruthfulQA evaluation. They are intended to generalize to the evaluation of new models on the fixed set of prompts, but they may fail to generalize to new prompts. You can try the model using the following scripts:

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makerallenai
TypeLanguage models
Context4k tokens
Variants1
Runs withtransformers
Released2024-02-07
Popularity27k downloads / month
Likes6
LicenceOpen weights

How it works

How language models work

Your prompttext / messagesTransformerattention over tokensNext-token loopgenerate + streamResponsetext · tool callsA language model reads your tokens and predicts the next one, again and again, streaming the reply back.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
truthfulqa-truth-judge-llama2-7BBF16Weights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys truthfulqa-truth-judge-llama2 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (truthfulqa-truth-judge-llama2 below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/chat/completions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"truthfulqa-truth-judge-llama2","messages":[{"role":"user","content":"Hello"}]}'

Details

Languages, data & research

Languages

en

Trained / evaluated on

truthful_qa

Tags

transformers pytorch llama text-generation en dataset:truthful_qa text-generation-inference endpoints_compatible deploy:azure

Licence

Open weights

Open weights under apache-2.0 — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗

Sources

Weights & code

Want truthfulqa-truth-judge-llama2 on EU-owned hardware?

Request this model on EU hardware See what’s served now

Explore

More language models

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms