Model reference · open weights

LFM2-RAG

Available as managed deployment Licence fee LLMs LiquidAI Text gen 1 variants 2k dl/mo

LFM2-RAG is an open-weight language model from LiquidAI. AxForge deploys and operates it for you on dedicated EU-owned hardware — with the licence handled where one is required.

Available as managed deployment — configured and operated for you on dedicated EU hardware, quoted per deployment.

What it is

Released byLiquidAI
TypeLanguage models
TaskText gen
Runs withtransformers
Based onLiquidAI/LFM2-1.2B-RAG
Released2025-09-05
Popularity2k downloads / month
LicenceCommercial licence needed

About

What LFM2-RAG is

src="https://cdn-uploads.huggingface.co/production/uploads/61b8e2ba285851687028d395/2b08LKpev0DNEk6DlnWkY.png" alt="Liquid AI" style="width: 100%; max-width: 100%; height: auto; display: inline-block; margin-bottom: 0.5em; margin-top: 0.5em;" />

Read the full model card

LFM2-1.2B-RAG-GGUF

Based on LFM2-1.2B, LFM2-1.2B-RAG is specialized in answering questions based on provided contextual documents, for use in RAG (Retrieval-Augmented Generation) systems.

Use cases:

  • Chatbot to ask questions about the documentation of a particular product.
  • Custom support with an internal knowledge base to provide grounded answers.
  • Academic research assistant with multi-turn conversations about research papers and course materials.

You can find more information about other task-specific models in this blog post.

🏃 How to run LFM2

Example usage with llama.cpp:

llama-cli -hf LiquidAI/LFM2-1.2B-RAG-GGUF

From the published model card. Full card on the HuggingFace links in the sidebar.

How it works

How language models work

Your prompttext / messagesTransformerattention over tokensNext-token loopgenerate + streamResponsetext · tool callsA language model reads your tokens and predicts the next one, again and again, streaming the reply back.

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys lfm2-rag for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (lfm2-rag below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/chat/completions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"lfm2-rag","messages":[{"role":"user","content":"Hello"}]}'

Create an account — your API key is available in the console. 3M free tokens every 30 days with every new account.

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms