Model reference · open weights
BAAR2 is an open-weight language model from aixk. AxForge deploys and operates it for you on dedicated EU-owned hardware — with the licence handled where one is required.
Available as managed deployment — configured and operated for you on dedicated EU hardware, quoted per deployment.
What it is
| Released by | aixk |
|---|---|
| Type | Language models |
| Task | Text gen |
| Parameters (lead) | 168M |
| Context | 512 tokens |
| Released | 2026-09-12 |
| Popularity | 1k downloads / month |
| Licence | Commercial licence needed |
About
🌐 Official Website: https://baar.uk/model/ 📄 License: Non-Commercial Only (Contact for Commercial Pricing)
BAAR2-150M is a lightweight multilingual language model designed for fast, resource-efficient inference across 11 languages. Optimized for low latency and minimal compute environments, it enables responsive real-time generation on edge devices, mobile platforms, and resource-constrained servers.
BAAR2-150M 은 저사양 하드웨어 및 리소스 제약 환경에서도 빠르고 안정적인 동작이 가능하도록 설계된 1.5억(150M) 파라미터 규모의 초경량 다국어 언어 모델입니다. 지연 시간(Latency)을 최소화하여 온디바이스(Edge), 실시간 대화, 경량 RAG(검색 증강 생성) 시스템 등에 최적화되어 있습니다.
Hardware: Single NVIDIA Tesla T4 | FP16 Inference
| Task / Category | Prompt | Generated Output | Latency |
|---|---|---|---|
| Korean General QA | 대한민국의 수도는 어디인가요? | 대한민국의 수도는 서울입니다. 개인적으로, 이 지역의 수도가 너 | ~1451.0ms |
| English General QA | What is artificial intelligence? | A comprehensive machine that is designed to perform tasks that normally require | ~671.2ms |
| KO-to-EN Translation (1) | 안녕이 영어로 뭐야? | 안녕은 영어로 Hello 입니다. | ~390.0ms |
| KO-to-EN Translation (2) | 사과가 영어로 뭐야? | 사과는 영어로 apple 입니다. | ~359.4ms |
| EN-to-KO Translation | What is 'Seoul' in Korean? | A company that is a company that operates in the United | ~571.1ms |
| Japanese General QA | 日本の首都はどこですか? | 日本の首都は、 東京都双方にあり | ~661.0ms |
This benchmark demonstrates the end-to-end performance of coupling a real-time vector retrieval pipeline with BAAR2-150M generation. It highlights the ultra-low latency response times and context-utilization capabilities achievable within an ultra-lightweight 150M-parameter footprint.
실시간 벡터 검색(Retrieval)과 BAAR2-150M 생성을 결합한 RAG 파이프라인 테스트 결과입니다. 150M 파라미터 경량 모델의 극대화된 추론 속도와 외부 문맥(Context) 활용 능력을 보여줍니다.
===========================================================================
🚀 [Starting Real-Time RAG Q&A]
===========================================================================
[1/4] 👤 질문: 아이폰 16 프로 가장 싼 모델 얼마부터 시작해?
🔍 검색된 문서 (유사도 0.67 | ⏱️ 231.1ms):
"아이폰 16 프로는 항공우주 등급 티타늄 디자인으로 제작되었으며 출고가는 128GB 모델 기준 155만 원부터 시작합..."
🤖 AI 답변: 제 개인적인 견해로는, 아이폰 16은 가격이 높고 가볍지 (⏱️ 928.1ms)
---------------------------------------------------------------------------
[2/4] 👤 질문: 신입사원이 연차 쓰려면 며칠 전에 승인받아야 해?
🔍 검색된 문서 (유사도 0.79 | ⏱️ 52.5ms):
"사내 취업규칙 제18조에 따라 근속 1년 미만 신규 입사자는 매월 개근 시 1일씩 총 11일의 연차가 부여됩니다. 연..."
🤖 AI 답변: 1. 연차 쓰려면 2. 인트라넷을 사용하는 것이 좋 (⏱️ 530.3ms)
---------------------------------------------------------------------------
[3/4] 👤 질문: 지구가 자전하는 과학적 이유가 뭐야?
🔍 검색된 문서 (유사도 0.65 | ⏱️ 16.0ms):
"태양계 형성 시기인 약 45억 년 전 거대한 가스와 먼지로 이루어진 원시 성운이 중력으로 수축하면서 각운동량 보존 법..."
🤖 AI 답변: 우주 공간에는 약 450,000년이 있습니다. (⏱️ 421.5ms)
---------------------------------------------------------------------------
[4/4] 👤 질문: 스페인의 수도는 어디인가요?
🔍 검색된 문서 (유사도 0.81 | ⏱️ 11.3ms):
"스페인은 남유럽 이베리아반도에 위치한 입헌군주제 국가입니다. 국토 면적은 약 50만 제곱킬로미터에 달하며, 국가의 행..."
🤖 AI 답변: 스페인의 수도는 마드리드입니다. (⏱️ 271.8ms)
---------------------------------------------------------------------------
Non-Commercial Use Only (비상업적 연구·개인 이용 한정): This model is strictly available for non-commercial, educational, and research purposes only. Any form of monetization (direct or indirect) without an explicit commercial license is strictly prohibited. 본 모델은 순수 연구, 교육 및 개인 비영리 목적에 한해 무료로 사용할 수 있습니다. 사전 허가 없는 일체의 수익 창출 행위(직·간접적 상업적 이용)는 엄격히 금지됩니다.
Monetization & Commercial Licensing Inquiries (수익 창출 및 상용 라이선스 문의): If you wish to monetize, integrate this model into commercial products, provide paid services/APIs, or require enterprise-level support, you must obtain a separate commercial license. Please reach out via email for licensing terms and enterprise adoption. 본 모델을 활용하여 수익을 창출하고자 하거나, 상용 서비스/유료 제품에 탑재, 또는 기업용 맞춤 적용 및 기술 지원이 필요한 경우 반드시 별도의 상용 라이선스를 취득하셔야 합니다. 도입 및 라이선스 발급은 아래 이메일로 문의해 주시기 바랍니다.
This project demonstrates practical competency in custom model architecture design, end-to-end distributed training optimization, and efficient multi-language serving under compute constraints.
I am actively seeking AI Engineering / Research opportunities, team recruitment offers, and project investment/partnerships.
제한된 하드웨어 리소스 환경에서 고효율 다국어 모델 아키텍처를 직접 설계하고 훈련 파이프라인을 엔드투엔드로 구축할 수 있는 AI 엔지니어입니다. 저의 기술적 역량 영입(채용)이나 프로젝트 협업/투자에 관심이 있으신 기업 및 팀의 연락을 기다립니다.
From the published model card. Full card on the HuggingFace links in the sidebar.
Using it via the API
Once AxForge deploys baar2 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (baar2 below is illustrative; you get the exact model name on deployment.)
$ curl -sS https://api.axforge.ai/v1/chat/completions \
-H "Authorization: Bearer $AXFORGE_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"baar2","messages":[{"role":"user","content":"Hello"}]}'
Create an account — your API key is available in the console. 3M free tokens every 30 days with every new account.