Catalogue · Language models

Open-source language models

1292 open-weight language models tracked from the labs releasing them, with who made each, its licence, what its weights weigh and the smallest GPU that runs it.

1292 models

Language models

Newest first. Search, filter with the dropdowns, or click any column heading to sort.

Model Maker Task Params Licence Runs on Released Downloads/mo
Qwen3-untiedNEW PrimeIntellect Text gen 752M Apache 2.0 RTX 3060 12 GB 2026-10-05 0
Qwen3.5-Reverse-Text-RLNEW PrimeIntellect Vision + text 1.1B Apache 2.0 RTX 3060 12 GB 2026-10-04 0
Qwen3.5-Reverse-Text-SFTNEW PrimeIntellect Vision + text 1.1B Apache 2.0 RTX 3060 12 GB 2026-10-04 0
DiarizationLM-Gemma-4NEW google Vision + text 8.0B Apache 2.0 RTX 3060 12 GB 2026-10-04 0
Kolibri-1NEW Aleph-Alpha Text gen 78.1B Apache 2.0 2× L40S 48 GB 2026-10-02 1k
PixelUMMNEW nvidia Image→text — Licence fee — 2026-10-01 0
TokleNEW techdotus Text gen 3M MIT RTX 3060 12 GB 2026-09-30 653
AREX-2NEW BAAI Vision + text 27.4B Apache 2.0 H100 80 GB 2026-09-29 0
ForgePlex-M2NEW ForgeWorks Text gen 10M Apache 2.0 RTX 3060 12 GB 2026-09-29 886
Tokle-SPABNEW techdotus Text gen 3M MIT RTX 3060 12 GB 2026-09-29 518
AdvancedMathBench-AutoVerifierNEW internlm Vision + text 36.0B — A100 80 GB 2026-09-29 0
Florence-2-large-FlashPackNEW fal Vision + text — MIT — 2026-09-29 5
IQuest-Q1NEW IQuestLab Text gen 320.3B Licence fee 4× B200 180 GB 2026-09-28 503
Naive-N0.5-Flash NaiveAI Text gen — MIT 4× B200 180 GB 2026-09-27 605
MiMo-Flash-MOPD XiaomiMiMo Text gen 310.8B MIT 2× H200 141 GB 2026-09-27 637
MiMo-Pro-MOPD XiaomiMiMo Text gen 1024.2B MIT 8× A100 80 GB 2026-09-27 763
Darwin-RSI FINAL-Bench Text gen 26.9B Apache 2.0 H100 80 GB 2026-09-27 619
Intern-Decision internlm Vision + text 4.5B Apache 2.0 RTX 3060 12 GB 2026-09-26 0
K2-Type IFM Text gen 1.1B Apache 2.0 RTX 3060 12 GB 2026-09-26 1k
clio-legacy NovelAI Text gen 3.0B Conditions RTX 3060 12 GB 2026-09-24 2k
Tev1-experimental togethercomputer Text gen 4.7B — RTX 4060 Ti 16 GB 2026-09-23 0
ThinkingCap-Qwen3.8 bottlecapai Vision + text 16.7B Licence fee 2× RTX 4060 Ti 16 GB 2026-09-23 755
tinctura bench-labs Text gen 96M Apache 2.0 RTX 3060 12 GB 2026-09-23 509
limite-violetto paradigma-inc Text gen 1.0B Apache 2.0 RTX 3060 12 GB 2026-09-21 1k
MiMo-Flash-RL XiaomiMiMo Text gen 310.8B MIT 2× H200 141 GB 2026-09-21 13k
MiMo-Pro-RL XiaomiMiMo Text gen 1024.2B MIT 8× A100 80 GB 2026-09-21 4k
Step-5 TypeSafeAI Vision + text 604.3B Licence fee 8× B200 180 GB 2026-09-20 524
LFM2.5-VL-DSpark LiquidAI Vision + text 279M Licence fee RTX 3060 12 GB 2026-09-18 546
GLM-5.3-Flash-Spark local-inference-lab Text gen 165.5B MIT 2× H200 141 GB 2026-09-17 8k
CORe-Pico-4 OpenCOReTechnologies Text gen 1.7B Apache 2.0 RTX 3060 12 GB 2026-09-16 885
Realtime-Venus inclusionAI Omni (any→any) — Apache 2.0 L40S 48 GB 2026-09-16 0
Xing4.0 XingChen-AGI Text gen · MoE 31.2B Apache 2.0 H100 80 GB 2026-09-16 3k
needle3 Cactus-Compute Text gen — Apache 2.0 RTX 3060 12 GB 2026-09-16 4k
olmo3-sdf-sft-clean150 EleutherAI Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2026-09-16 370
olmo3-sdf-sft-scrub-b1reset150 EleutherAI Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2026-09-16 363
llm-jp-4.1-thinking llm-jp Text gen · MoE 8.6B Apache 2.0 2× RTX 3060 12 GB 2026-09-15 921
open-sft Gensyn Text gen 1.6B Apache 2.0 RTX 3060 12 GB 2026-09-15 713
Evo2 Aquiles-ai Text gen 1.1B Apache 2.0 RTX 3060 12 GB 2026-09-14 892
gollem-pl SlayerLab Text gen — Licence fee RTX 5090 32 GB 2026-09-14 1k
Simple-Attention-Sparsification tencent Text gen — — RTX 3060 12 GB 2026-09-14 0
FlyGPT QuixiAI Text gen 2M Open RTX 3060 12 GB 2026-09-14 589
Ivme-Conversate IvmeLabs Text gen 25M Apache 2.0 RTX 3060 12 GB 2026-09-13 546
Aurora1.0 AuroraAI-Research Text gen 149M Apache 2.0 RTX 3060 12 GB 2026-09-13 506
NVIDIA-Nemotron-3-Super-UD-Q4_K_XL-MTPv2-layers meshllm Text gen · MoE — Licence fee 4× RTX 4090 24 GB 2026-09-12 5k
songgot-l palette-lab Text gen 303M Apache 2.0 RTX 3060 12 GB 2026-09-12 585
AliceAI-Foundation yandex Text gen · MoE 81.3B Apache 2.0 4× L40S 48 GB 2026-09-12 676
Underdog-Woof-1.1 ConwayResearch Text gen 4.2B Apache 2.0 RTX 3060 12 GB 2026-09-12 558
cagliostro bench-labs Text gen 146M Apache 2.0 RTX 3060 12 GB 2026-09-11 1k
AstaBrief_8B_SFT allenai Text gen — Apache 2.0 2× RTX 3060 12 GB 2026-09-10 0
Haidass-Translate DALabCommunity Text gen 143M Apache 2.0 RTX 3060 12 GB 2026-09-10 723
Ling-3.0-flash-Fin-fp4 inclusionAI Text gen 65.6B MIT H100 80 GB 2026-09-10 30
Ling-3.0-flash-Fin inclusionAI Text gen 127.5B MIT 2× L40S 48 GB 2026-09-09 13
MiniCPM5-DSpark openbmb Text gen — Apache 2.0 RTX 3060 12 GB 2026-09-09 430
tiny-aya-32K CohereLabs Text gen 3.4B Licence fee RTX 3060 12 GB 2026-09-08 0
glm-4-voice-of-reason-stitch kyutai Omni (any→any) 9.5B Licence fee 2× RTX 3060 12 GB 2026-09-08 0
Nex-N2.5-Pro nex-agi Text gen 396.8B Apache 2.0 4× H200 141 GB 2026-09-08 30k
Ling-3.0-flash-VL-fp4 inclusionAI Vision + text 64.5B MIT A100 80 GB 2026-09-08 597
Nex-N2.5-mini nex-agi Text gen 35.1B Apache 2.0 A100 80 GB 2026-09-08 2k
NV-Reason-CT nvidia Vision + text 5.3B Conditions RTX 4060 Ti 16 GB 2026-09-08 210
qwen3-djinnsdf-dolci EleutherAI Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2026-09-08 414
SupraNeo SupraLabs Text gen 4M Apache 2.0 RTX 3060 12 GB 2026-09-07 535
ZGCM-1 zgcagi Text gen 7.4B MIT 2× RTX 3060 12 GB 2026-09-07 1k
GigaChat3.5-Reasoning ai-sage Text gen · MoE — MIT 4× H100 80 GB 2026-09-07 1k
JustRL-II-model openbmb Text gen — — RTX 3060 12 GB 2026-09-07 0
Nex-N2.5-Max nex-agi Text gen 1600.8B Apache 2.0 more than the reference cards 2026-09-07 11k
LLaDA-UI inclusionAI Vision + text 16.9B — L40S 48 GB 2026-09-06 36
Surjo SurjoLabs Text gen 54M Apache 2.0 RTX 3060 12 GB 2026-09-05 590
LLaDA2.2-mini inclusionAI Text gen 16.3B Apache 2.0 L40S 48 GB 2026-09-05 0
Z1T-0 Extropic-AI Text gen — — — 2026-09-04 531
Ling-3.0-flash-VL inclusionAI Vision + text 124.8B MIT 2× H200 141 GB 2026-09-04 1k
NVIDIA-Nemotron-Labs-3-Competitive-Coding nvidia Text gen · MoE 335.0B Licence fee more than the reference cards 2026-09-04 0
ConsiSpace BAAI Vision + text 5.7B Open RTX 4060 Ti 16 GB 2026-09-04 0
Recon2Reason-Reasoning BAAI Vision + text 4.4B Apache 2.0 RTX 4060 Ti 16 GB 2026-09-03 0
Pebble basically-ai Text gen 26M Apache 2.0 RTX 3060 12 GB 2026-09-02 795
Nemotron-3-Labs-Ultra-Math-RL nvidia Text gen 560.5B Licence fee more than the reference cards 2026-09-02 335
Nemotron-3-Labs-Ultra-Math-SFT nvidia Text gen 560.5B Licence fee more than the reference cards 2026-09-02 470
tiny-aya-en-thinker CohereLabs Text gen 3.4B Licence fee RTX 3060 12 GB 2026-09-02 7
tiny-aya-l2-thinker CohereLabs Text gen 3.4B Licence fee RTX 3060 12 GB 2026-09-02 0
RWKV7-G1j-20260831 RWKV Text gen 13.3B Apache 2.0 RTX 5090 32 GB 2026-09-02 1k
O2 ant-intl Text gen 9.4B Apache 2.0 RTX 4090 24 GB 2026-09-02 615
K2-Horizon IFM Text gen · MoE 5.1B Apache 2.0 RTX 4060 Ti 16 GB 2026-09-01 1k
K2-Horizon-MoVA IFM Text gen · MoE 37.4B Apache 2.0 A100 80 GB 2026-09-01 1k
NVIDIA-NemotronLabs-AI-for-Media-Sports-Tennis nvidia Omni (any→any) 33.0B Licence fee H100 80 GB 2026-09-01 2
Myosotis-1 FWKV Text gen 102M Apache 2.0 RTX 3060 12 GB 2026-09-01 1k
ssiat-1.0 MOGODIK Text gen 255M Apache 2.0 RTX 3060 12 GB 2026-09-01 2k
jina-ocr jinaai Vision + text 3.4B Licence fee RTX 3060 12 GB 2026-09-01 595
Speck2 specklabs Text gen 141M MIT RTX 3060 12 GB 2026-09-01 731
Blaze-SFT SurjoLabs Text gen 48M MIT RTX 3060 12 GB 2026-09-01 1k
MiniCPM5-Midtrain openbmb Text gen 2.5B Apache 2.0 RTX 3060 12 GB 2026-09-01 1
Nanbeige4.2-DSpark Nanbeige Text gen 848M Apache 2.0 RTX 3060 12 GB 2026-08-31 542
DeepSeek-Flash-Vision-Exp deepseek-ai Vision + text 304.6B MIT 2× H200 141 GB 2026-08-31 0
Millie-11GB llmsforall Text gen · MoE — Apache 2.0 RTX 3060 12 GB 2026-08-30 727
ParallelTubeDecoding-Qwen3-VL MBZUAI Vision + text 4.4B Apache 2.0 RTX 4060 Ti 16 GB 2026-08-29 16
TLM McGill-NLP Text gen 229M Apache 2.0 RTX 3060 12 GB 2026-08-28 216
NCP_ArchPreview_dolma3_8.9B_Stage2 ArchSpace-Collection Text gen 8.9B Apache 2.0 RTX 4090 24 GB 2026-08-27 920
NCP_ArchPreview_dolma3_8.9B_Stage1 ArchSpace-Collection Text gen 8.9B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-08-27 630
Hy4 tencent Text gen 780.0B Apache 2.0 more than the reference cards 2026-08-27 2k
Qwen-Drive-1.0 Qwen Vision + text 4.5B Apache 2.0 RTX 3060 12 GB 2026-08-27 47
ContextPilot tencent Text gen 7.9B Licence fee 2× RTX 3060 12 GB 2026-08-27 0
CASA-Helium1-VL-Shared kyutai Vision + text 2.7B Licence fee RTX 3060 12 GB 2026-08-26 9
CASA-Qwen2_5-VL-Shared kyutai Vision + text 3.8B Licence fee RTX 3060 12 GB 2026-08-26 11
UI-Venus-2 inclusionAI Vision + text 1M — RTX 4090 24 GB 2026-08-26 1k
Moderato-Pro nitrai-research Text gen 113.3B Apache 2.0 2× H200 141 GB 2026-08-25 2k
GLM-5.3-Flash zai-org Text gen 321.3B MIT 4× RTX PRO 6000 Blackwell 96 GB 2026-08-25 346k
GLM-5.3 zai-org Text gen 753.3B Licence fee 8× H200 141 GB 2026-08-25 50k
Qwen3.8-Flash-Next Qwen Vision + text 180.0B Licence fee 4× H200 141 GB 2026-08-24 121k
Spark-X2.5 XHToken Text gen 4.1B Apache 2.0 RTX 3060 12 GB 2026-08-24 1k
Muse2 Muse-research Text gen 196M Apache 2.0 RTX 3060 12 GB 2026-08-23 802
Supra2-Medium SupraLabs Text gen 25M Apache 2.0 RTX 3060 12 GB 2026-08-21 1k
gpt2-custom EleutherAI Text gen 124M — RTX 3060 12 GB 2026-08-21 281
ArmorOCR inclusionAI Vision + text 8.8B Apache 2.0 2× RTX 3060 12 GB 2026-08-20 277
Aurora-80K AuroraAI-Research Text gen — Apache 2.0 RTX 3060 12 GB 2026-08-20 560
SenseNova-U1.5-MoT-SFT sensenova Omni (any→any) 17.5B Apache 2.0 L40S 48 GB 2026-08-19 552
LFM2.5-DSpark LiquidAI Text gen — Licence fee RTX 3060 12 GB 2026-08-19 105k
Qwen3.8-RTX5090-LMHead4 gittensor-model-hub Vision + text 15.0B Apache 2.0 RTX 4090 24 GB 2026-08-19 12k
SenseNova-U1.5-MoT sensenova Omni (any→any) 17.5B Apache 2.0 L40S 48 GB 2026-08-19 5k
Ornith-1.5 ornith-ai Text gen · MoE 19.5B MIT 2× RTX 4060 Ti 16 GB 2026-08-18 824k
Apodex-1.1-mini apodex Text gen 19.8B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-08-17 1k
Puro thu-pacman Text gen 2.0B Apache 2.0 RTX 3060 12 GB 2026-08-16 1k
llm-jp-4-vl llm-jp Vision + text 9.1B Apache 2.0 RTX 4090 24 GB 2026-08-15 1k
NVIDIA-Nemotron-Labs-Teacher nvidia Text gen 560.5B Licence fee more than the reference cards 2026-08-14 1k
NVIDIA-Nemotron-Labs-Teacher-Competition-Coding nvidia Text gen 560.5B Licence fee more than the reference cards 2026-08-14 2k
NVIDIA-Nemotron-Labs-Teacher-Instruction-Following nvidia Text gen 560.5B Licence fee more than the reference cards 2026-08-14 1k
NVIDIA-Nemotron-Labs-Teacher-General-Reasoning nvidia Text gen 560.5B Licence fee more than the reference cards 2026-08-14 600
NVIDIA-Nemotron-Labs-Teacher-STEM nvidia Text gen 560.5B Licence fee more than the reference cards 2026-08-14 1k
BananaMind-2-Pro BananaMind Text gen 160M Licence fee RTX 3060 12 GB 2026-08-14 1k
TeleOCR XingChen-AGI Vision + text 1.4B Apache 2.0 RTX 3060 12 GB 2026-08-14 27k
TeleOCR StarDoc-AI Vision + text 1.4B Apache 2.0 RTX 3060 12 GB 2026-08-14 16k
Qwen3.8-DSpark RadixArk Text gen 1.9B Licence fee RTX 3060 12 GB 2026-08-14 368k
UI-Mate-democua tencent Vision + text 27.4B Apache 2.0 H100 80 GB 2026-08-14 231
UI-Mate tencent Vision + text 9.4B Apache 2.0 RTX 4090 24 GB 2026-08-14 723
MathForm openbmb Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2026-08-14 777
Qwen3.8 Qwen Vision + text 27.8B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-08-13 5.1M
DeepSeek-Pro-0813 deepseek-ai Text gen 1650.5B MIT 8× H200 141 GB 2026-08-13 127k
Ling-3.0-flash-30T inclusionAI Text gen 127.5B MIT 2× H200 141 GB 2026-08-11 1k
Ling-3.0-flash-midtrain inclusionAI Text gen 127.5B MIT 2× H200 141 GB 2026-08-11 1k
North-Micro-Vision CohereLabs Vision + text 2.5B Apache 2.0 RTX 3060 12 GB 2026-08-10 31k
Muse-Glimmer-W4A4 Inferact Vision + text 6.0B Apache 2.0 RTX 5090 32 GB 2026-08-10 41k
gnani-evon gnani Text gen · MoE 31.7B Apache 2.0 H100 80 GB 2026-08-10 701
Muse-Glimmer-ExecuTorch-PTE meta-models Vision + text — Apache 2.0 RTX 3060 12 GB 2026-08-10 15k
Ling-3.0-tiny inclusionAI Text gen 7.9B MIT 2× RTX 3060 12 GB 2026-08-10 21k
Muse-Glimmer-assistant meta-models Vision + text 2.6B Apache 2.0 RTX 3060 12 GB 2026-08-09 27k
Muse-Glimmer meta-models Vision + text 29.8B Apache 2.0 H100 80 GB 2026-08-09 591k
Ling-3.0-flash-dspark inclusionAI Text gen 1.4B Licence fee RTX 3060 12 GB 2026-08-09 2k
title desert-ant-labs Text gen 352M Licence fee RTX 3060 12 GB 2026-08-09 974
dots3-note-prev dots-studio Vision + text 288.4B Apache 2.0 8× A100 80 GB 2026-08-09 2k
Qwen3.8-2.4T Qwen Text gen · MoE 2446.2B Licence fee more than the reference cards 2026-08-08 35k
granite-4.2 ibm-granite Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2026-08-07 7k
Azra-1-Mini OttomanNLP Image→text 853M Apache 2.0 RTX 3060 12 GB 2026-08-07 746
indic-ocr bodhan-ai Image→text — Licence fee RTX 3060 12 GB 2026-08-07 626
Motif-3 Motif-Technologies Text gen 314.8B MIT 4× B200 180 GB 2026-08-07 5k
BTL-4-Compact badtheorylabs Text gen — Apache 2.0 RTX 3060 12 GB 2026-08-06 133k
NVIDIA-Nemotron-3.5-Lightning-DSpark nvidia Text gen · MoE 764M Licence fee RTX 3060 12 GB 2026-08-05 191k
NVIDIA-Nemotron-3.5-Lightning-DFlash nvidia Text gen · MoE 663M Licence fee RTX 3060 12 GB 2026-08-05 3k
maple deepgrove Text gen 20.2B MIT L40S 48 GB 2026-08-04 4k
NVIDIA-Nemotron-3.5-Lightning nvidia Text gen · MoE 17.8B Licence fee RTX 4090 24 GB 2026-08-04 921k
glm-4-voice-of-reason kyutai Omni (any→any) 9.5B Licence fee 2× RTX 3060 12 GB 2026-08-04 0
Ling-3.0-flash-fp4 inclusionAI Text gen 65.6B MIT H100 80 GB 2026-08-04 8k
Supra2 SupraLabs Text gen 101M Apache 2.0 RTX 3060 12 GB 2026-08-03 26k
DFM-Mimir danish-foundation-models Text gen 1.8B Apache 2.0 RTX 3060 12 GB 2026-08-03 7k
JiRackUltra CMSManhattan Text gen 1.8B MIT RTX 3060 12 GB 2026-08-02 910k
Ling-3.0-flash inclusionAI Text gen 127.5B MIT 2× H200 141 GB 2026-08-02 19k
bagpiper-sft espnet Omni (any→any) — — RTX 4090 24 GB 2026-08-02 0
LFM2.5 LiquidAI Text gen · MoE — Licence fee RTX 3060 12 GB 2026-08-01 856k
DeepSeek-Flash-0731 deepseek-ai Text gen 304.2B MIT 4× L40S 48 GB 2026-07-31 4.6M
K-EXAONE-2.0-DSpark LGAI-EXAONE Text gen · MoE 751.4B Apache 2.0 more than the reference cards 2026-07-30 1k
Inkling-Small-DSpark RadixArk Text gen 1.4B — RTX 3060 12 GB 2026-07-30 35k
needle2 Cactus-Compute Text gen — Apache 2.0 — 2026-07-29 41k
Intern-S2-Mobius internlm Vision + text 36.0B Apache 2.0 A100 80 GB 2026-07-29 1k
ThinkingCap-Qwen3.6 bottlecapai Vision + text 16.7B Apache 2.0 RTX 4090 24 GB 2026-07-29 52k
K-EXAONE-2.0 LGAI-EXAONE Text gen · MoE 749.4B Apache 2.0 more than the reference cards 2026-07-29 4k
A.X-K2 skt Text gen 691.7B Apache 2.0 8× H200 141 GB 2026-07-28 118k
bagpiper espnet Omni (any→any) — — RTX 4090 24 GB 2026-07-28 0
Inkling-Small thinkingmachines Vision + text 156.0B Apache 2.0 2× H200 141 GB 2026-07-27 248k
Kimi-K3-DSpark RadixArk Text gen 2.2B — RTX 3060 12 GB 2026-07-27 3.3M
A.X-K2-Raon-Speech KRAFTON Omni (any→any) · MoE 21.2B Licence fee L40S 48 GB 2026-07-27 19k
Qwen3-VL amd Vision + text · MoE 118.8B Apache 2.0 B200 180 GB 2026-07-27 23k
Mage-VL microsoft Vision + text 4.7B Apache 2.0 RTX 4060 Ti 16 GB 2026-07-25 417k
Apertus swiss-ai Vision + text 8.9B Apache 2.0 RTX 4090 24 GB 2026-07-24 14k
Instella-MoE-Think amd Text gen · MoE 15.9B Licence fee L40S 48 GB 2026-07-23 2k
AREX BAAI Text gen 4.5B Apache 2.0 RTX 3060 12 GB 2026-07-23 748
KAT-Coder-Dev Kwaipilot Text gen 34.7B Apache 2.0 A100 80 GB 2026-07-23 48k
Macaron-Tall mindlab-research Text gen 36.0B MIT A100 80 GB 2026-07-22 555
VLX-Seek-1.5 omlab Vision + text 10.0B Apache 2.0 RTX 4090 24 GB 2026-07-22 33k
Solar-Open2 upstage Text gen 250.3B Licence fee 4× H200 141 GB 2026-07-22 24k
G9v3-39A5B ai9stars Text gen · MoE 39.0B Apache 2.0 2× L40S 48 GB 2026-07-21 2k
G9v3 ai9stars Text gen 3.0B Apache 2.0 RTX 3060 12 GB 2026-07-21 2k
HPD-Parsing PaddlePaddle Vision + text 1.1B Apache 2.0 RTX 3060 12 GB 2026-07-21 2k
Nanbeige4.2 Nanbeige Text gen 4.2B Apache 2.0 RTX 3060 12 GB 2026-07-21 34k
LLaDA2.2-flash inclusionAI Text gen 102.9B Apache 2.0 2× H200 141 GB 2026-07-16 1k
SenseNova-U1-MoT-Infographic sensenova Omni (any→any) 17.6B Apache 2.0 L40S 48 GB 2026-07-15 8k
Inkling thinkingmachines Vision + text 952.4B Apache 2.0 more than the reference cards 2026-07-14 308k
Hy-Embodied-VLM-1.0 tencent Vision + text 30.5B Apache 2.0 H100 80 GB 2026-07-14 446
Hy-Embodied-RxBrain-1.0 tencent Omni (any→any) 6.2B Apache 2.0 RTX 4060 Ti 16 GB 2026-07-14 147
GigaChat3.1-Audio ai-sage Text gen — MIT RTX 3060 12 GB 2026-07-13 280k
UniVR-Planning ByteDance Vision + text — Open B200 180 GB 2026-07-13 3
Agents-A1 InternScience Text gen 4.5B Apache 2.0 RTX 3060 12 GB 2026-07-13 332k
SingGuard-NSFA inclusionAI Vision + text 1.1B Apache 2.0 RTX 3060 12 GB 2026-07-10 953
HiLS-Attention tencent Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2026-07-09 369
Qwen3.6-Magic-Prompt fal Text gen · MoE 36.0B Apache 2.0 L40S 48 GB 2026-07-08 726
LightOn-rerank-PW lightonai Vision + text 4.5B Apache 2.0 RTX 3060 12 GB 2026-07-08 1k
LightOn-rerank-LW lightonai Vision + text 2.2B Apache 2.0 RTX 3060 12 GB 2026-07-08 873
GPT-2-wikitext-chunks EleutherAI Text gen 124M MIT RTX 3060 12 GB 2026-07-08 292
Nemotron-Labs-Audex nvidia Text gen — Licence fee RTX 4060 Ti 16 GB 2026-07-06 1k
Lumma FrontiersMind Text gen 649M Apache 2.0 RTX 3060 12 GB 2026-07-06 885
GigaChat3.5 ai-sage Text gen · MoE 433.8B MIT 4× H200 141 GB 2026-07-05 632
LongCat-2.0 meituan-longcat Text gen 1775.6B MIT more than the reference cards 2026-07-05 1k
Laguna-S-2.1 poolside Text gen 117.6B Conditions 4× RTX 5090 32 GB 2026-07-02 449k
HARC-Qwen2.5 microsoft Text gen 7.6B Apache 2.0 2× RTX 3060 12 GB 2026-07-02 517
HARC-Llama-3.1 microsoft Text gen 8.0B Conditions 2× RTX 3060 12 GB 2026-07-02 531
granite-swash ibm-granite Text gen 2.1B Apache 2.0 RTX 3060 12 GB 2026-07-01 6k
Soofi-S Soofi-Project Text gen 31.6B Licence fee H100 80 GB 2026-07-01 519
moondream3.1 moondream Vision + text · MoE 9.3B Licence fee RTX 3060 12 GB 2026-06-30 11k
NVIDIA-Nemotron-Parse-2.0 nvidia Vision + text 903M Conditions RTX 3060 12 GB 2026-06-30 25k
Laguna-XS-2.1 poolside Text gen — Conditions 2× RTX 3060 12 GB 2026-06-29 232k
react-native-executorch-pp-ocrv6 software-mansion Image→text — Apache 2.0 — 2026-06-29 554
react-native-executorch-easy-ocr software-mansion Image→text — Apache 2.0 — 2026-06-29 1k
DeepSeek-Pro-DSpark deepseek-ai Text gen 1650.5B MIT 8× H200 141 GB 2026-06-27 4k
DeepSeek-Flash-DSpark deepseek-ai Text gen 165.3B MIT 4× L40S 48 GB 2026-06-27 377k
NVIDIA-Nemotron-Labs-3-Puzzle nvidia Text gen · MoE 44.5B Licence fee H100 80 GB 2026-06-24 437k
Qwen-AgentWorld Qwen Text gen · MoE 34.7B Apache 2.0 A100 80 GB 2026-06-22 74k
Ornith-1.0 ornith-ai Text gen 1M MIT A100 80 GB 2026-06-21 3M
AfriqueQwen3.5-50Langs McGill-NLP Text gen 5.2B Open RTX 4060 Ti 16 GB 2026-06-20 1k
lift datalab-to Vision + text 9.7B Open RTX 4090 24 GB 2026-06-19 161k
Unlimited-OCR baidu Vision + text 3.3B MIT RTX 3060 12 GB 2026-06-19 3.1M
UniAR-SFT ShareLab-SII Image→text 9.6B Apache 2.0 RTX 4090 24 GB 2026-06-16 544
UniAR-RL ShareLab-SII Image→text 9.6B Apache 2.0 RTX 4090 24 GB 2026-06-16 788
GLM-5.2 zai-org Text gen 753.3B MIT more than the reference cards 2026-06-16 1.9M
Sa2VA-LLaVA-1.5 ByteDance Vision + text 7.3B Apache 2.0 2× RTX 3060 12 GB 2026-06-15 27
LFM2.5-ONNX LiquidAI Text gen — Licence fee — 2026-06-15 1k
MolmoMotion-H3-F30 allenai Vision + text 4.9B Apache 2.0 RTX 3060 12 GB 2026-06-15 331
MolmoMotion-H1-F32 allenai Vision + text 4.9B Apache 2.0 RTX 3060 12 GB 2026-06-15 228
Kimi-K3 moonshotai Vision + text 2779.9B Licence fee more than the reference cards 2026-06-13 2.8M
Kimi-K2.7-Code moonshotai Vision + text 1026.9B Licence fee 4× B200 180 GB 2026-06-11 250k
PP-OCRv6_tiny_rec PaddlePaddle Image→text — Apache 2.0 — 2026-06-11 2k
MobileMoE-S facebook Text gen · MoE 1.3B Licence fee RTX 3060 12 GB 2026-06-11 73
MobileMoE-M facebook Text gen · MoE 2.8B Licence fee RTX 3060 12 GB 2026-06-11 31
MobileMoE-M-SFT facebook Text gen · MoE 2.8B Licence fee RTX 3060 12 GB 2026-06-11 9
MobileMoE-S-SFT facebook Text gen · MoE 1.3B Licence fee RTX 3060 12 GB 2026-06-11 13
MobileMoE-S-QAT facebook Text gen · MoE 672M Licence fee RTX 3060 12 GB 2026-06-11 2
MobileMoE-M-QAT facebook Text gen · MoE 1.5B Licence fee RTX 3060 12 GB 2026-06-11 2
MobileMoE-L-QAT facebook Text gen · MoE 2.8B Licence fee RTX 3060 12 GB 2026-06-11 2
MobileMoE-L-SFT facebook Text gen · MoE 5.3B Licence fee RTX 3060 12 GB 2026-06-11 11
MobileMoE-L facebook Text gen · MoE 5.3B Licence fee RTX 3060 12 GB 2026-06-10 56
PP-OCRv6_small_rec PaddlePaddle Image→text — Apache 2.0 — 2026-06-10 9k
PP-OCRv6_medium_rec PaddlePaddle Image→text — Apache 2.0 — 2026-06-10 51k
PP-OCRv6_tiny_det PaddlePaddle Image→text — Apache 2.0 — 2026-06-10 1k
PP-OCRv6_small_det PaddlePaddle Image→text — Apache 2.0 — 2026-06-10 9k
PP-OCRv6_medium_det PaddlePaddle Image→text — Apache 2.0 — 2026-06-10 58k
Supra1.5-exp SupraLabs Text gen 52M Apache 2.0 RTX 3060 12 GB 2026-06-10 649
UltraX openbmb Text gen — Apache 2.0 RTX 3060 12 GB 2026-06-10 0
EvoQuality ByteDance Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2026-06-10 150
KDL-Frontier-Parser-nano KDLAI Vision + text 1.2B Licence fee RTX 3060 12 GB 2026-06-10 12k
Sa2VA-Qwen3-VL-SAM3 ByteDance Vision + text 5.3B Apache 2.0 RTX 4060 Ti 16 GB 2026-06-10 395
diffusiongemma google Vision + text · MoE 25.8B Apache 2.0 H100 80 GB 2026-06-09 1.4M
PP-OCRv6_tiny_rec_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-06-09 3k
PP-OCRv6_small_rec_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-06-09 10k
PP-OCRv6_medium_rec_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-06-09 70k
PP-OCRv6_medium_det_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-06-09 32k
PP-OCRv6_small_det_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-06-09 9k
PP-OCRv6_tiny_det_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-06-09 3k
PP-OCRv6_small_det_safetensors PaddlePaddle Image→text 2M Apache 2.0 RTX 3060 12 GB 2026-06-09 606
PP-OCRv6_medium_det_safetensors PaddlePaddle Image→text 22M Apache 2.0 RTX 3060 12 GB 2026-06-09 2k
PP-OCRv6_medium_rec_safetensors PaddlePaddle Image→text 19M Apache 2.0 RTX 3060 12 GB 2026-06-08 1k
PP-OCRv6_small_rec_safetensors PaddlePaddle Image→text 5M Apache 2.0 RTX 3060 12 GB 2026-06-08 634
gemma-4-qat-ct google Omni (any→any) 13.3B Apache 2.0 RTX 4060 Ti 16 GB 2026-06-05 1.4M
North-Mini-Code-1.0 CohereLabs Text gen 30.5B Apache 2.0 H100 80 GB 2026-06-05 22k
gemma-4-qat-q4_0 google Omni (any→any) — Apache 2.0 RTX 3060 12 GB 2026-06-05 670k
gemma-4-qat-q4_0-unquantized-assistant google Omni (any→any) · MoE 423M Apache 2.0 RTX 3060 12 GB 2026-06-04 53k
gemma-4-qat-q4_0-unquantized google Omni (any→any) · MoE 12.0B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-06-04 271k
Nex-N2-mini nex-agi Text gen 35.1B Apache 2.0 A100 80 GB 2026-06-04 1k
NVIDIA-Nemotron-3-Ultra nvidia Text gen · MoE 302.8B Licence fee more than the reference cards 2026-06-03 403k
gemma-4-qat-mobile-transformers google Omni (any→any) 2.3B Apache 2.0 RTX 3060 12 GB 2026-06-02 6k
MiniMax-M3 MiniMaxAI Vision + text 440.3B Licence fee 4× H200 141 GB 2026-06-02 370k
gemma-4-qat-mobile-ct google Omni (any→any) 5.5B Apache 2.0 RTX 3060 12 GB 2026-06-01 47k
Fast-dDrive Efficient-Large-Model Vision + text 235M Apache 2.0 RTX 3060 12 GB 2026-05-30 68
DeepSeek-OCR-2 deepseek-community Vision + text 3.4B Apache 2.0 RTX 3060 12 GB 2026-05-27 47k
PaddleOCR-VL-1.6 PaddlePaddle Vision + text 959M Apache 2.0 RTX 3060 12 GB 2026-05-27 32k
Mellum2-Thinking JetBrains Text gen 12.1B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-05-26 2k
Qwen3-ZH-Swap lightonai Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2026-05-26 9
Qwen3-ES-Swap lightonai Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2026-05-26 10
Qwen3-DE-Swap lightonai Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2026-05-26 11
Qwen3-FR-Swap lightonai Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2026-05-26 8
Mellum2 JetBrains Text gen 12.1B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-05-26 16k
Qwen3-SW-Swap lightonai Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2026-05-26 12
UVDoc_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 595
PP-OCRv5_mobile_rec_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 3k
PP-OCRv5_server_rec_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 604
PP-OCRv5_mobile_det_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 17k
PP-OCRv5_server_det_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 514
PP-DocLayout_plus-L_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 2k
PP-LCNet_x1_0_textline_ori_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 1k
PP-LCNet_x1_0_doc_ori_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 2k
PP-LCNet_x0_25_textline_ori_onnx PaddlePaddle Image→text — Apache 2.0 — 2026-05-26 642
LFM2.5-VL-Extract LiquidAI Vision + text 1.6B Licence fee RTX 3060 12 GB 2026-05-26 1k
LFM2.5-JP-202606 LiquidAI Text gen 1.2B Licence fee RTX 3060 12 GB 2026-05-26 15k
LFM2-Longevity LiquidAI Text gen 2.6B Licence fee RTX 3060 12 GB 2026-05-25 5k
SingGuard inclusionAI Vision + text 2.1B Apache 2.0 RTX 3060 12 GB 2026-05-25 1k
Step-3.7-Flash stepfun-ai Vision + text 201.4B Apache 2.0 4× H200 141 GB 2026-05-23 40k
MiniCPM5 openbmb Text gen 1.1B Apache 2.0 RTX 3060 12 GB 2026-05-21 881k
MiniCPM5-SFT openbmb Text gen 1.1B Apache 2.0 RTX 3060 12 GB 2026-05-21 15k
Supra SupraLabs Text gen 52M Apache 2.0 RTX 3060 12 GB 2026-05-21 1k
Qwen-Image-Bench Qwen Vision + text 27.4B Apache 2.0 H100 80 GB 2026-05-21 42k
MinerU2.5-Pro-2605 opendatalab Vision + text 1.2B Apache 2.0 RTX 3060 12 GB 2026-05-20 89k
BitCPM-CANN-unquantized openbmb Text gen — Apache 2.0 RTX 3060 12 GB 2026-05-18 10k
command-a-plus-05-2026-w4a4 CohereLabs Vision + text 218.8B Apache 2.0 2× H100 80 GB 2026-05-18 4k
HRM-Text sapientinc Text gen 1.2B Apache 2.0 RTX 3060 12 GB 2026-05-17 26k
BitCPM-CANN openbmb Text gen — Apache 2.0 RTX 3060 12 GB 2026-05-15 10k
Cola-DLM ByteDance-Seed Text gen — Apache 2.0 RTX 3060 12 GB 2026-05-15 113
Intern-S2 internlm Vision + text 36.1B Apache 2.0 A100 80 GB 2026-05-15 932
surya-ocr-2 datalab-to Vision + text 686M Open RTX 3060 12 GB 2026-05-14 1.3M
Prompt-Refine HiDream-ai Vision + text 31.3B MIT A100 80 GB 2026-05-13 389
MagenticBrain microsoft Text gen 14.8B MIT L40S 48 GB 2026-05-12 778
Fara1.5 microsoft Vision + text 9.4B MIT RTX 4090 24 GB 2026-05-12 5k
command-a-plus-05-2026 CohereLabs Vision + text 218.8B Apache 2.0 4× H200 141 GB 2026-05-11 41k
LLaVA-OneVision-2 lmms-lab-encoder Vision + text 8.5B Apache 2.0 2× RTX 3060 12 GB 2026-05-11 11k
Domyn-Small domyn Text gen 9.8B Licence fee 2× RTX 3060 12 GB 2026-05-09 689
MiniCPM-V-4.6-Thinking-BNB openbmb Vision + text 1.3B Apache 2.0 RTX 3060 12 GB 2026-05-09 36k
MiniCPM-V-4.6-BNB openbmb Vision + text 1.3B Apache 2.0 RTX 3060 12 GB 2026-05-09 43k
MiniCPM-V-4.6-Thinking openbmb Vision + text 1.3B Apache 2.0 RTX 3060 12 GB 2026-05-08 85k
Molmo2-ER allenai Vision + text 4.9B Apache 2.0 RTX 3060 12 GB 2026-05-04 7k
gemma-4-DFlash z-lab Text gen 1.5B Apache 2.0 RTX 3060 12 GB 2026-04-30 18k
Falcon3-prequantized tiiuae Text gen 10.3B Licence fee RTX 4090 24 GB 2026-04-30 322
OpenSeek-Mid BAAI Text gen 10.6B — RTX 4090 24 GB 2026-04-30 1
Ling-2.6-flash inclusionAI Text gen 107.5B MIT 2× H200 141 GB 2026-04-28 3k
MiMo XiaomiMiMo Text gen 310.8B MIT 4× RTX PRO 6000 Blackwell 96 GB 2026-04-27 316k
MiMo-Pro XiaomiMiMo Text gen 1023.2B MIT 8× B200 180 GB 2026-04-27 25k
Nemotron-3-Nano-Omni-Reasoning nvidia Omni (any→any) · MoE 18.3B Licence fee 2× RTX 4060 Ti 16 GB 2026-04-24 1.1M
nanowhale HuggingFaceTB Text gen 110M Apache 2.0 RTX 3060 12 GB 2026-04-24 949
gemma-4-assistant google Omni (any→any) · MoE 470M Apache 2.0 RTX 3060 12 GB 2026-04-23 545k
Qwen3.6-DFlash z-lab Text gen 1.7B MIT RTX 3060 12 GB 2026-04-23 150k
rnj-1.5 EssentialAI Text gen 8.3B Apache 2.0 2× RTX 3060 12 GB 2026-04-22 678
LLaDA2.0-Uni inclusionAI Omni (any→any) 16.3B Apache 2.0 L40S 48 GB 2026-04-22 4k
Falcon-E-prequantized tiiuae Text gen 3.1B Licence fee RTX 3060 12 GB 2026-04-22 149
DeepSeek-Pro deepseek-ai Text gen 1598.8B MIT 8× H200 141 GB 2026-04-22 843k
DeepSeek-Flash deepseek-ai Text gen 290.9B MIT 4× L40S 48 GB 2026-04-22 1.7M
SenseNova-U1-MoT sensenova Omni (any→any) 17.6B Apache 2.0 L40S 48 GB 2026-04-22 18k
AfriqueQwen3.5-ExtendedCM McGill-NLP Text gen 5.2B Open RTX 4060 Ti 16 GB 2026-04-20 201
AfriqueQwen3.5 McGill-NLP Text gen 5.2B Open RTX 4060 Ti 16 GB 2026-04-20 758
AfriqueQwen McGill-NLP Text gen 4.0B Open RTX 3060 12 GB 2026-04-20 1k
granite-guardian-4.1 ibm-granite Text gen 8.4B Apache 2.0 2× RTX 3060 12 GB 2026-04-16 68k
granite-vision-4.1 ibm-granite Vision + text 4.0B Apache 2.0 RTX 3060 12 GB 2026-04-16 106k
dialseg-ar-gemma3 MBZUAI Text gen 4.3B Licence fee RTX 3060 12 GB 2026-04-15 0
Qwen3.6 Qwen Vision + text · MoE 36.0B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-04-15 12.9M
Kimi-K2.6 moonshotai Vision + text 1026.9B Licence fee 4× B200 180 GB 2026-04-14 685k
MiniCPM-V-4.6 openbmb Vision + text 1.3B Apache 2.0 RTX 3060 12 GB 2026-04-13 580k
Hy3 tencent Text gen 298.8B Licence fee 8× A100 80 GB 2026-04-13 60k
MiniMax-M2.7 MiniMaxAI Text gen 228.7B Licence fee 2× H200 141 GB 2026-04-09 1.1M
Infinity-Parser2-Pro infly Vision + text 35.1B Apache 2.0 A100 80 GB 2026-04-08 189k
SuperApriel ServiceNow-AI Text gen — MIT DGX Spark (GB10) 128 GB unified 2026-04-07 126
A3-Qwen3.5 McGill-NLP Vision + text 9.4B — RTX 4090 24 GB 2026-04-06 1k
granite-4.1 ibm-granite Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2026-04-06 1.3M
DeepSeek-OCR-2 TitanML Vision + text 3.4B Apache 2.0 RTX 3060 12 GB 2026-04-06 13k
EXAONE-4.5 LGAI-EXAONE Vision + text 34.4B Licence fee H100 80 GB 2026-04-04 87k
GLM-5.1 zai-org Text gen 753.9B MIT 8× H200 141 GB 2026-04-03 396k
SmolLM3-GSM8K-SFT HuggingFaceTB Text gen 3.1B Apache 2.0 RTX 3060 12 GB 2026-04-03 317
MinerU2.5-Pro-2604 opendatalab Vision + text 1.2B Apache 2.0 RTX 3060 12 GB 2026-04-02 170k
nemotron-ocr nvidia Image→text — Licence fee RTX 3060 12 GB 2026-04-01 1k
CoME-VL MBZUAI Vision + text — Apache 2.0 L40S 48 GB 2026-04-01 21
react-native-executorch-lfm-2.5 software-mansion Vision + text — Licence fee — 2026-04-01 11k
Trinity-Large-Thinking-Block arcee-ai Text gen 398.7B Licence fee 4× H200 141 GB 2026-04-01 35
Trinity-Large-Thinking arcee-ai Text gen 398.6B Licence fee 8× H200 141 GB 2026-04-01 5k
Raon-Speech KRAFTON Omni (any→any) 9.0B Licence fee RTX 4090 24 GB 2026-03-30 3k
llm-jp-4 llm-jp Text gen 8.6B Apache 2.0 2× RTX 3060 12 GB 2026-03-30 5k
qwen3-gsm8k-sft HuggingFaceTB Text gen 1.7B Apache 2.0 RTX 3060 12 GB 2026-03-25 897
Trinity-Large-Block arcee-ai Text gen 398.7B Licence fee 4× H200 141 GB 2026-03-24 23
sarashina2.2-ocr sbintuitions Image→text 3.9B MIT RTX 3060 12 GB 2026-03-22 1k
Trinity-Mini-Block arcee-ai Text gen 26.1B Licence fee 2× RTX 4060 Ti 16 GB 2026-03-22 447
Trinity-Nano-Block arcee-ai Text gen 6.1B Licence fee RTX 3060 12 GB 2026-03-22 239
GigaChat3.1 ai-sage Text gen — MIT RTX 3060 12 GB 2026-03-21 6k
UniRG-CXR microsoft Vision + text 8.8B Apache 2.0 2× RTX 3060 12 GB 2026-03-19 976
dots.mocr-svg dots-studio Vision + text 3.0B MIT RTX 3060 12 GB 2026-03-19 1k
dots.mocr dots-studio Vision + text 3.0B MIT RTX 3060 12 GB 2026-03-19 627k
Nemotron-Cascade-2 nvidia Text gen · MoE 31.6B Licence fee H100 80 GB 2026-03-18 32k
Nemotron-Labs-Diffusion nvidia Text gen 8.5B Licence fee 2× RTX 3060 12 GB 2026-03-18 163k
PP-OCRv5_mobile_rec_safetensors PaddlePaddle Image→text 8M Apache 2.0 RTX 3060 12 GB 2026-03-18 756
Qianfan-OCR baidu Vision + text 4.7B Apache 2.0 RTX 4060 Ti 16 GB 2026-03-18 76k
chandra-ocr-2 datalab-to Vision + text 5.3B Open RTX 4060 Ti 16 GB 2026-03-16 2.9M
cyrillic-large-handwritten Kansallisarkisto Image→text 611M Apache 2.0 RTX 3060 12 GB 2026-03-16 1k
MolmoPoint allenai Vision + text 8.7B Apache 2.0 2× RTX 3060 12 GB 2026-03-16 7k
llm-jp-4-thinking llm-jp Text gen 8.6B Apache 2.0 2× RTX 3060 12 GB 2026-03-16 18k
Gemma-4 google Vision + text · MoE 31.3B Apache 2.0 RTX 4090 24 GB 2026-03-11 8.5M
reka-edge-2603 RekaAI Vision + text 7.1B Licence fee 2× RTX 3060 12 GB 2026-03-11 533
NVIDIA-Nemotron-3-Super nvidia Text gen · MoE 67.2B Licence fee DGX Spark (GB10) 128 GB unified 2026-03-10 1.1M
PP-Chart2Table_safetensors PaddlePaddle Image→text 561M Apache 2.0 RTX 3060 12 GB 2026-03-10 851
gemma-3-qat-q4_0-unquantized Lightricks Vision + text 12.2B Conditions RTX 5090 32 GB 2026-03-04 23k
granite-4.0-vision ibm-granite Vision + text 4.0B Apache 2.0 RTX 4060 Ti 16 GB 2026-03-03 12k
sarvam sarvamai Text gen 32.2B Apache 2.0 H100 80 GB 2026-03-03 14k
LocateAnything nvidia Vision + text 3.8B Licence fee RTX 3060 12 GB 2026-03-02 95k
Step-3.5-Flash-Midtrain stepfun-ai Text gen 197.8B Apache 2.0 4× H200 141 GB 2026-03-02 152
XCurOS0.1 XCurOS Text gen 7.6B Licence fee 2× RTX 3060 12 GB 2026-02-28 57k
Qwen3.5 Qwen Vision + text · MoE 9.7B Apache 2.0 RTX 4090 24 GB 2026-02-27 12.6M
MediX-R1 MBZUAI Vision + text — Licence fee 2× RTX 3060 12 GB 2026-02-27 1k
Infinity-Parser2-Flash infly Vision + text 2.2B Apache 2.0 RTX 3060 12 GB 2026-02-27 43k
XCurOS1.2-VLBF16 XCurOS Vision + text 8.8B Licence fee 2× RTX 3060 12 GB 2026-02-25 92k
Spatial-SSRL internlm Vision + text 4.1B Apache 2.0 RTX 3060 12 GB 2026-02-25 238
MedMO-Next MBZUAI Vision + text 8.8B Apache 2.0 2× RTX 3060 12 GB 2026-02-23 565
Falcon-OCR tiiuae Image→text 270M Apache 2.0 RTX 3060 12 GB 2026-02-22 3k
Aurora-Spec-Minimax-M2.5 togethercomputer Text gen 858M Apache 2.0 RTX 3060 12 GB 2026-02-19 272
OpenEarthAgent MBZUAI Text gen 4.0B Apache 2.0 RTX 3060 12 GB 2026-02-19 316
Olmo-Hybrid-SFT allenai Text gen 7.4B Apache 2.0 2× RTX 3060 12 GB 2026-02-19 6k
MiniMax-M2.5 PrimeIntellect Text gen 228.7B Licence fee 4× H200 141 GB 2026-02-18 659
minimax-m2-tiny PrimeIntellect Text gen 252M Apache 2.0 RTX 3060 12 GB 2026-02-18 113
qwen3-moe-tiny PrimeIntellect Text gen · MoE 670M Apache 2.0 RTX 3060 12 GB 2026-02-18 1k
glm4-moe-tiny PrimeIntellect Text gen · MoE 543M Apache 2.0 RTX 3060 12 GB 2026-02-18 243
WebWorld Qwen Text gen 1M Apache 2.0 H100 80 GB 2026-02-13 1k
tiny-aya-fire CohereLabs Text gen 3.3B Licence fee RTX 3060 12 GB 2026-02-13 723
tiny-aya-earth CohereLabs Text gen 3.3B Licence fee RTX 3060 12 GB 2026-02-13 1k
tiny-aya-water CohereLabs Text gen 3.3B Licence fee RTX 3060 12 GB 2026-02-13 371
tiny-aya-global CohereLabs Text gen 3.3B Licence fee RTX 3060 12 GB 2026-02-13 5k
tiny-aya CohereLabs Text gen 3.3B Licence fee RTX 3060 12 GB 2026-02-13 12k
Tucano2-qwen Polygl0t Text gen 3.8B Apache 2.0 RTX 3060 12 GB 2026-02-12 2k
Ovis2.6 ATH-MaaS Vision + text · MoE 31.4B Apache 2.0 H100 80 GB 2026-02-12 1k
MiniMax-M2.5 MiniMaxAI Text gen 228.7B Licence fee 2× H200 141 GB 2026-02-12 493k
MiniCPM-SALA openbmb Text gen 9.5B Apache 2.0 2× RTX 3060 12 GB 2026-02-11 26k
GLM-5 zai-org Text gen 753.9B MIT 8× H200 141 GB 2026-02-11 871k
Ring-2.5-1T inclusionAI Text gen 1012.5B MIT 8× B200 180 GB 2026-02-10 8k
Ming-flash-omni-2.0 inclusionAI Omni (any→any) 104.2B MIT 2× H200 141 GB 2026-02-10 3k
Nanbeige4.1 Nanbeige Text gen 3.9B Apache 2.0 RTX 3060 12 GB 2026-02-10 17k
AstaBrief allenai Text gen — Apache 2.0 2× RTX 3060 12 GB 2026-02-09 880
UI-Venus-1.5 inclusionAI Vision + text · MoE 2.4B Apache 2.0 RTX 3060 12 GB 2026-02-09 3k
LLaDA2.1-mini inclusionAI Text gen 16.3B Apache 2.0 L40S 48 GB 2026-02-09 105k
Baichuan-M3-Q4_K_M baichuan-inc Text gen — Apache 2.0 4× L40S 48 GB 2026-02-06 121
Baichuan-M2-Q4_K_M baichuan-inc Text gen — Apache 2.0 2× RTX 3060 12 GB 2026-02-06 706
MedMO MBZUAI Vision + text 4.4B Apache 2.0 RTX 4060 Ti 16 GB 2026-02-06 131
Aurora-Spec-Minimax-M2.1 togethercomputer Text gen 858M Apache 2.0 RTX 3060 12 GB 2026-02-04 52
MiniCPM-o-4_5 openbmb Omni (any→any) 9.4B Apache 2.0 RTX 4090 24 GB 2026-02-03 963k
Sequential_Helium kyutai Text gen 6.3B Open RTX 4060 Ti 16 GB 2026-02-03 240
KD-Tinker HuggingFaceH4 Text gen 8.2B — 2× RTX 3060 12 GB 2026-02-03 21
Aurora-Spec-Qwen3-Coder-Next togethercomputer Text gen 519M Apache 2.0 RTX 3060 12 GB 2026-02-03 151
Intern-S1-Pro internlm Vision + text — Apache 2.0 8× H200 141 GB 2026-02-02 76k
Qwen3-Coder-Next Qwen Text gen 79.7B Apache 2.0 2× L40S 48 GB 2026-02-01 1.8M
Step-3.5-Flash-Q4_K_S stepfun-ai Text gen — Apache 2.0 4× RTX 5090 32 GB 2026-02-01 1k
Step-3.5-Flash stepfun-ai Text gen 199.4B Apache 2.0 4× H200 141 GB 2026-02-01 155k
GLM-OCR zai-org Vision + text 1.3B MIT RTX 3060 12 GB 2026-01-30 2.2M
PaddleOCR-VL-1.5 PaddlePaddle Vision + text 959M Apache 2.0 RTX 3060 12 GB 2026-01-28 15k
Olmo-Hybrid allenai Text gen 7.4B Apache 2.0 2× RTX 3060 12 GB 2026-01-28 21k
Trinity-Large arcee-ai Text gen 398.6B Licence fee 8× H200 141 GB 2026-01-27 391
Trinity-Large-TrueBase arcee-ai Text gen 398.6B Licence fee 8× H200 141 GB 2026-01-27 234
DeepSeek-OCR-2 deepseek-ai Vision + text 3.4B Apache 2.0 RTX 3060 12 GB 2026-01-27 1.1M
EuroLLM-2512 utter-project Text gen 9.2B Apache 2.0 2× RTX 3060 12 GB 2026-01-26 29k
vllm-translategemma Infomaniak-AI Vision + text 5.0B Conditions RTX 3060 12 GB 2026-01-26 577k
Phi-4-reasoning-vision microsoft Vision + text 15.1B MIT L40S 48 GB 2026-01-23 6k
Youtu-VL tencent Vision + text 5.3B Licence fee RTX 4060 Ti 16 GB 2026-01-23 3k
GLM-4.7-Flash-REAP cerebras Text gen · MoE — MIT RTX 4060 Ti 16 GB 2026-01-23 26k
INTELLECT-3.1 PrimeIntellect Text gen 106.9B MIT 2× H200 141 GB 2026-01-20 740
GLM-4.7-Flash zai-org Text gen 31.2B MIT H100 80 GB 2026-01-19 2M
LightOnOCR-2-bbox-soup lightonai Vision + text 1.0B Apache 2.0 RTX 3060 12 GB 2026-01-16 14k
LightOnOCR-2-ocr-soup lightonai Vision + text 1.0B Apache 2.0 RTX 3060 12 GB 2026-01-16 3k
LightOnOCR-2-bbox lightonai Vision + text 1.0B Apache 2.0 RTX 3060 12 GB 2026-01-16 9k
LightOnOCR-2 lightonai Vision + text 1.0B Apache 2.0 RTX 3060 12 GB 2026-01-16 427k
PP-OCRv5_server_det_safetensors PaddlePaddle Image→text 22M Apache 2.0 RTX 3060 12 GB 2026-01-16 759
Stable-DiffCoder ByteDance-Seed Text gen 8.3B MIT 2× RTX 3060 12 GB 2026-01-15 403
PP-OCRv5_mobile_det_safetensors PaddlePaddle Image→text 4M Apache 2.0 RTX 3060 12 GB 2026-01-15 871
Unipic3 Skywork Omni (any→any) — MIT 2× RTX 3060 12 GB 2026-01-13 43
Falcon-H1-Tiny-Coder tiiuae Text gen 91M Licence fee RTX 3060 12 GB 2026-01-13 325
Step3-VL stepfun-ai Vision + text 10.2B Apache 2.0 2× RTX 4060 Ti 16 GB 2026-01-13 39k
Baichuan-M3 baichuan-inc Text gen 235.1B Apache 2.0 H200 141 GB 2026-01-13 5k
translategemma google Vision + text 5.0B Conditions RTX 3060 12 GB 2026-01-12 31k
Falcon-H1-Tiny-R-pre-GRPO tiiuae Text gen 622M Licence fee RTX 3060 12 GB 2026-01-12 254
Falcon-H1-Tiny-R tiiuae Text gen 622M Licence fee RTX 3060 12 GB 2026-01-12 4k
Falcon-H1-Tiny-Multilingual tiiuae Text gen 108M Licence fee RTX 3060 12 GB 2026-01-12 2k
Falcon-H1-Tiny tiiuae Text gen 91M Licence fee RTX 3060 12 GB 2026-01-12 1k
Unipic3-DMD Skywork Omni (any→any) — MIT 2× RTX 3060 12 GB 2026-01-12 26
Unipic3-Consistency-Model Skywork Omni (any→any) — MIT 4× RTX 4090 24 GB 2026-01-12 18
medgemma-1.5 google Vision + text 4.3B Licence fee RTX 3060 12 GB 2026-01-07 248k
AI21-Jamba2 ai21labs Text gen 3.0B Apache 2.0 RTX 3060 12 GB 2026-01-06 8k
LFM2.5-VL LiquidAI Vision + text 1.6B Licence fee RTX 3060 12 GB 2026-01-05 228k
Kimi-K2.5 moonshotai Vision + text 1026.9B Licence fee 4× B200 180 GB 2026-01-01 576k
Youtu-LLM tencent Text gen 2.0B Licence fee RTX 3060 12 GB 2025-12-31 11k
K-EXAONE LGAI-EXAONE Text gen · MoE 237.1B Licence fee 4× H200 141 GB 2025-12-26 21k
MAI-UI Tongyi-MAI Vision + text 8.8B Apache 2.0 2× RTX 3060 12 GB 2025-12-25 2k
CapRL-Qwen3VL internlm Vision + text 4.4B Apache 2.0 RTX 4060 Ti 16 GB 2025-12-24 3k
MiniCPM-o-2_6 FriendliAI Omni (any→any) 8.7B Apache 2.0 2× RTX 3060 12 GB 2025-12-24 637
HyperCLOVAX-SEED-Think naver-hyperclovax Text gen 33.3B Licence fee H100 80 GB 2025-12-23 13k
NousCoder NousResearch Text gen 14.8B Apache 2.0 L40S 48 GB 2025-12-23 480
GLM-4.7 zai-org Text gen 358.3B MIT 8× H200 141 GB 2025-12-22 68k
MiniMax-M2.1 MiniMaxAI Text gen 228.7B Licence fee 2× H200 141 GB 2025-12-20 31k
Helium1-VL kyutai Vision + text 2.7B Licence fee RTX 3060 12 GB 2025-12-18 44
MiMo-Flash XiaomiMiMo Text gen 309.8B MIT 2× B200 180 GB 2025-12-16 93k
EuroMoE-2512 utter-project Text gen · MoE 2.6B Apache 2.0 RTX 3060 12 GB 2025-12-15 539
paligemma-mix-224 fal Vision + text 2.9B Conditions RTX 3060 12 GB 2025-12-15 5k
Molmo2-O allenai Vision + text 7.8B Apache 2.0 2× RTX 3060 12 GB 2025-12-14 48k
Molmo2 allenai Vision + text 8.7B Apache 2.0 2× RTX 3060 12 GB 2025-12-14 121k
Cosmos-Reason2 nvidia Vision + text 2.4B Licence fee RTX 3060 12 GB 2025-12-12 995k
circuit-sparsity openai Text gen 419M Apache 2.0 RTX 3060 12 GB 2025-12-11 465
RealVideo zai-org Omni (any→any) — MIT H100 80 GB 2025-12-11 0
Olmo-3.1 allenai Text gen 32.2B Apache 2.0 H100 80 GB 2025-12-10 15k
CASA-Qwen2_5-VL kyutai Vision + text 4.1B Licence fee RTX 3060 12 GB 2025-12-10 47
CASA-Helium1-VL kyutai Vision + text 3.0B Licence fee RTX 3060 12 GB 2025-12-10 64
Olmo-3.1-Think allenai Text gen 32.2B Apache 2.0 H100 80 GB 2025-12-10 6k
Solar-Open upstage Text gen 102.7B Licence fee 2× H200 141 GB 2025-12-10 15k
RLVR-0926 stepfun-ai Text gen 8.2B MIT 2× RTX 3060 12 GB 2025-12-09 131
PaCoRe stepfun-ai Text gen 8.2B MIT 2× RTX 3060 12 GB 2025-12-09 137
AutoGLM-Phone-Multilingual zai-org Vision + text 1M MIT RTX 4090 24 GB 2025-12-09 437
AutoGLM-Phone zai-org Vision + text 1M MIT RTX 4090 24 GB 2025-12-08 25k
GLM-4.6V zai-org Vision + text 107.7B MIT 2× H200 141 GB 2025-12-07 7k
GLM-4.6V-Flash zai-org Vision + text 10.3B MIT RTX 4090 24 GB 2025-12-07 102k
nomos-1 NousResearch Text gen 30.5B Apache 2.0 H100 80 GB 2025-12-07 389
NVIDIA-Nemotron-3-Nano nvidia Text gen · MoE 31.6B Licence fee H100 80 GB 2025-12-04 876k
R1V4 Skywork Vision + text — MIT — 2025-12-03 0
Trinity-Nano arcee-ai Text gen 6.1B Licence fee RTX 4060 Ti 16 GB 2025-12-01 16k
Trinity-Mini arcee-ai Text gen 26.1B Licence fee 2× RTX 5090 32 GB 2025-12-01 20k
DeepSeek deepseek-ai Text gen 685.4B MIT 8× H200 141 GB 2025-12-01 1.4M
Apriel-1.6-Thinker ServiceNow-AI Vision + text 14.9B MIT L40S 48 GB 2025-11-28 5k
GELab-Zero stepfun-ai Vision + text 4.4B Apache 2.0 RTX 4060 Ti 16 GB 2025-11-28 220
Falcon-H1R tiiuae Text gen — Licence fee RTX 3060 12 GB 2025-11-28 3k
DeepSeek-Speciale deepseek-ai Text gen 685.4B MIT 8× H200 141 GB 2025-11-28 6k
Rax-4.5 raxcore-dev Vision + text 2.3B Apache 2.0 RTX 3060 12 GB 2025-11-27 1.1M
INTELLECT-3 PrimeIntellect Text gen 106.9B MIT 2× H200 141 GB 2025-11-26 8k
LLaDA2.0-flash inclusionAI Text gen 102.9B Apache 2.0 2× H200 141 GB 2025-11-25 3k
LLaDA2.0-mini inclusionAI Text gen 16.3B Apache 2.0 L40S 48 GB 2025-11-25 214k
Hermes-4.3 NousResearch Text gen — Apache 2.0 2× RTX 3060 12 GB 2025-11-25 9k
Ministral-3-2512-ONNX mistralai Vision + text — Apache 2.0 — 2025-11-24 408
Spatial-SSRL-Qwen3VL internlm Vision + text 4.8B Apache 2.0 RTX 4060 Ti 16 GB 2025-11-21 204
AprielGuard ServiceNow-AI Text gen 7.9B MIT 2× RTX 3060 12 GB 2025-11-21 3k
Olmo-3 allenai Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2025-11-19 289k
sarashina2.2-vision sbintuitions Image→text 3.8B MIT RTX 3060 12 GB 2025-11-19 521
Olmo-3-DPO allenai Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2025-11-19 15k
Olmo-3-Think-DPO allenai Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2025-11-18 7k
HunyuanOCR tencent Vision + text 1.1B Licence fee RTX 3060 12 GB 2025-11-18 684k
Olmo-3-Think allenai Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2025-11-18 72k
Foundation-Sec-1.1 fdtn-ai Text gen 8.0B Licence fee 2× RTX 3060 12 GB 2025-11-18 38k
Olmo-3-SFT allenai Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2025-11-17 44k
NVIDIA-Nemotron-Parse nvidia Vision + text 957M Licence fee RTX 3060 12 GB 2025-11-15 141k
Olmo-3-Think-SFT allenai Text gen 32.2B Apache 2.0 H100 80 GB 2025-11-14 42k
Bielik speakleash Text gen 11.2B Apache 2.0 RTX 4090 24 GB 2025-11-07 4k
ERNIE-4.5-VL-Thinking baidu Vision + text · MoE 29.7B Apache 2.0 H100 80 GB 2025-11-07 2k
Foundation-Sec-Reasoning fdtn-ai Text gen 8.0B Licence fee 2× RTX 3060 12 GB 2025-11-06 20k
Olmo-3-1125 allenai Text gen 32.2B Apache 2.0 H100 80 GB 2025-11-04 43k
Kimi-K2-Thinking moonshotai Text gen 1026.4B Licence fee 8× A100 80 GB 2025-11-04 43k
Emu3.5 BAAI Omni (any→any) 34.1B Apache 2.0 H100 80 GB 2025-10-31 399
xRouter Salesforce Text gen 7.6B Licence fee 2× RTX 3060 12 GB 2025-10-30 78
Kimi-Linear moonshotai Text gen · MoE 49.1B MIT DGX Spark (GB10) 128 GB unified 2025-10-30 181k
Ouro-Thinking ByteDance Text gen 2.7B Apache 2.0 RTX 3060 12 GB 2025-10-28 13k
Ouro ByteDance Text gen 1.4B Apache 2.0 RTX 3060 12 GB 2025-10-28 21k
Apriel-H1-Thinker-SFT ServiceNow-AI Text gen 15.7B MIT L40S 48 GB 2025-10-28 71
JanusCoderV internlm Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-10-27 581
t5gemma-2 google Vision + text 2.1B Conditions RTX 3060 12 GB 2025-10-25 50k
NVIDIA-Nemotron-Nano-VL-QAD nvidia Vision + text 7.7B Licence fee RTX 4060 Ti 16 GB 2025-10-22 10k
NVIDIA-Nemotron-Nano-VL nvidia Vision + text 13.2B Licence fee 2× RTX 3060 12 GB 2025-10-22 124k
MiniMax-M2 MiniMaxAI Text gen 228.7B Licence fee 2× H200 141 GB 2025-10-22 254k
Sa2VA-Qwen3-VL ByteDance Vision + text 5.1B Apache 2.0 RTX 4060 Ti 16 GB 2025-10-21 1k
chandra datalab-to Vision + text 8.8B Open 2× RTX 3060 12 GB 2025-10-21 17k
LightOnOCR-1025 lightonai Image→text 1.2B Apache 2.0 RTX 3060 12 GB 2025-10-20 95k
DeepSeek-OCR deepseek-ai Vision + text 3.3B MIT RTX 3060 12 GB 2025-10-17 2.4M
PaddleOCR-VL PaddlePaddle Vision + text 959M Apache 2.0 RTX 3060 12 GB 2025-10-16 9k
devanagari_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-10-16 1k
arabic_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-10-16 1k
cyrillic_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-10-16 984
ta_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-10-16 636
Ming-flash-omni inclusionAI Omni (any→any) 104.4B MIT 2× H200 141 GB 2025-10-14 1k
Qwen3-VL-Thinking Qwen Vision + text · MoE 8.8B Apache 2.0 2× RTX 3060 12 GB 2025-10-11 121k
Qwen3-VL Qwen Vision + text · MoE 8.8B Apache 2.0 2× RTX 3060 12 GB 2025-10-11 8.6M
KAT-Dev-Exp Kwaipilot Text gen 72.7B Apache 2.0 4× L40S 48 GB 2025-10-10 222
functiongemma google Text gen 268M Conditions RTX 3060 12 GB 2025-10-08 26k
AHN-Mamba2-for-Qwen-2.5 ByteDance-Seed Text gen 12M Licence fee RTX 3060 12 GB 2025-10-08 202
AHN-GDN-for-Qwen-2.5 ByteDance-Seed Text gen 13M Licence fee RTX 3060 12 GB 2025-10-08 70
AHN-DN-for-Qwen-2.5 ByteDance-Seed Text gen 51M Apache 2.0 RTX 3060 12 GB 2025-10-08 57
granite-4.0 ibm-granite Text gen 352M Apache 2.0 RTX 3060 12 GB 2025-10-07 17k
olmOCR-2-1025 allenai Vision + text 8.3B Apache 2.0 RTX 4060 Ti 16 GB 2025-10-06 276k
Ling-1T inclusionAI Text gen 999.7B MIT more than the reference cards 2025-10-02 3k
GTA1 Salesforce Vision + text 8.3B MIT 2× RTX 3060 12 GB 2025-10-01 220
SDLM-D8 OpenGVLab Text gen 3.4B Apache 2.0 RTX 3060 12 GB 2025-09-30 66
GLM-4.6 zai-org Text gen 356.8B MIT 8× H200 141 GB 2025-09-29 17k
SDLM-D4 OpenGVLab Text gen 3.4B Apache 2.0 RTX 3060 12 GB 2025-09-29 74
DeepSeek-Exp deepseek-ai Text gen 685.4B MIT 8× H200 141 GB 2025-09-29 89k
InternVL3_5-Flash OpenGVLab Vision + text · MoE 8.8B Apache 2.0 2× RTX 3060 12 GB 2025-09-28 508
HiPO Kwaipilot Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2025-09-26 189
CoDA Salesforce Text gen 2.0B Licence fee RTX 3060 12 GB 2025-09-25 787
Ring-mini-linear-2.0 inclusionAI Text gen 16.4B MIT L40S 48 GB 2025-09-24 3k
Apriel-1.5-Thinker ServiceNow-AI Vision + text 14.9B MIT L40S 48 GB 2025-09-24 352
Qwen3Guard-Gen Qwen Text gen 752M Apache 2.0 RTX 3060 12 GB 2025-09-23 160k
typhoon2.5-qwen3 typhoon-ai Text gen · MoE 30.5B Apache 2.0 H100 80 GB 2025-09-23 132k
GLM PrimeIntellect Text gen 543M MIT RTX 3060 12 GB 2025-09-22 3k
Qwen3-Fast PrimeIntellect Text gen · MoE 30.5B — H100 80 GB 2025-09-22 136
DeepSeek-Terminus deepseek-ai Text gen 684.5B MIT 8× H200 141 GB 2025-09-22 16k
Qwen3-Omni Qwen Omni (any→any) · MoE 35.3B Licence fee H100 80 GB 2025-09-20 910k
MiMo-Audio XiaomiMiMo Omni (any→any) 8.0B MIT 2× RTX 3060 12 GB 2025-09-18 5k
gpt-oss-safeguard openai Text gen 21.5B Apache 2.0 2× RTX 3060 12 GB 2025-09-18 79k
MinerU2.5-2509 opendatalab Vision + text 1.2B Licence fee RTX 3060 12 GB 2025-09-17 10k
granite-4.0-h-tiny ibm-granite Text gen 6.9B Apache 2.0 RTX 4060 Ti 16 GB 2025-09-16 95k
granite-4.0-h-small ibm-granite Text gen 32.2B Apache 2.0 H100 80 GB 2025-09-16 26k
granite-4.0-micro ibm-granite Text gen 3.4B Apache 2.0 RTX 3060 12 GB 2025-09-16 59k
granite-4.0-h-micro ibm-granite Text gen 3.2B Apache 2.0 RTX 3060 12 GB 2025-09-16 21k
Qianfan-VL baidu Vision + text 3.7B Licence fee RTX 3060 12 GB 2025-09-16 199
ScaleCUA OpenGVLab Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-09-16 69
Tongyi-DeepResearch Alibaba-NLP Text gen · MoE 30.5B Apache 2.0 H100 80 GB 2025-09-16 48k
Qwen3-Omni-Thinking Qwen Omni (any→any) · MoE 31.7B Licence fee H100 80 GB 2025-09-15 421k
Qwen3-Omni-Captioner Qwen Omni (any→any) · MoE 31.7B Licence fee H100 80 GB 2025-09-15 16k
KAT-Dev Kwaipilot Text gen 32.8B Apache 2.0 H100 80 GB 2025-09-15 404
Olmo-3-1025 allenai Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2025-09-12 114k
k2-merged-3.5T NousResearch Text gen 3468.1B Licence fee more than the reference cards 2025-09-12 330
moondream3 moondream Vision + text 9.3B Licence fee L40S 48 GB 2025-09-11 130k
Florence-2 florence-community Vision + text 232M MIT RTX 3060 12 GB 2025-09-11 64k
Florence-2-ft florence-community Vision + text 232M MIT RTX 3060 12 GB 2025-09-11 19k
Florence-2-large florence-community Vision + text 777M MIT RTX 3060 12 GB 2025-09-11 319k
Qwen3-Next-Thinking Qwen Text gen · MoE 81.3B Apache 2.0 4× L40S 48 GB 2025-09-09 45k
Qwen3-Next Qwen Text gen · MoE 81.3B Apache 2.0 4× L40S 48 GB 2025-09-09 289k
Lumina-DiMOO Alpha-VLLM Omni (any→any) 8.1B Apache 2.0 RTX 3060 12 GB 2025-09-09 1k
ERNIE-4.5-Thinking baidu Text gen · MoE 21.8B Apache 2.0 2× RTX 5090 32 GB 2025-09-08 16k
Ling-mini-2.0 inclusionAI Text gen 16.3B MIT L40S 48 GB 2025-09-08 12k
LFM2-RAG LiquidAI Text gen — Licence fee RTX 3060 12 GB 2025-09-05 2k
LFM2-Tool LiquidAI Text gen 1.2B Licence fee RTX 3060 12 GB 2025-09-03 1k
Kimi-K2-0905 moonshotai Text gen 1026.5B Licence fee 8× B200 180 GB 2025-09-03 39k
SmolLM3-QAT-Baseline-Q HuggingFaceH4 Text gen — — RTX 3060 12 GB 2025-09-02 18
MiniCPM4.1 openbmb Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2025-09-02 48k
Hermes-4 NousResearch Text gen 14.8B Apache 2.0 2× RTX 3060 12 GB 2025-08-29 7k
InternVL3_5-GPT-OSS OpenGVLab Vision + text · MoE 21.2B Apache 2.0 2× RTX 5090 32 GB 2025-08-29 71k
Step-Audio-2-mini stepfun-ai Omni (any→any) 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-08-28 2k
command-a-translate-08-2025 CohereLabs Text gen 111.1B Licence fee more than the reference cards 2025-08-27 23
InternVL3_5 OpenGVLab Vision + text · MoE 1.1B Apache 2.0 RTX 3060 12 GB 2025-08-25 81k
MiniCPM-V-4_5 openbmb Vision + text 8.7B Apache 2.0 2× RTX 3060 12 GB 2025-08-24 389k
LFM2-Extract LiquidAI Text gen 1.2B Licence fee RTX 3060 12 GB 2025-08-22 2k
th_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-08-21 508
en_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-08-21 501k
Seed-OSS ByteDance-Seed Text gen 36.2B Apache 2.0 A100 80 GB 2025-08-20 35k
Intern-S1-mini internlm Vision + text 8.5B Apache 2.0 2× RTX 3060 12 GB 2025-08-18 8k
OmniNeural NexaAI Omni (any→any) — Open — 2025-08-15 594
Ovis2.5 ATH-MaaS Vision + text 2.6B Apache 2.0 RTX 3060 12 GB 2025-08-15 3k
olmOCR-0825 allenai Vision + text 8.3B Apache 2.0 RTX 4060 Ti 16 GB 2025-08-13 18k
Apertus-2509 swiss-ai Text gen 8.1B Apache 2.0 2× RTX 3060 12 GB 2025-08-13 508k
UniPic2-Metaquery-GRPO-Flash Skywork Omni (any→any) — MIT RTX 4060 Ti 16 GB 2025-08-13 12
UniPic2-SD3-Kontext-GRPO Skywork Omni (any→any) — MIT 2× RTX 3060 12 GB 2025-08-13 11
UniPic2-Metaquery-GRPO Skywork Omni (any→any) — MIT RTX 3060 12 GB 2025-08-13 12
NVIDIA-Nemotron-Nano nvidia Text gen 8.9B Licence fee 2× RTX 3060 12 GB 2025-08-12 411k
Qwen3 PrimeIntellect Text gen 752M Apache 2.0 RTX 3060 12 GB 2025-08-12 53k
LFM2-VL LiquidAI Vision + text 1.6B Licence fee RTX 3060 12 GB 2025-08-12 43k
UniPic2-Metaquery-Flash Skywork Omni (any→any) — MIT RTX 4060 Ti 16 GB 2025-08-12 9
UniPic2-SD3-Kontext Skywork Omni (any→any) — MIT 2× RTX 3060 12 GB 2025-08-12 23
TowerVision utter-project Vision + text 3.0B Licence fee RTX 3060 12 GB 2025-08-11 148
UniPic2-Metaquery Skywork Omni (any→any) — MIT RTX 3060 12 GB 2025-08-11 21
GLM-4.5V zai-org Vision + text 107.7B MIT 2× H200 141 GB 2025-08-10 65k
Baichuan-M2 baichuan-inc Text gen 32.8B Apache 2.0 H100 80 GB 2025-08-10 1k
Qwen3-Thinking-2507 Qwen Text gen · MoE 4.0B Apache 2.0 RTX 3060 12 GB 2025-08-05 358k
Qwen3-2507 Qwen Text gen · MoE 4.0B Apache 2.0 RTX 3060 12 GB 2025-08-05 3.4M
gpt-oss openai Text gen 20.9B Apache 2.0 2× RTX 3060 12 GB 2025-08-04 6.5M
MindLink-0801 Skywork Text gen 32.8B Apache 2.0 H100 80 GB 2025-08-01 32
Llama-3_3-Nemotron-Super-v1_5 nvidia Text gen 49.9B Licence fee A100 80 GB 2025-07-31 253k
Qwen3-Coder Qwen Text gen · MoE 30.5B Apache 2.0 L40S 48 GB 2025-07-31 1.1M
dots.ocr dots-studio Vision + text 3.0B MIT RTX 3060 12 GB 2025-07-30 322k
AFM arcee-ai Text gen 4.6B Apache 2.0 RTX 3060 12 GB 2025-07-29 11k
EXAONE-4.0.1 LGAI-EXAONE Text gen 32.0B Licence fee H100 80 GB 2025-07-29 5k
Skywork-UniPic Skywork Omni (any→any) — MIT RTX 3060 12 GB 2025-07-29 25
step3 stepfun-ai Vision + text 321.0B Apache 2.0 8× H200 141 GB 2025-07-28 39k
command-a-vision-07-2025 CohereLabs Vision + text 111.9B Licence fee more than the reference cards 2025-07-28 35k
Meta-Llama-3.1-bnb fixie-ai Text gen 8.2B — RTX 3060 12 GB 2025-07-25 18
GLM-4.1V-Thinking THUDM Vision + text — MIT RTX 3060 12 GB 2025-07-25 4k
Intern-S1 internlm Vision + text 240.7B Apache 2.0 4× H200 141 GB 2025-07-24 11k
kanana-1.5-v kakaocorp Vision + text 3.7B Licence fee RTX 3060 12 GB 2025-07-23 28k
KAT Kwaipilot Text gen 40.6B Licence fee 4× RTX 4090 24 GB 2025-07-20 177
GLM-4.5-Air zai-org Text gen 110.5B MIT 2× H200 141 GB 2025-07-20 116k
GLM-4.5 zai-org Text gen 358.3B MIT 8× H200 141 GB 2025-07-20 96k
HiDream-E1-1 HiDream-ai Omni (any→any) 17.1B MIT 2× RTX 5090 32 GB 2025-07-16 96
Ming-Lite-Omni-1.5 inclusionAI Omni (any→any) 18.9B MIT L40S 48 GB 2025-07-15 1k
MiniCPM-V-4 openbmb Vision + text 4.1B Apache 2.0 RTX 3060 12 GB 2025-07-12 81k
EXAONE-4.0 LGAI-EXAONE Text gen 32.0B Licence fee H100 80 GB 2025-07-11 26k
Kimi-K2 moonshotai Text gen 1026.4B Licence fee 8× B200 180 GB 2025-07-11 163k
LFM2 LiquidAI Text gen · MoE 1.2B Licence fee RTX 3060 12 GB 2025-07-10 67k
SmolLM3-ONNX HuggingFaceTB Text gen — Apache 2.0 — 2025-07-08 214
ERNIE-4.5-2Bits-TP4-Paddle baidu Text gen · MoE — Apache 2.0 DGX Spark (GB10) 128 GB unified 2025-07-08 35
ERNIE-4.5-2Bits-TP2-Paddle baidu Text gen · MoE — Apache 2.0 DGX Spark (GB10) 128 GB unified 2025-07-08 46
Devstral-Small-2507 mistralai Text gen — Apache 2.0 RTX 4060 Ti 16 GB 2025-07-07 3k
Kimina-Prover AI-MO Text gen 2.0B Apache 2.0 RTX 3060 12 GB 2025-07-04 1k
A.X-4.0-Light skt Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2025-07-02 35k
eslav_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-07-02 4k
korean_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-07-02 5k
latin_PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-07-02 26k
DeepSWE agentica-org Text gen 32.8B MIT H100 80 GB 2025-07-01 1k
ERNIE-4.5-Paddle baidu Text gen · MoE 361M Apache 2.0 RTX 3060 12 GB 2025-06-29 213
ERNIE-4.5-VL-Paddle baidu Vision + text · MoE 423.5B Apache 2.0 8× H200 141 GB 2025-06-28 65
GLM-4.1V-Thinking zai-org Vision + text 10.3B MIT RTX 4090 24 GB 2025-06-28 228k
ERNIE-4.5-2Bits-Paddle baidu Text gen · MoE 91.7B Apache 2.0 DGX Spark (GB10) 128 GB unified 2025-06-28 56
ERNIE-4.5-W4A8C8-TP4-Paddle baidu Text gen · MoE — Apache 2.0 4× L40S 48 GB 2025-06-28 50
ERNIE-4.5-PT baidu Text gen · MoE 21.9B Apache 2.0 2× RTX 5090 32 GB 2025-06-28 28k
ERNIE-4.5-VL-PT baidu Vision + text · MoE 29.4B Apache 2.0 H100 80 GB 2025-06-28 60k
ShotVL Vchitect Vision + text 3.8B Apache 2.0 RTX 3060 12 GB 2025-06-27 3k
Hunyuan tencent Text gen · MoE 80.4B Licence fee 4× L40S 48 GB 2025-06-25 48k
WebDancer Alibaba-NLP Text gen 32.8B MIT H100 80 GB 2025-06-23 3k
Phi-tiny-MoE microsoft Text gen · MoE 3.8B MIT RTX 3060 12 GB 2025-06-23 294k
Kimi-VL-Thinking-2506 moonshotai Vision + text · MoE 16.4B MIT L40S 48 GB 2025-06-21 32k
Mistral-Small-3.2-2506 mistralai Vision + text — Apache 2.0 RTX 4060 Ti 16 GB 2025-06-20 40k
t5gemma-ul2 google Text gen 5.6B Conditions RTX 4060 Ti 16 GB 2025-06-19 50k
SmolLM3 HuggingFaceTB Text gen 3.1B Apache 2.0 RTX 3060 12 GB 2025-06-19 589k
deep-ignorance-e2e-strong-filter EleutherAI Text gen 6.9B Apache 2.0 2× RTX 3060 12 GB 2025-06-19 5k
Kimi-Dev moonshotai Text gen 72.7B MIT 4× L40S 48 GB 2025-06-16 1k
Nanonets-OCR-s nanonets Vision + text — — RTX 3060 12 GB 2025-06-16 9k
MiniMax-M1-80k MiniMaxAI Text gen 456.1B Apache 2.0 8× H200 141 GB 2025-06-13 900
gemma-3n google Vision + text 5.4B Conditions RTX 4060 Ti 16 GB 2025-06-12 210k
deepseek-vl deepseek-community Vision + text 2.0B Licence fee RTX 3060 12 GB 2025-06-12 21k
PP-LCNet_x1_0_textline_ori PaddlePaddle Image→text — Apache 2.0 — 2025-06-12 153k
PP-LCNet_x0_25_textline_ori PaddlePaddle Image→text — Apache 2.0 — 2025-06-12 1k
FlexOlmo-7x7B-1T allenai Text gen 33.3B Apache 2.0 H100 80 GB 2025-06-11 9k
Flex-reddit-2x7B-1T allenai Text gen 11.6B Apache 2.0 2× RTX 4060 Ti 16 GB 2025-06-11 5k
EuroMoE utter-project Text gen · MoE 2.6B Apache 2.0 RTX 3060 12 GB 2025-06-09 1k
gemma-3n-litert-lm google Text gen — Conditions — 2025-06-06 14k
OmniGen2 OmniGen2 Omni (any→any) 4.0B Apache 2.0 RTX 4090 24 GB 2025-06-06 2k
PP-DocLayout-L PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 1k
PP-LCNet_x1_0_table_cls PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 7k
RT-DETR-L_wireless_table_cell_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 7k
UVDoc PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 425k
PP-OCRv4_server_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 10k
en_PP-OCRv4_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 15k
SLANeXt_wired PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 7k
PP-FormulaNet_plus-S PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 604
RT-DETR-L_wired_table_cell_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 7k
PP-DocBlockLayout PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 4k
SLANet_plus PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 7k
PP-OCRv3_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 2k
PP-FormulaNet_plus-M PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 516
PP-OCRv4_mobile_seal_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 729
PP-OCRv4_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 3k
PP-FormulaNet_plus-L PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 2k
latin_PP-OCRv3_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 849
PP-OCRv4_mobile_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 4k
PP-DocLayout_plus-L PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 9k
en_PP-OCRv3_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 9k
PP-OCRv4_server_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 621
PP-LCNet_x1_0_doc_ori PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 338k
SLANeXt_wireless PaddlePaddle Image→text — Apache 2.0 — 2025-06-06 856
PP-Chart2Table PaddlePaddle Image→text — Apache 2.0 — 2025-06-05 6k
PP-OCRv3_mobile_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-05 10k
MiniMax-M1-40k MiniMaxAI Text gen 456.1B Apache 2.0 8× H200 141 GB 2025-06-05 5k
show-o2 showlab Omni (any→any) — Apache 2.0 RTX 3060 12 GB 2025-06-05 1k
Lingshu lingshu-medical-mllm Vision + text 8.3B MIT 2× RTX 3060 12 GB 2025-06-05 6k
MiniCPM4-MCP openbmb Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2025-06-05 14k
MiniCPM4 openbmb Text gen 434M Apache 2.0 RTX 3060 12 GB 2025-06-05 42k
PP-OCRv5_mobile_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-04 23k
PP-OCRv5_server_rec PaddlePaddle Image→text — Apache 2.0 — 2025-06-04 141k
PP-OCRv5_mobile_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-04 53k
PP-OCRv5_server_det PaddlePaddle Image→text — Apache 2.0 — 2025-06-04 740k
granite-guardian-3.3 ibm-granite Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2025-06-03 217k
granite-vision-3.3 ibm-granite Image→text 3.0B Apache 2.0 RTX 3060 12 GB 2025-06-03 254k
MiniMax-Text-01 MiniMaxAI Text gen 456.1B Licence fee 8× H200 141 GB 2025-06-03 17k
SynLogic MiniMaxAI Text gen 7.6B MIT 2× RTX 3060 12 GB 2025-06-03 502
Llama-3.1-Nemotron-Nano-VL nvidia Vision + text 8.7B Licence fee RTX 4090 24 GB 2025-06-03 342k
Fanar-1 QCRI Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2025-06-01 159k
MMaDA-MixCoT Gen-Verse Omni (any→any) 8.1B MIT RTX 4090 24 GB 2025-06-01 1k
SynLogic-Mix-3 MiniMaxAI Text gen 32.8B MIT H100 80 GB 2025-05-30 208
DeepSeek-R1-0528-Qwen3 deepseek-ai Text gen 8.2B MIT 2× RTX 3060 12 GB 2025-05-29 1M
DeepSeek-R1-0528 deepseek-ai Text gen 684.5B MIT 8× H200 141 GB 2025-05-28 197k
LLaDA-V GSAI-ML Vision + text 8.4B — 2× RTX 3060 12 GB 2025-05-28 14k
sarvam-m sarvamai Text gen 23.6B Apache 2.0 2× RTX 5090 32 GB 2025-05-20 4k
BAGEL-MoT ByteDance-Seed Omni (any→any) 14.7B Apache 2.0 RTX 3060 12 GB 2025-05-19 955
medgemma-text google Text gen 27.0B Licence fee H100 80 GB 2025-05-19 25k
medgemma google Vision + text 4.3B Licence fee RTX 3060 12 GB 2025-05-19 963k
medgemma-pt google Vision + text 4.3B Licence fee RTX 3060 12 GB 2025-05-19 1k
granite-docling ibm-granite Vision + text 258M Apache 2.0 RTX 3060 12 GB 2025-05-19 266k
MMaDA Gen-Verse Omni (any→any) 8.1B MIT RTX 4090 24 GB 2025-05-19 1k
aion polymathic-ai Omni (any→any) 318M MIT RTX 3060 12 GB 2025-05-16 2k
Ling-lite-1.5 inclusionAI Text gen 16.8B MIT L40S 48 GB 2025-05-11 10k
Apriel-Nemotron-Thinker ServiceNow-AI Text gen 15.0B MIT L40S 48 GB 2025-05-06 5k
M1 togethercomputer Text gen 3.4B MIT RTX 3060 12 GB 2025-05-02 1k
Falcon-H1-Deep tiiuae Text gen 1.6B Licence fee RTX 3060 12 GB 2025-05-01 10k
Falcon-H1 tiiuae Text gen 521M Licence fee RTX 3060 12 GB 2025-05-01 67k
VLM2Vec VLM2Vec Vision + text — Apache 2.0 RTX 3060 12 GB 2025-04-30 6k
granite-4.0-tiny ibm-granite Text gen 6.7B Apache 2.0 RTX 4060 Ti 16 GB 2025-04-30 74k
Qwen2.5-Omni Qwen Omni (any→any) 5.5B Licence fee RTX 4060 Ti 16 GB 2025-04-30 661k
Phi-4-mini-reasoning microsoft Text gen 3.8B MIT RTX 3060 12 GB 2025-04-29 59k
OLMo-2-0425-DPO allenai Text gen — Apache 2.0 RTX 3060 12 GB 2025-04-28 10k
Seed-Coder ByteDance-Seed Text gen 8.3B MIT 2× RTX 3060 12 GB 2025-04-27 3k
Qwen3 Qwen Text gen · MoE 752M Apache 2.0 RTX 4090 24 GB 2025-04-27 22.5M
HiDream-E1-Full HiDream-ai Omni (any→any) 17.1B MIT 2× RTX 5090 32 GB 2025-04-27 207
OLMo-2-0425-SFT allenai Text gen — Apache 2.0 RTX 3060 12 GB 2025-04-24 6k
Llama-Guard-4 meta-llama Vision + text 12.0B Licence fee RTX 5090 32 GB 2025-04-23 90k
DAM-Self-Contained nvidia Vision + text — Licence fee RTX 3060 12 GB 2025-04-21 17k
DAM nvidia Vision + text — Licence fee RTX 3060 12 GB 2025-04-21 14k
Cosmos-Reason1 nvidia Vision + text 8.3B Licence fee 2× RTX 3060 12 GB 2025-04-18 91k
InternVL2_5-MPO OpenGVLab Vision + text 2.2B — RTX 3060 12 GB 2025-04-18 7k
OLMo-2-0425 allenai Text gen 1.5B Apache 2.0 RTX 3060 12 GB 2025-04-17 891k
MAI-DS-R1 microsoft Text gen 671.0B MIT more than the reference cards 2025-04-16 561
UI-TARS-1.5 ByteDance-Seed Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-04-16 741k
bitnet-b1.58-4T microsoft Text gen 850M MIT RTX 3060 12 GB 2025-04-15 21k
Skywork-VL-Reward Skywork Vision + text 8.3B MIT RTX 4090 24 GB 2025-04-14 3k
Kimina-Autoformalizer AI-MO Text gen 7.6B Apache 2.0 2× RTX 3060 12 GB 2025-04-13 544
Eagle2.5 nvidia Vision + text 8.1B Licence fee RTX 4090 24 GB 2025-04-12 15k
Apriel ServiceNow-AI Text gen 4.8B MIT RTX 3060 12 GB 2025-04-11 1k
Qari-OCR-VL NAMAA-Space Vision + text 2.2B Apache 2.0 RTX 3060 12 GB 2025-04-10 84k
InternVL3 OpenGVLab Vision + text 938M Apache 2.0 RTX 3060 12 GB 2025-04-10 207k
Phi-4-reasoning microsoft Text gen 14.7B MIT L40S 48 GB 2025-04-09 27k
academic-ds ByteDance-Seed Text gen 9.4B Apache 2.0 2× RTX 3060 12 GB 2025-04-09 34k
granite-3.3 ibm-granite Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2025-04-09 55k
Kimi-VL-Thinking moonshotai Vision + text · MoE 16.4B MIT L40S 48 GB 2025-04-09 48k
Kimi-VL moonshotai Vision + text · MoE 16.4B MIT L40S 48 GB 2025-04-09 323k
gemma-3-qat-q4_0-unquantized google Vision + text 12.2B Conditions RTX 5090 32 GB 2025-04-08 69k
GLM-Z1-0414 zai-org Text gen 32.6B MIT H100 80 GB 2025-04-08 22k
GLM-4-0414 zai-org Text gen 9.4B MIT 2× RTX 3060 12 GB 2025-04-07 32k
DeepCoder agentica-org Text gen 1.8B MIT RTX 3060 12 GB 2025-04-07 2k
BaichuanMed-OCR baichuan-inc Vision + text 8.3B Licence fee 2× RTX 3060 12 GB 2025-04-07 279
Dream Dream-org Text gen 7.6B Apache 2.0 2× RTX 3060 12 GB 2025-04-03 96k
Llama-4-Scout-16E meta-llama Vision + text 108.6B Licence fee more than the reference cards 2025-04-02 247k
Llama-4-Maverick-128E meta-llama Vision + text 401.6B Licence fee more than the reference cards 2025-04-01 78k
qwen2.5-omni tiny-random Omni (any→any) 5M — RTX 3060 12 GB 2025-03-28 1k
Llama-xLAM-2-fc-r Salesforce Text gen 8.0B Licence fee 2× RTX 3060 12 GB 2025-03-27 19k
Mistral-Small-3.1-2503-dynamic mistralai Vision + text 24.0B Apache 2.0 RTX 5090 32 GB 2025-03-27 15k
DeepSeek-0324 deepseek-ai Text gen 684.5B MIT 8× H200 141 GB 2025-03-24 1.1M
Nemotron-H-8K nvidia Text gen 8.1B Licence fee 2× RTX 3060 12 GB 2025-03-19 71k
Skywork-R1V Skywork Vision + text 38.4B MIT 4× RTX 4090 24 GB 2025-03-17 36k
Llama-3.1-Nemotron-Nano nvidia Text gen 8.0B Licence fee 2× RTX 3060 12 GB 2025-03-16 218k
Llama-3_3-Nemotron-Super nvidia Text gen 49.9B Licence fee H200 141 GB 2025-03-16 178k
gemma-3 tiny-random Vision + text 9M — RTX 3060 12 GB 2025-03-15 10k
aya-vision CohereForAI Vision + text 33.1B Licence fee H100 80 GB 2025-03-14 12k
gemma-3-qat-q4_0 google Vision + text — Conditions RTX 3060 12 GB 2025-03-12 2k
EXAONE-Deep LGAI-EXAONE Text gen 7.8B Licence fee 2× RTX 3060 12 GB 2025-03-12 3k
gemma-3 google Text gen 1.0B Conditions RTX 3060 12 GB 2025-03-10 3.5M
QwQ Qwen Text gen 32.8B Apache 2.0 H100 80 GB 2025-03-05 75k
shieldgemma-2 google Vision + text 4.3B Conditions RTX 3060 12 GB 2025-03-04 9k
aya-vision CohereLabs Vision + text 8.6B Licence fee RTX 4090 24 GB 2025-03-02 5k
Janus-Pro deepseek-community Omni (any→any) 2.1B MIT RTX 3060 12 GB 2025-03-01 26k
sarashina2.2 sbintuitions Text gen 793M MIT RTX 3060 12 GB 2025-02-26 220k
OLMo-2-0325 allenai Text gen 32.2B Apache 2.0 H100 80 GB 2025-02-23 8k
Moonlight moonshotai Text gen · MoE 16.0B MIT L40S 48 GB 2025-02-22 62k
SmolLM2-16k HuggingFaceTB Text gen 1.7B Apache 2.0 RTX 3060 12 GB 2025-02-21 151
gemma-3-pt google Vision + text 4.3B Conditions RTX 3060 12 GB 2025-02-20 84k
LLaDA GSAI-ML Text gen 8.0B MIT RTX 4090 24 GB 2025-02-19 250k
Phi-4-mini microsoft Text gen 3.8B MIT RTX 3060 12 GB 2025-02-19 457k
granite-3.2 ibm-granite Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2025-02-17 11k
ALLaM humain-ai Text gen 7.0B Apache 2.0 2× RTX 3060 12 GB 2025-02-13 9k
SmolVLM2-Video HuggingFaceTB Vision + text 507M Apache 2.0 RTX 3060 12 GB 2025-02-11 1.5M
Ovis2 ATH-MaaS Vision + text 1.3B Apache 2.0 RTX 3060 12 GB 2025-02-10 15k
SmolVLM2 HuggingFaceTB Vision + text 2.2B Apache 2.0 RTX 3060 12 GB 2025-02-08 169k
Meta-Llama-3.1-unsloth-bnb unsloth Text gen 8.2B Conditions RTX 3060 12 GB 2025-02-02 36k
Mistral-Small-2501 mistralai Text gen 23.6B Apache 2.0 2× RTX 5090 32 GB 2025-01-30 45k
DeepScaleR agentica-org Text gen 1.8B MIT RTX 3060 12 GB 2025-01-29 7k
OLMoE-0125 allenai Text gen · MoE 6.9B Apache 2.0 RTX 4060 Ti 16 GB 2025-01-27 302k
llm-jp-3 llm-jp Text gen 152M Apache 2.0 RTX 3060 12 GB 2025-01-27 124k
YuE-s1-anneal-en-cot m-a-p Text gen 6.2B Apache 2.0 RTX 4060 Ti 16 GB 2025-01-26 6k
Janus-Pro deepseek-ai Omni (any→any) — MIT RTX 3060 12 GB 2025-01-26 13k
Qwen2.5-VL Qwen Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-01-26 8.3M
KwaiCoder Kwaipilot Text gen · MoE 23.3B MIT 2× RTX 5090 32 GB 2025-01-22 57
UI-TARS-DPO ByteDance-Seed Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-01-22 4k
DeepSeek-R1-Qwen deepseek-ai Text gen 32.8B MIT H100 80 GB 2025-01-20 595k
DeepSeek-R1-Llama deepseek-ai Text gen 8.0B MIT 2× RTX 3060 12 GB 2025-01-20 391k
DeepSeek-R1 deepseek-ai Text gen 684.5B MIT 8× H200 141 GB 2025-01-20 2.2M
DeepSeek-R1-Zero deepseek-ai Text gen 684.5B MIT 8× H200 141 GB 2025-01-20 6k
UI-TARS-SFT ByteDance-Seed Vision + text 2.4B Apache 2.0 RTX 3060 12 GB 2025-01-20 5k
SmolVLM HuggingFaceTB Vision + text 256M Apache 2.0 RTX 3060 12 GB 2025-01-17 596k
olmOCR-0225 allenai Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-01-15 5k
helium-1 kyutai Text gen 2.2B Open RTX 3060 12 GB 2025-01-13 11k
internlm3 internlm Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2025-01-13 85k
MiniMax-VL-01 MiniMaxAI Vision + text 456.4B — 8× H200 141 GB 2025-01-12 18k
MiniCPM-o-2_6 openbmb Omni (any→any) 8.7B Apache 2.0 2× RTX 3060 12 GB 2025-01-12 186k
starvector-im2svg starvector Text gen 7.5B Apache 2.0 2× RTX 3060 12 GB 2025-01-11 3k
huginn-0125 tomg-group-umd Text gen 3.9B Apache 2.0 RTX 3060 12 GB 2025-01-08 29k
AIN MBZUAI Vision + text 8.3B MIT 2× RTX 3060 12 GB 2025-01-07 939
Sa2VA ByteDance Vision + text 8.3B Apache 2.0 2× RTX 3060 12 GB 2025-01-07 6k
llava-llama-3-v1_1-imat city96 Vision + text — — RTX 3060 12 GB 2024-12-20 875
OLMo-2-1124-SFT allenai Text gen — Apache 2.0 2× RTX 3060 12 GB 2024-12-18 33k
OLMo-2-1124 allenai Text gen 7.3B Apache 2.0 2× RTX 3060 12 GB 2024-12-18 65k
llama3.1-typhoon2 typhoon-ai Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-12-15 64k
deepseek-vl2 deepseek-ai Vision + text 27.5B Licence fee H100 80 GB 2024-12-13 4k
deepseek-vl2-small deepseek-ai Vision + text 16.1B Licence fee L40S 48 GB 2024-12-13 8k
deepseek-vl2-tiny deepseek-ai Vision + text 3.4B Licence fee RTX 3060 12 GB 2024-12-13 302k
gemma-2 Efficient-Large-Model Text gen 2.6B Conditions RTX 4060 Ti 16 GB 2024-12-12 149k
c4ai-command-r7b-12-2024 CohereLabs Text gen 8.0B Licence fee RTX 4090 24 GB 2024-12-11 13k
phi-4 microsoft Text gen 14.7B MIT L40S 48 GB 2024-12-11 640k
VisionReward-Video zai-org Text gen 12.5B Licence fee 2× RTX 4060 Ti 16 GB 2024-12-10 3k
DeepSeek-1210 deepseek-ai Text gen 235.7B Licence fee 4× H200 141 GB 2024-12-10 770
Mini-InternVL2-DA-Medical OpenGVLab Vision + text 4.1B MIT RTX 4060 Ti 16 GB 2024-12-07 7k
granite-3.1 ibm-granite Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2024-12-06 47k
gregg-vision.1 grascii Image→text 36M MIT RTX 3060 12 GB 2024-12-06 1k
gemma-2-MoAA-DPO togethercomputer Text gen 9.2B — 2× RTX 3060 12 GB 2024-12-06 53
Llama-3.1-MoAA-DPO togethercomputer Text gen 8.0B — 2× RTX 3060 12 GB 2024-12-05 52
Hermes-3-Llama-3.2 NousResearch Text gen 3.2B Conditions RTX 3060 12 GB 2024-12-03 5k
KwaiCoder-DS-Lite Kwaipilot Text gen 15.7B MIT L40S 48 GB 2024-12-02 72
EXAONE-3.5 LGAI-EXAONE Text gen 7.8B Licence fee RTX 3060 12 GB 2024-12-01 399k
Aria-8K rhymes-ai Vision + text 25.3B Apache 2.0 H100 80 GB 2024-12-01 30
Aria-64K rhymes-ai Vision + text 25.3B Apache 2.0 H100 80 GB 2024-11-30 31
Falcon3 tiiuae Text gen 7.5B Licence fee 2× RTX 3060 12 GB 2024-11-29 18k
glm-edge zai-org Text gen — Licence fee RTX 3060 12 GB 2024-11-27 35k
glm-edge-v zai-org Vision + text — Licence fee RTX 3060 12 GB 2024-11-27 837
Llama-3.3 meta-llama Text gen 70.6B Conditions 4× L40S 48 GB 2024-11-26 339k
GOT-OCR-2.0 stepfun-ai Vision + text 561M Apache 2.0 RTX 3060 12 GB 2024-11-22 178k
EuroLLM utter-project Text gen 9.2B Apache 2.0 RTX 4090 24 GB 2024-11-22 35k
paligemma2-pt-224 google Vision + text 3.0B Conditions RTX 3060 12 GB 2024-11-21 9k
paligemma2-mix-448 google Vision + text 3.0B Conditions RTX 3060 12 GB 2024-11-21 3k
paligemma2-ft-docci-448 google Vision + text 3.0B Conditions RTX 3060 12 GB 2024-11-21 13k
paligemma2-mix-224 google Vision + text 3.0B Conditions RTX 3060 12 GB 2024-11-21 23k
InternVL2_5 OpenGVLab Vision + text 3.7B MIT RTX 3060 12 GB 2024-11-20 260k
Llama-3.1-Tulu-3 allenai Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-11-20 10k
Llama-3.1-Tulu-3-DPO allenai Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-11-20 14k
Llama-3.1-Tulu-3-SFT allenai Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-11-18 19k
Athene-Agent Nexusflow Text gen 72.7B Licence fee 4× L40S 48 GB 2024-11-12 58
Athene Nexusflow Text gen 72.7B Licence fee 4× L40S 48 GB 2024-11-12 3k
JanusFlow deepseek-ai Omni (any→any) 2.0B MIT RTX 3060 12 GB 2024-11-12 1k
Llama3.1-PRM-Deepseek-Data RLHFlow Text gen 8.0B — 2× RTX 3060 12 GB 2024-11-08 16k
Hymba nvidia Text gen 1.5B Licence fee RTX 3060 12 GB 2024-10-31 1k
SmolLM2 HuggingFaceTB Text gen 135M Apache 2.0 RTX 3060 12 GB 2024-10-31 2.4M
Emu3 BAAI Vision + text 8.8B Apache 2.0 RTX 4090 24 GB 2024-10-24 42k
Llama-3.2-SpinQuant_INT4_EO8 meta-llama Text gen — Conditions RTX 3060 12 GB 2024-10-23 212
Llama-3.2-QLORA_INT4_EO8 meta-llama Text gen — Conditions RTX 3060 12 GB 2024-10-23 216
sarvam-1 sarvamai Text gen 2.5B — RTX 3060 12 GB 2024-10-23 8k
aya-expanse CohereLabs Text gen 8.0B Licence fee RTX 4090 24 GB 2024-10-23 15k
StructTable-InternVL2 InternScience Image→text 938M Apache 2.0 RTX 3060 12 GB 2024-10-18 603
Janus deepseek-ai Omni (any→any) 2.1B MIT RTX 3060 12 GB 2024-10-18 4k
Aria-sequential_mlp rhymes-ai Vision + text 25.3B Apache 2.0 H100 80 GB 2024-10-17 24
h2ovl-mississippi h2oai Text gen 826M Apache 2.0 RTX 3060 12 GB 2024-10-16 47k
LLaMmlein_1B_prerelease LSX-UniWue Text gen 1.1B Licence fee RTX 3060 12 GB 2024-10-15 184k
falcon-mamba-tiny-dev tiiuae Text gen 9M — RTX 3060 12 GB 2024-10-13 68k
Llama-3.1-Dragonfly-Med togethercomputer Vision + text — Conditions 2× RTX 3060 12 GB 2024-10-10 56
Llama-3.1-Dragonfly togethercomputer Vision + text — Conditions 2× RTX 3060 12 GB 2024-10-10 56
granite-3.0 ibm-granite Text gen 8.2B Apache 2.0 2× RTX 3060 12 GB 2024-10-02 154k
Mistral-NeMo-Minitron nvidia Text gen 8.4B Licence fee 2× RTX 3060 12 GB 2024-10-02 44k
NVLM-D nvidia Vision + text 79.4B Licence fee 2× H200 141 GB 2024-09-30 11k
openthaigpt1.5 openthaigpt Text gen 7.6B Licence fee 2× RTX 3060 12 GB 2024-09-30 83k
salamandra BSC-LT Text gen 7.8B Apache 2.0 2× RTX 3060 12 GB 2024-09-30 27k
thai-trocr openthaigpt Image→text 103M Apache 2.0 RTX 3060 12 GB 2024-09-29 638
hf-moshiko kmhf Text gen 7.8B — 2× RTX 5090 32 GB 2024-09-27 107k
Llama-3.2 NousResearch Text gen 1.2B Conditions RTX 3060 12 GB 2024-09-27 45k
Aria rhymes-ai Vision + text 25.3B Apache 2.0 H100 80 GB 2024-09-26 64k
gemma-2-jpn google Text gen 2.6B Conditions RTX 3060 12 GB 2024-09-25 6k
Emu3-Gen BAAI Omni (any→any) 8.5B Apache 2.0 2× RTX 3060 12 GB 2024-09-25 989
Molmo-0924 allenai Vision + text 73.3B Apache 2.0 4× L40S 48 GB 2024-09-25 10k
Molmo-D-0924 allenai Vision + text 8.0B Apache 2.0 2× RTX 3060 12 GB 2024-09-25 22k
gemma-2-MoAA-SFT togethercomputer Text gen 9.2B — 2× RTX 3060 12 GB 2024-09-23 53
Llama-3.1-MoAA-SFT togethercomputer Text gen 8.0B — 2× RTX 3060 12 GB 2024-09-20 55
Zamba2 Zyphra Text gen 1.2B Apache 2.0 RTX 3060 12 GB 2024-09-19 179k
Llama-3.2-Vision meta-llama Vision + text 10.7B Conditions 2× RTX 4060 Ti 16 GB 2024-09-18 110k
Llama-3.2 meta-llama Text gen 1.2B Conditions RTX 3060 12 GB 2024-09-18 6.7M
Qwen2.5-Coder Qwen Text gen 7.6B Apache 2.0 2× RTX 3060 12 GB 2024-09-17 2.4M
Qwen2.5-Math Qwen Text gen 1.5B Apache 2.0 RTX 3060 12 GB 2024-09-16 370k
Qwen2.5 Qwen Text gen 7.6B Apache 2.0 2× RTX 3060 12 GB 2024-09-16 10.8M
pixtral mistral-experimental Vision + text 12.7B Apache 2.0 RTX 5090 32 GB 2024-09-13 137k
GOT-OCR2_0 stepfun-ai Vision + text 716M Apache 2.0 RTX 3060 12 GB 2024-09-12 707k
Nemotron-Mini nvidia Text gen — Licence fee RTX 3060 12 GB 2024-09-10 72k
solar-pro upstage Text gen 22.1B MIT 2× RTX 5090 32 GB 2024-09-09 11k
DeepSeek-Coder-0724 deepseek-ai Text gen 235.7B Licence fee 4× H200 141 GB 2024-09-05 598
openvla-finetuned-libero-spatial openvla Vision + text 7.5B MIT RTX 4090 24 GB 2024-09-03 13k
MiniCPM3 openbmb Text gen — Apache 2.0 RTX 3060 12 GB 2024-09-03 27k
solar-pro-pretrained upstage Text gen 22.1B — H100 80 GB 2024-08-29 0
Qwen2-VL Qwen Vision + text 8.3B Apache 2.0 RTX 3060 12 GB 2024-08-29 1.7M
xLAM-8x22b-r Salesforce Text gen 140.6B Licence fee 4× H100 80 GB 2024-08-28 17k
Yi-Coder 01-ai Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2024-08-21 9k
Phi-3.5-MoE microsoft Text gen · MoE 41.9B MIT 4× RTX 4090 24 GB 2024-08-17 132k
SILMA silma-ai Text gen 9.2B Conditions 2× RTX 3060 12 GB 2024-08-17 1k
Phi-3.5-vision microsoft Vision + text 4.1B MIT RTX 4060 Ti 16 GB 2024-08-16 697k
Phi-3.5-mini microsoft Text gen 3.8B MIT RTX 3060 12 GB 2024-08-16 290k
PowerMoE ibm-research Text gen · MoE 3.4B Apache 2.0 RTX 3060 12 GB 2024-08-14 1.2M
PowerLM ibm-research Text gen 3.5B Apache 2.0 RTX 3060 12 GB 2024-08-14 52k
llava-onevision-qwen2-ov llava-hf Vision + text 894M Apache 2.0 RTX 3060 12 GB 2024-08-13 583k
CheXagent-2-srrg-impression StanfordAIMI Image→text 3.1B MIT RTX 3060 12 GB 2024-08-12 1k
CheXagent-2-srrg-findings StanfordAIMI Image→text 3.1B MIT RTX 3060 12 GB 2024-08-12 1k
xflux_text_encoders XLabs-AI Text gen 4.8B Apache 2.0 RTX 4060 Ti 16 GB 2024-08-11 325k
Idefics3-Llama3 HuggingFaceM4 Vision + text 8.5B Apache 2.0 2× RTX 3060 12 GB 2024-08-05 92k
Lumina-mGPT-512 Alpha-VLLM Omni (any→any) 7.0B — 2× RTX 3060 12 GB 2024-08-04 482
Lumina-mGPT-1024 Alpha-VLLM Omni (any→any) 7.0B — 2× RTX 3060 12 GB 2024-08-04 14
MiniCPM-V-2_6 openbmb Vision + text 8.1B — 2× RTX 3060 12 GB 2024-08-04 46k
EXAONE-3.0 LGAI-EXAONE Text gen 7.8B Licence fee RTX 4090 24 GB 2024-07-31 12k
gemma-2-bnb unsloth Text gen 2.7B Conditions RTX 3060 12 GB 2024-07-31 41k
internlm2_5-1 internlm Text gen 1.9B Licence fee RTX 3060 12 GB 2024-07-30 6k
Lumina-mGPT-512-MultiImage Alpha-VLLM Omni (any→any) 7.0B — 2× RTX 3060 12 GB 2024-07-29 14
maira-2 microsoft Text gen 6.9B Licence fee RTX 4060 Ti 16 GB 2024-07-29 3k
Lumina-mGPT-768 Alpha-VLLM Omni (any→any) 7.0B — 2× RTX 3060 12 GB 2024-07-29 2k
Chameleon_7B_mGPT Alpha-VLLM Omni (any→any) 7.0B — RTX 4060 Ti 16 GB 2024-07-28 0
Lumina-mGPT-768-Omni Alpha-VLLM Omni (any→any) 7.0B — 2× RTX 3060 12 GB 2024-07-28 13
Hermes-3-Llama-3.1 NousResearch Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-07-28 412k
Meta-Llama-3.1-quantized RedHatAI Text gen 8.0B Conditions RTX 3060 12 GB 2024-07-26 65k
EleutherAI_pythia-deduped__sft__tldr HuggingFaceH4 Text gen 6.9B — 2× RTX 3060 12 GB 2024-07-25 23
Meta-Llama-3.1 NousResearch Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-07-24 288k
Llama-3.1-AlternateTokenizer teknium Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-07-23 1k
Meta-Llama-3.1 meta-llama Text gen 70.6B Conditions A100 80 GB 2024-07-23 142k
Meta-Llama-3.1 unsloth Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-07-23 363k
Meta-Llama-3.1-bnb unsloth Text gen 8.2B Conditions RTX 3060 12 GB 2024-07-23 101k
Llama-Guard-3 meta-llama Text gen 8.0B Conditions RTX 4090 24 GB 2024-07-22 201k
OLMoE-0924 allenai Text gen · MoE 6.9B Apache 2.0 RTX 4060 Ti 16 GB 2024-07-20 155k
llama3-llava-next llava-hf Vision + text 8.4B Conditions RTX 4090 24 GB 2024-07-19 22k
llama3-CLA-3 answerdotai Text gen — — 2× RTX 3060 12 GB 2024-07-18 11
llama3-CLA-2 answerdotai Text gen — — 2× RTX 3060 12 GB 2024-07-18 9
Llama-3.1 meta-llama Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-07-18 5.9M
DeepSeek-0628 deepseek-ai Text gen 235.7B Licence fee 4× H200 141 GB 2024-07-18 4k
xLAM-fc-r Salesforce Text gen 1.3B Licence fee RTX 3060 12 GB 2024-07-17 3k
DeepSeek-Coder-Lite RedHatAI Text gen 15.7B Licence fee 2× RTX 3060 12 GB 2024-07-17 129k
falcon-mamba tiiuae Text gen 7.3B Licence fee RTX 4060 Ti 16 GB 2024-07-17 65k
shieldgemma google Text gen 2.6B Conditions RTX 3060 12 GB 2024-07-16 3k
gemma-2 google Text gen 2.6B Conditions RTX 3060 12 GB 2024-07-16 758k
NuminaMath-CoT AI-MO Text gen 6.9B Apache 2.0 2× RTX 4060 Ti 16 GB 2024-07-15 86
SmolLM HuggingFaceTB Text gen 135M Apache 2.0 RTX 3060 12 GB 2024-07-14 150k
OLMo-0724 allenai Text gen 6.9B Apache 2.0 2× RTX 3060 12 GB 2024-07-12 7k
llava-interleave-qwen llava-hf Vision + text 864M Licence fee RTX 3060 12 GB 2024-07-10 38k
Infinity-0625-Llama3 BAAI Text gen 8.0B Apache 2.0 2× RTX 3060 12 GB 2024-07-09 8k
Infinity-0625-Yi-1.5 BAAI Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2024-07-09 8k
codegeex4-all zai-org Text gen 9.4B Licence fee 2× RTX 3060 12 GB 2024-07-05 10k
h2o-danube3 h2oai Text gen 514M Apache 2.0 RTX 3060 12 GB 2024-07-04 75k
NuminaMath-TIR AI-MO Text gen 6.9B Apache 2.0 2× RTX 3060 12 GB 2024-07-04 406
Llama3-Med42 m42-health Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-07-02 4k
llava-onevision-qwen2-ov lmms-lab Text gen 8.0B Apache 2.0 2× RTX 3060 12 GB 2024-06-29 35k
gemma-2 unsloth Text gen 27.2B Conditions H100 80 GB 2024-06-27 43k
internlm2_5 internlm Text gen 7.7B Licence fee 2× RTX 3060 12 GB 2024-06-27 38k
InternVL2 OpenGVLab Vision + text 2.2B MIT RTX 3060 12 GB 2024-06-27 649k
AceGPT FreedomIntelligence Text gen 8.0B Apache 2.0 2× RTX 3060 12 GB 2024-06-20 1k
Llama-3-SEC arcee-ai Text gen — Conditions L40S 48 GB 2024-06-18 3k
wildguard allenai Text gen 7.2B Apache 2.0 RTX 4060 Ti 16 GB 2024-06-15 150k
Florence-2-ft microsoft Vision + text 232M MIT RTX 3060 12 GB 2024-06-15 241k
Florence-2-large-ft microsoft Vision + text 770M MIT RTX 3060 12 GB 2024-06-15 38k
Florence-2 microsoft Vision + text 232M MIT RTX 3060 12 GB 2024-06-15 2.8M
Florence-2-large microsoft Vision + text 777M MIT RTX 3060 12 GB 2024-06-15 626k
DeepSeek-Coder-Lite deepseek-ai Text gen 15.7B Licence fee L40S 48 GB 2024-06-14 690k
4M-21_B EPFL-VILAB Omni (any→any) 843M Licence fee RTX 3060 12 GB 2024-06-12 916
4M-21_L EPFL-VILAB Omni (any→any) 1.6B Licence fee RTX 3060 12 GB 2024-06-12 510
crnn-fa-license-plate-recognition hezarai Image→text — — RTX 3060 12 GB 2024-06-09 3k
glm-4 zai-org Text gen 9.4B Licence fee 2× RTX 3060 12 GB 2024-06-04 11k
Llama-3-Dragonfly-Med togethercomputer Vision + text — Conditions RTX 4090 24 GB 2024-06-03 2
Llama-3-Dragonfly togethercomputer Vision + text — Conditions RTX 4090 24 GB 2024-06-03 2
Qwen2 Qwen Text gen 494M Apache 2.0 RTX 3060 12 GB 2024-05-31 796k
TrOCR_german_handwritten fhswf Image→text 558M Licence fee RTX 3060 12 GB 2024-05-29 967
aya-23 CohereLabs Text gen 8.0B Licence fee RTX 4090 24 GB 2024-05-19 9k
Phi-3-vision-128k microsoft Text gen 4.1B MIT RTX 4060 Ti 16 GB 2024-05-19 65k
MiniCPM-Llama3-V-2_5 openbmb Vision + text 8.5B — 2× RTX 3060 12 GB 2024-05-19 16k
cogvlm2-llama3 zai-org Text gen 19.5B Licence fee H100 80 GB 2024-05-16 7k
Yi-1.5-16K 01-ai Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2024-05-15 9k
Yi-1.5-32K 01-ai Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2024-05-15 8k
DeepSeek-Lite deepseek-ai Text gen 15.7B Licence fee L40S 48 GB 2024-05-15 448k
llava-med-mistral microsoft Vision + text 7.6B Apache 2.0 2× RTX 3060 12 GB 2024-05-14 13k
Mini-InternVL-5 OpenGVLab Vision + text 2.2B MIT RTX 3060 12 GB 2024-05-13 6k
deepseekcoder-codeqwen-align-subset bigcode Text gen 33.3B — H100 80 GB 2024-05-13 144
kosmos-2.5 microsoft Vision + text 1.4B MIT RTX 3060 12 GB 2024-05-13 18k
paligemma-ft-cococap-448 google Vision + text 2.9B Conditions RTX 3060 12 GB 2024-05-13 253k
paligemma-mix-224 google Vision + text 2.9B Conditions RTX 3060 12 GB 2024-05-12 103k
paligemma-pt-224 google Vision + text 2.9B Conditions RTX 3060 12 GB 2024-05-12 204k
Yi-1.5 01-ai Text gen 8.8B Apache 2.0 2× RTX 3060 12 GB 2024-05-10 16k
openchat-3.6-20240522 openchat Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-05-07 10k
Hermes-2-Theta-Llama-3 NousResearch Text gen 8.0B Apache 2.0 2× RTX 3060 12 GB 2024-05-05 10k
Llama-3-VILA1.5 Efficient-Large-Model Text gen — Licence fee RTX 4090 24 GB 2024-04-30 3k
Hermes-2-Pro-Llama-3 NousResearch Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-04-30 4k
Llama-3-Gradient-1048k gradientai Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-04-29 36k
VILA1.5 Efficient-Large-Model Text gen — Licence fee RTX 3060 12 GB 2024-04-27 22k
Llama-3-UltraMedical TsinghuaC3I Text gen 8.0B Conditions 2× RTX 3060 12 GB 2024-04-27 106k
llava-llama-3-v1_1-transformers xtuner Vision + text 8.4B — RTX 4090 24 GB 2024-04-26 16k
tiny-mixtral TitanML Text gen 247M — RTX 3060 12 GB 2024-04-24 174k
granite-code-2k ibm-granite Text gen 3.5B Apache 2.0 RTX 3060 12 GB 2024-04-23 39k
Phi-3-mini-128k microsoft Text gen 3.8B MIT RTX 3060 12 GB 2024-04-22 295k
Phi-3-mini-4k microsoft Text gen 3.8B MIT RTX 3060 12 GB 2024-04-22 628k
snowflake-arctic Snowflake Text gen 478.6B Apache 2.0 8× H200 141 GB 2024-04-21 6k
Minerva sapienzanlp Text gen 2.9B Apache 2.0 RTX 4060 Ti 16 GB 2024-04-19 3k
Meta-Llama-3 NousResearch Text gen 8.0B Licence fee 2× RTX 3060 12 GB 2024-04-18 99k
llama-3-bnb meta-llama Text gen 8.2B Conditions RTX 3060 12 GB 2024-04-18 61k
InternVL-5 OpenGVLab Vision + text 25.5B MIT H100 80 GB 2024-04-18 10k
K2 IFM Text gen 65.3B Apache 2.0 4× L40S 48 GB 2024-04-17 599
Meta-Llama-3 meta-llama Text gen 8.0B Conditions RTX 4090 24 GB 2024-04-17 1.7M
TinySolar-4k-code upstage Text gen 248M Apache 2.0 RTX 3060 12 GB 2024-04-15 119
OpenELM apple Text gen 3.0B Licence fee RTX 3060 12 GB 2024-04-12 610
OpenELM-1 apple Text gen 1.1B Licence fee RTX 3060 12 GB 2024-04-12 1.4M
OLMo allenai Text gen 1.2B Apache 2.0 RTX 3060 12 GB 2024-04-12 50k
vsft-llava-1.5-trl HuggingFaceH4 Image→text 7.1B — RTX 4060 Ti 16 GB 2024-04-11 45
zephyr-orpo HuggingFaceH4 Text gen · MoE 140.6B Apache 2.0 4× H100 80 GB 2024-04-10 151
idefics2 HuggingFaceM4 Vision + text 8.4B Apache 2.0 RTX 4090 24 GB 2024-04-09 114k
MiniCPM-128k openbmb Text gen — — RTX 3060 12 GB 2024-04-09 503
MiniCPM-MoE-8x2B openbmb Text gen · MoE — — 2× RTX 4060 Ti 16 GB 2024-04-07 4k
c4ai-command-r-plus CohereLabs Text gen 103.8B Licence fee more than the reference cards 2024-04-03 10k
transformer fla-hub Text gen 1.4B MIT RTX 3060 12 GB 2024-04-01 101k
uform-gen2-dpo unum-cloud Image→text 1.3B Apache 2.0 RTX 3060 12 GB 2024-03-27 4k
gemma-1.1 google Text gen 2.5B Conditions RTX 3060 12 GB 2024-03-26 93k
chameleon facebook Vision + text 7.0B Licence fee RTX 4060 Ti 16 GB 2024-03-26 28k
Starling-LM-beta Nexusflow Text gen 7.2B Apache 2.0 2× RTX 3060 12 GB 2024-03-19 1k
grok-1 hpcai-tech Text gen — Apache 2.0 4× B200 180 GB 2024-03-19 810
llava llava-hf Vision + text 34.8B — H100 80 GB 2024-03-17 8k
llava-vicuna llava-hf Vision + text 7.1B Conditions RTX 4060 Ti 16 GB 2024-03-17 26k
grok-1 xai-org Text gen — Apache 2.0 — 2024-03-17 929
CodeLlama meta-llama Text gen 6.7B Conditions RTX 4060 Ti 16 GB 2024-03-13 3k
c4ai-command-r CohereLabs Text gen 35.0B Licence fee H100 80 GB 2024-03-11 28k
Hermes-2-Pro-Mistral NousResearch Text gen 7.2B Apache 2.0 2× RTX 3060 12 GB 2024-03-11 4k
openchat-3.5-0106-gemma openchat Text gen 8.5B Licence fee 2× RTX 3060 12 GB 2024-03-09 1k
deepseek-vl deepseek-ai Vision + text 7.3B Licence fee RTX 4090 24 GB 2024-03-07 9k
mamba state-spaces Text gen 129M — RTX 3060 12 GB 2024-03-06 322k
Qwen1.5-MoE Qwen Text gen · MoE 14.3B Licence fee L40S 48 GB 2024-02-29 716k
geochat MBZUAI Text gen — Apache 2.0 2× RTX 3060 12 GB 2024-02-26 3k
udop-large microsoft Vision + text 742M MIT RTX 3060 12 GB 2024-02-26 93k
TrOCR qualcomm Image→text — MIT — 2024-02-25 2k
evo-1-8k togethercomputer Text gen 6.5B Apache 2.0 2× RTX 3060 12 GB 2024-02-24 3k
gemma unsloth Text gen 2.5B Apache 2.0 RTX 3060 12 GB 2024-02-21 35k
evo-1-131k togethercomputer Text gen 6.5B Apache 2.0 2× RTX 3060 12 GB 2024-02-20 2k
Viking LumiOpen Text gen 33.1B Apache 2.0 H100 80 GB 2024-02-20 4k
llava-mistral llava-hf Vision + text 7.6B Apache 2.0 RTX 4090 24 GB 2024-02-20 525k
uform-gen2-qwen unum-cloud Image→text 1.3B Apache 2.0 RTX 3060 12 GB 2024-02-15 510
BioMistral BioMistral Text gen — Apache 2.0 RTX 4060 Ti 16 GB 2024-02-14 53k
InternVL-2 OpenGVLab Vision + text 40.1B MIT 4× RTX 4090 24 GB 2024-02-11 7k
gemma google Text gen 2.5B Conditions RTX 3060 12 GB 2024-02-08 134k
truthfulqa-info-judge-llama2 allenai Text gen — Apache 2.0 2× RTX 3060 12 GB 2024-02-07 24k
truthfulqa-truth-judge-llama2 allenai Text gen — Apache 2.0 2× RTX 3060 12 GB 2024-02-07 26k
TinySolar-4k upstage Text gen 248M Apache 2.0 RTX 3060 12 GB 2024-02-07 410
TinySolar-4k-py upstage Text gen 248M Apache 2.0 RTX 3060 12 GB 2024-02-07 114
sqlcoder-2 defog Text gen 6.7B Open RTX 3060 12 GB 2024-02-05 10k
deepseek-math deepseek-ai Text gen — Licence fee 2× RTX 3060 12 GB 2024-02-05 6k
pythia-seed9 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-02-04 6k
pythia-seed8 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-02-04 5k
pythia-seed7 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-02-04 4k
HarmBench-Llama-2-cls cais Text gen 13.0B MIT 2× RTX 4060 Ti 16 GB 2024-02-03 138k
tofu_ft_phi-1.5 locuslab Text gen 1.4B Apache 2.0 RTX 3060 12 GB 2024-01-31 36k
internlm2-1 internlm Text gen — Licence fee RTX 3060 12 GB 2024-01-30 5k
MiniCPM-sft openbmb Text gen — — RTX 3060 12 GB 2024-01-29 33k
deepseek-coder deepseek-ai Text gen 6.9B Licence fee 2× RTX 3060 12 GB 2024-01-25 822k
internlm-xcomposer2 internlm Text gen — Licence fee 2× RTX 3060 12 GB 2024-01-25 6k
pythia-data-seed3 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-24 8k
Qwen1.5 Qwen Text gen 7.7B Licence fee 2× RTX 3060 12 GB 2024-01-22 108k
pythia-data-seed2 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-22 12k
pythia-data-seed1 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-22 52k
stablelm-2-zephyr-1 stabilityai Text gen 1.6B Licence fee RTX 3060 12 GB 2024-01-19 11k
v0-llama2-100k delphi-suite Text gen — MIT RTX 3060 12 GB 2024-01-19 54k
stablelm-2-1 stabilityai Text gen 1.6B Licence fee RTX 3060 12 GB 2024-01-18 8k
pythia-weight-seed3 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-17 5k
pythia-weight-seed2 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-17 12k
pythia-weight-seed1 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-16 22k
pythia-seed6 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-14 15k
pythia-seed5 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2024-01-14 15k
internlm2-sft internlm Text gen 7.7B Licence fee 2× RTX 3060 12 GB 2024-01-11 3k
Nous-Hermes-2-Mixtral-8x7B-DPO NousResearch Text gen 46.7B Apache 2.0 DGX Spark (GB10) 128 GB unified 2024-01-11 21k
internlm2 internlm Text gen 7.7B Licence fee 2× RTX 3060 12 GB 2024-01-10 72k
deepseek-moe deepseek-ai Text gen · MoE 16.4B Licence fee L40S 48 GB 2024-01-09 32k
stable-code stabilityai Text gen 2.8B Licence fee RTX 3060 12 GB 2024-01-09 6k
openchat-3.5-0106 openchat Text gen 7.2B Apache 2.0 2× RTX 3060 12 GB 2024-01-07 10k
Nous-Hermes-2-SOLAR NousResearch Text gen 10.7B Apache 2.0 RTX 4090 24 GB 2024-01-01 9k
TinyLlama TinyLlama Text gen 1.1B Apache 2.0 RTX 3060 12 GB 2023-12-30 1.9M
TinyLlama-intermediate-step-1431k-3T TinyLlama Text gen 1.1B Apache 2.0 RTX 3060 12 GB 2023-12-28 85k
Nous-Hermes-2-Mixtral-8x7B-SFT NousResearch Text gen 46.7B Apache 2.0 DGX Spark (GB10) 128 GB unified 2023-12-26 4k
Yi-VL 01-ai Vision + text — Apache 2.0 RTX 4060 Ti 16 GB 2023-12-25 2k
Nous-Hermes-2-Yi NousResearch Text gen 34.4B Apache 2.0 H100 80 GB 2023-12-23 8k
WhiteRabbitNeo WhiteRabbitNeo Text gen — Conditions L40S 48 GB 2023-12-17 517
phi-2 microsoft Text gen 2.8B MIT RTX 3060 12 GB 2023-12-13 1.5M
OpenHathi-Hi sarvamai Text gen 6.9B Conditions RTX 3060 12 GB 2023-12-13 1k
SOLAR upstage Text gen 10.7B Licence fee RTX 4090 24 GB 2023-12-12 31k
openchat-3.5-1210 openchat Text gen 7.2B Apache 2.0 2× RTX 3060 12 GB 2023-12-12 1k
Mistral mistralai Text gen 7.2B Apache 2.0 2× RTX 3060 12 GB 2023-12-11 1.2M
vip-llava llava-hf Vision + text 7.1B — RTX 4060 Ti 16 GB 2023-12-10 20k
Amber IFM Text gen 6.7B Apache 2.0 RTX 4060 Ti 16 GB 2023-12-07 2k
LlamaGuard meta-llama Text gen 6.7B Conditions RTX 4060 Ti 16 GB 2023-12-05 3k
llava-1.5 llava-hf Vision + text 7.1B Conditions RTX 4060 Ti 16 GB 2023-12-05 2.1M
NexusRaven Nexusflow Text gen 13.0B Licence fee L40S 48 GB 2023-12-04 126
StripedHyena-Nous togethercomputer Text gen 7.6B Apache 2.0 2× RTX 3060 12 GB 2023-12-04 167
Qwen-Audio Qwen Text gen 8.4B — 2× RTX 3060 12 GB 2023-11-30 960
Qwen-1 Qwen Text gen 1.8B — RTX 3060 12 GB 2023-11-30 1k
starcoder2 bigcode Text gen 3.0B Licence fee RTX 3060 12 GB 2023-11-29 119k
deepseek-llm deepseek-ai Text gen — Licence fee 2× RTX 3060 12 GB 2023-11-29 37k
crnn-fa hezarai Image→text — Apache 2.0 RTX 3060 12 GB 2023-11-27 3k
Qwen Qwen Text gen 72.3B Licence fee 4× L40S 48 GB 2023-11-26 2.4M
GPT-Prompt-Expansion-Fooocus LykosAI Text gen — Licence fee RTX 3060 12 GB 2023-11-25 37k
Starling-LM-alpha berkeley-nest Text gen 7.2B Apache 2.0 2× RTX 3060 12 GB 2023-11-25 1k
Yi-4bits 01-ai Text gen 6.1B Apache 2.0 RTX 3060 12 GB 2023-11-22 399
Yi-8bits 01-ai Text gen 6.1B Apache 2.0 RTX 3060 12 GB 2023-11-22 366
stablelm-zephyr stabilityai Text gen 2.8B Licence fee RTX 3060 12 GB 2023-11-21 15k
StripedHyena-Hessian togethercomputer Text gen 7.6B Apache 2.0 2× RTX 3060 12 GB 2023-11-21 97
Orca-2 microsoft Text gen — Licence fee H100 80 GB 2023-11-14 2k
tulu-2 allenai Text gen — — 2× RTX 3060 12 GB 2023-11-13 8k
sabia maritaca-ai Text gen 6.7B — RTX 4060 Ti 16 GB 2023-11-08 909
pythia-seed4 EleutherAI Text gen — Apache 2.0 RTX 3060 12 GB 2023-11-08 15k
Yi-200K 01-ai Text gen 6.1B Apache 2.0 RTX 4060 Ti 16 GB 2023-11-06 16k
hyenadna-medium-450k-seqlen LongSafari Text gen 28M BSD RTX 3060 12 GB 2023-11-03 87k
Hermes-Trismegistus-Mistral teknium Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-11-02 33
Yi 01-ai Text gen 6.1B Apache 2.0 RTX 4060 Ti 16 GB 2023-11-01 32k
openchat_3.5 openchat Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-10-30 1k
OpenHermes-2.5-Mistral teknium Text gen 7.2B Apache 2.0 2× RTX 3060 12 GB 2023-10-29 4k
mistral-sft-beta HuggingFaceH4 Text gen — MIT 2× RTX 3060 12 GB 2023-10-26 3k
zephyr-beta HuggingFaceH4 Text gen 7.2B MIT 2× RTX 3060 12 GB 2023-10-26 78k
Poro LumiOpen Text gen 34.2B Apache 2.0 4× RTX 4090 24 GB 2023-10-19 526
SD-PrompTune teknium Text gen — MIT 2× RTX 3060 12 GB 2023-10-18 15
fuyu adept Vision + text 9.4B Licence fee 2× RTX 4060 Ti 16 GB 2023-10-17 21k
MistralLite amazon Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-10-16 5k
OpenHermes-2-Mistral teknium Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-10-12 947
zephyr-alpha HuggingFaceH4 Text gen 7.2B MIT 2× RTX 3060 12 GB 2023-10-09 3k
Mistral-Trismegistus teknium Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-10-07 67
CollectiveCognition-Mistral teknium Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-10-04 31
airoboros-mistral2.2 teknium Text gen — MIT 2× RTX 3060 12 GB 2023-10-03 24
kosmos-2-patch14-224 microsoft Image→text 1.7B MIT RTX 3060 12 GB 2023-10-02 189k
Mistral-OpenOrca Open-Orca Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-09-29 856
vit-roberta-fa-image-captioning-flickr30k hezarai Image→text — — RTX 3060 12 GB 2023-09-29 2k
stablelm-4e1t stabilityai Text gen 2.8B Open RTX 3060 12 GB 2023-09-29 47k
internlm-xcomposer internlm Text gen — Apache 2.0 2× RTX 3060 12 GB 2023-09-26 5k
nougat-small facebook Image→text 247M Open RTX 3060 12 GB 2023-09-21 24k
nougat facebook Image→text 349M Licence fee RTX 3060 12 GB 2023-09-21 87k
Colossal-LLaMA-2 hpcai-tech Text gen — Conditions 2× RTX 3060 12 GB 2023-09-18 128
OpenHermes teknium Text gen — MIT 2× RTX 3060 12 GB 2023-09-14 41
Phi-Hermes teknium Text gen 1.4B Licence fee RTX 3060 12 GB 2023-09-13 17
Puffin-Phi teknium Text gen — Licence fee RTX 3060 12 GB 2023-09-12 34
phi-1_5 microsoft Text gen 1.4B MIT RTX 3060 12 GB 2023-09-10 57k
openmoe hpcai-tech Text gen · MoE — Apache 2.0 2× RTX 5090 32 GB 2023-09-05 63
openchat.2_super openchat Text gen — Conditions 2× RTX 4060 Ti 16 GB 2023-09-04 148
Yarn-Llama-2-128k NousResearch Text gen — — 2× RTX 3060 12 GB 2023-08-31 3k
Baichuan2-4bits baichuan-inc Text gen — Licence fee 2× RTX 3060 12 GB 2023-08-30 215
Baichuan2 baichuan-inc Text gen — Licence fee 2× RTX 3060 12 GB 2023-08-29 29k
CodeLlama codellama Text gen 6.7B Conditions 2× RTX 3060 12 GB 2023-08-24 258k
SOLAR-0 upstage Text gen — — 2× A100 80 GB 2023-07-30 22k
vicuna lmsys Text gen — Conditions 2× RTX 3060 12 GB 2023-07-29 34k
LLaMA-2-32K togethercomputer Text gen — Conditions 2× RTX 3060 12 GB 2023-07-26 6k
Nous-Hermes-llama-2 NousResearch Text gen 6.7B MIT 2× RTX 3060 12 GB 2023-07-25 21k
Nous-Hermes-Llama2 NousResearch Text gen 13.0B MIT 2× RTX 4060 Ti 16 GB 2023-07-20 3k
Llama-2 NousResearch Text gen 6.7B — 2× RTX 3060 12 GB 2023-07-18 249k
openchat_v2_openorca openchat Text gen — Licence fee RTX 5090 32 GB 2023-07-14 0
Llama-2 meta-llama Text gen 6.7B Conditions RTX 4060 Ti 16 GB 2023-07-13 804k
openchat_v2_w openchat Text gen — Licence fee 2× RTX 4060 Ti 16 GB 2023-07-07 111
internlm internlm Text gen — — 2× RTX 3060 12 GB 2023-07-06 35k
trocr-small-korean team-lucid Image→text 55M Apache 2.0 RTX 3060 12 GB 2023-06-30 1k
opencoderplus openchat Text gen — — L40S 48 GB 2023-06-30 118
openchat_8192 openchat Text gen — — L40S 48 GB 2023-06-22 58
openchat openchat Text gen — Licence fee 2× RTX 4060 Ti 16 GB 2023-06-22 133
Baichuan baichuan-inc Text gen — — 2× RTX 3060 12 GB 2023-06-13 23k
open_llama openlm-research Text gen — Apache 2.0 RTX 4060 Ti 16 GB 2023-06-07 82k
instructblip-flan-t5-xl Salesforce Vision + text 4.0B MIT RTX 3060 12 GB 2023-05-28 48k
instructblip-vicuna Salesforce Vision + text 7.9B Licence fee RTX 4090 24 GB 2023-05-22 7k
tiny_starcoder_py bigcode Text gen 164M Licence fee RTX 3060 12 GB 2023-05-15 220k
RedPajama-INCITE togethercomputer Text gen — Apache 2.0 RTX 3060 12 GB 2023-05-04 3k
starcoderbase bigcode Text gen — Licence fee H100 80 GB 2023-05-03 3k
replit-code replit Text gen — Open RTX 3060 12 GB 2023-04-28 657
codegen2-16B_P Salesforce Text gen — Apache 2.0 H100 80 GB 2023-04-26 22k
falcon-rw tiiuae Text gen — Apache 2.0 RTX 3060 12 GB 2023-04-26 3k
falcon tiiuae Text gen 7.2B Apache 2.0 RTX 4060 Ti 16 GB 2023-04-24 606k
starcoder bigcode Text gen 15.8B Licence fee L40S 48 GB 2023-04-24 6k
stablelm-alpha stabilityai Text gen — Open 2× RTX 3060 12 GB 2023-04-17 3k
LaMini-GPT MBZUAI Text gen — Licence fee RTX 3060 12 GB 2023-04-14 9k
LaMini-Flan-T5 MBZUAI Text gen — Licence fee RTX 3060 12 GB 2023-04-10 4k
gpt_bigcode-santacoder bigcode Text gen 1.1B Open RTX 3060 12 GB 2023-04-06 39k
vicuna-delta lmsys Text gen — — 2× RTX 4060 Ti 16 GB 2023-04-03 522
pix2struct google Image→text 282M Apache 2.0 RTX 3060 12 GB 2023-03-13 19k
pix2struct-textcaps google Image→text 282M Apache 2.0 RTX 3060 12 GB 2023-03-01 2k
pythia-seed3 EleutherAI Text gen 213M Apache 2.0 RTX 3060 12 GB 2023-02-15 30k
pythia-seed2 EleutherAI Text gen 213M Apache 2.0 RTX 3060 12 GB 2023-02-15 37k
pythia-seed1 EleutherAI Text gen 213M Apache 2.0 RTX 3060 12 GB 2023-02-15 40k
pythia-deduped EleutherAI Text gen 96M Apache 2.0 RTX 3060 12 GB 2023-02-13 839k
pythia EleutherAI Text gen 213M Apache 2.0 RTX 3060 12 GB 2023-02-08 3.5M
blip2-flan-t5-xl-coco Salesforce Image→text 3.9B MIT RTX 3060 12 GB 2023-02-07 832
blip2-opt-coco Salesforce Image→text 3.9B MIT RTX 3060 12 GB 2023-02-07 3k
blip2-flan-t5-xl Salesforce Vision + text 3.9B MIT RTX 3060 12 GB 2023-02-06 132k
blip2-opt Salesforce Vision + text 3.7B MIT RTX 4060 Ti 16 GB 2023-02-06 447k
mscoco_finetuned_CoCa-ViT-L-14-laion2B-s13B-b90k laion Image→text — MIT RTX 3060 12 GB 2023-02-03 18k
bloom-7b1-petals bigscience Text gen 1.0B — RTX 3060 12 GB 2023-01-22 204
bloomz-petals bigscience Text gen — — L40S 48 GB 2023-01-16 37
cpm-ant openbmb Text gen — — H100 80 GB 2023-01-15 12k
git-large-textcaps microsoft Image→text — MIT RTX 3060 12 GB 2023-01-02 559
git-large-coco microsoft Image→text 394M MIT RTX 3060 12 GB 2023-01-02 2k
git-large microsoft Image→text — MIT RTX 3060 12 GB 2023-01-02 654
BioMedLM stanford-crfm Text gen — Licence fee RTX 4060 Ti 16 GB 2022-12-14 1k
blip-image-captioning-large Salesforce Image→text 470M BSD RTX 3060 12 GB 2022-12-13 476k
blip-image-captioning Salesforce Image→text — BSD RTX 3060 12 GB 2022-12-12 1.9M
git-coco microsoft Image→text — MIT RTX 3060 12 GB 2022-12-06 5k
git microsoft Image→text 177M MIT RTX 3060 12 GB 2022-12-06 32k
santacoder bigcode Text gen — Licence fee RTX 3060 12 GB 2022-12-02 7k
mgp-str alibaba-damo Image→text 148M — RTX 3060 12 GB 2022-11-23 104k
biogpt microsoft Text gen — MIT RTX 3060 12 GB 2022-11-20 101k
mt0-xxl-p3 bigscience Text gen 13.9B Apache 2.0 L40S 48 GB 2022-10-28 95
mt0-xxl-mt bigscience Text gen 13.9B Apache 2.0 L40S 48 GB 2022-10-27 55
mt0-xl bigscience Text gen 3.7B Apache 2.0 RTX 3060 12 GB 2022-10-27 480
mt0-large bigscience Text gen 1.2B Apache 2.0 RTX 3060 12 GB 2022-10-27 696
mt0-small bigscience Text gen 300M Apache 2.0 RTX 3060 12 GB 2022-10-27 6k
mt0 bigscience Text gen 582M Apache 2.0 RTX 3060 12 GB 2022-10-27 1k
mt0-xxl bigscience Text gen 13.9B Apache 2.0 L40S 48 GB 2022-10-19 241
bloomz-1b7 bigscience Text gen 1.7B Licence fee RTX 3060 12 GB 2022-10-08 1k
bloomz bigscience Text gen 559M Licence fee RTX 3060 12 GB 2022-10-08 961k
vlt5-keywords Voicelab Text gen 275M Open RTX 3060 12 GB 2022-09-27 279k
bloomz-7b1 bigscience Text gen 7.1B Licence fee RTX 4060 Ti 16 GB 2022-09-27 7k
polyglot-ko EleutherAI Text gen 1.4B Apache 2.0 RTX 3060 12 GB 2022-09-15 4k
trocr-large-str microsoft Image→text — — RTX 3060 12 GB 2022-09-08 1k
trocr-str microsoft Image→text — — RTX 3060 12 GB 2022-09-08 1k
japanese-gpt-neox-small rinna Text gen 204M MIT RTX 3060 12 GB 2022-08-31 557k
gpt-neox-japanese abeja Text gen — MIT RTX 3060 12 GB 2022-08-29 90k
donut-finetuned-rvlcdip naver-clova-ix Image→text — MIT RTX 3060 12 GB 2022-07-19 1k
donut naver-clova-ix Image→text — MIT RTX 3060 12 GB 2022-07-19 54k
donut-finetuned-cord naver-clova-ix Image→text — MIT RTX 3060 12 GB 2022-07-19 9k
bloom-7b1 bigscience Text gen 7.1B Licence fee 2× RTX 3060 12 GB 2022-05-19 10k
bloom-1b7 bigscience Text gen 1.7B Licence fee RTX 3060 12 GB 2022-05-19 51k
bloom-1b1 bigscience Text gen 1.1B Licence fee RTX 3060 12 GB 2022-05-19 10k
bloom bigscience Text gen 559M Licence fee RTX 3060 12 GB 2022-05-19 454k
opt facebook Text gen — Licence fee RTX 3060 12 GB 2022-05-11 11.4M
RITA_s lightonai Text gen 85M — RTX 3060 12 GB 2022-04-25 627
codegen-multi Salesforce Text gen — BSD RTX 3060 12 GB 2022-04-11 13k
codegen-mono Salesforce Text gen — BSD RTX 3060 12 GB 2022-04-11 126k
gpt-neox EleutherAI Text gen 20.7B Apache 2.0 2× RTX 5090 32 GB 2022-04-07 785k
trocr-small-handwritten microsoft Image→text — — RTX 3060 12 GB 2022-03-02 512k
trocr-printed microsoft Image→text 333M — RTX 3060 12 GB 2022-03-02 194k
trocr-handwritten microsoft Image→text 333M MIT RTX 3060 12 GB 2022-03-02 173k
DialoGPT-medium microsoft Text gen — MIT RTX 3060 12 GB 2022-03-02 168k
trocr-large-handwritten microsoft Image→text — — RTX 3060 12 GB 2022-03-02 132k
vit-gpt2-image-captioning nlpconnect Image→text — Apache 2.0 RTX 3060 12 GB 2022-03-02 102k
xglm facebook Text gen — MIT RTX 3060 12 GB 2022-03-02 102k
reformer-crime-and-punishment google Text gen — — RTX 3060 12 GB 2022-03-02 96k
trocr-stage1 microsoft Image→text 384M — RTX 3060 12 GB 2022-03-02 58k
trocr-large-printed microsoft Image→text 608M — RTX 3060 12 GB 2022-03-02 56k
DialoGPT-small microsoft Text gen 176M MIT RTX 3060 12 GB 2022-03-02 54k
alias-gpt2-small-x21 stanford-crfm Text gen — Apache 2.0 RTX 3060 12 GB 2022-03-02 36k
trocr-small-printed microsoft Image→text 61M — RTX 3060 12 GB 2022-03-02 35k
trocr-small-stage1 microsoft Image→text — — RTX 3060 12 GB 2022-03-02 12k
DialoGPT-large microsoft Text gen — MIT RTX 3060 12 GB 2022-03-02 2k
trocr-large-stage1 microsoft Image→text 608M — RTX 3060 12 GB 2022-03-02 1k
gpt2 openai-community Text gen 137M MIT RTX 3060 12 GB 2022-03-02 14.3M
distilgpt2 distilbert Text gen 88M Apache 2.0 RTX 3060 12 GB 2022-03-02 2.1M
gpt2-large openai-community Text gen 812M MIT RTX 3060 12 GB 2022-03-02 1.1M
gpt-neo EleutherAI Text gen 150M MIT RTX 3060 12 GB 2022-03-02 513k
gpt2-medium openai-community Text gen 380M MIT RTX 3060 12 GB 2022-03-02 352k
gpt-j EleutherAI Text gen — Apache 2.0 2× RTX 4060 Ti 16 GB 2022-03-02 259k
openai-gpt openai-community Text gen 120M MIT RTX 3060 12 GB 2022-03-02 202k
xlnet-cased xlnet Text gen — MIT RTX 3060 12 GB 2022-03-02 172k
gpt2-small-dutch GroNLP Text gen 129M — RTX 3060 12 GB 2022-03-02 165k
gpt2-xl openai-community Text gen 1.6B MIT RTX 3060 12 GB 2022-03-02 103k
ctrl Salesforce Text gen — BSD RTX 3060 12 GB 2022-03-02 97k
Mistral Small mistralai Text gen 24B dense Apache 2.0 RTX 3060 12 GB

Downloads are HuggingFace 30-day pulls of the most popular variant, a popularity signal, not an AxForge metric. "Runs on" is an estimate from each lead build's files and config; the model's page has the whole picture, build by build.

Explore

Other categories

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms