Best open source LLM rankings
150 open-weights models you can self-host or fine-tune, ranked on the same LLM Index Score as everything else — 7 of them carry a computed score. Weights are downloadable for every model listed here.
| # | Model | LLM Index Scoreours | Reasoning | Coding | Agentic Coding | Mathematics | Data Analysis | Language | IF | Price/M in | Context |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 1 | Z.ai: GLM 5.2 z-ai | 73.84 | 78.6 | 79.7 | 51.9 | 89.8 | 73.7 | 76.2 | 62.3 | $0.952 | 1.0M |
| 2 | DeepSeek: DeepSeek V4 Pro deepseek | 72.26 | 82.7 | 70.0 | 42.6 | 90.7 | 74.5 | 78.1 | 62.4 | $0.435 | 1.0M |
| 3 | MoonshotAI: Kimi K2.6 moonshotai | 71.06 | 79.4 | 78.6 | 46.9 | 84.3 | 65.1 | 75.1 | 64.4 | $0.684 | 262K |
| 4 | MoonshotAI: Kimi K2.7 Code moonshotai | 69.26 | 82.8 | 74.0 | 45.7 | 79.6 | 62.7 | 77.9 | 56.3 | $0.820 | 262K |
| 5 | MiniMax: MiniMax M3 minimax | 67.63 | 74.5 | 68.2 | 40.7 | 77.0 | 76.2 | 76.8 | 57.5 | $0.300 | 1.0M |
| 6 | DeepSeek: DeepSeek V4 Flash deepseek | 65.62 | 70.6 | 69.2 | 37.6 | 79.7 | 68.0 | 70.1 | 63.1 | $0.094 | 1.0M |
| 7 | Qwen: Qwen3.6 27B qwen | 64.91 | 70.3 | 71.8 | 39.3 | 79.9 | 70.4 | 63.3 | 53.2 | $0.450 | 262K |
| — | Meituan: LongCat 2.0 meituan | not evaluated | — | — | — | — | — | — | — | $0.300 | 1.0M |
| — | Thinking Machines: Inkling thinkingmachines | not evaluated | — | — | — | — | — | — | — | $1.00 | 1.0M |
| — | Tencent: Hy3 tencent | not evaluated | — | — | — | — | — | — | — | $0.140 | 262K |
| — | Poolside: Laguna XS 2.1 poolside | not evaluated | — | — | — | — | — | — | — | $0.060 | 262K |
| — | Poolside: Laguna XS 2.1 (free) poolside | not evaluated | — | — | — | — | — | — | — | Free | 262K |
| — | Nex AGI: Nex-N2-Mini nex-agi | not evaluated | — | — | — | — | — | — | — | $0.025 | 262K |
| — | Cohere: North Mini Code (free) cohere | not evaluated | — | — | — | — | — | — | — | Free | 256K |
| — | Nex AGI: Nex-N2-Pro nex-agi | not evaluated | — | — | — | — | — | — | — | $0.250 | 262K |
| — | NVIDIA: Nemotron 3.5 Content Safety (free) nvidia | not evaluated | — | — | — | — | — | — | — | Free | 128K |
| — | NVIDIA: Nemotron 3 Ultra nvidia | not evaluated | — | — | — | — | — | — | — | $0.600 | 1M |
| — | NVIDIA: Nemotron 3 Ultra (free) nvidia | not evaluated | — | — | — | — | — | — | — | Free | 1M |
| — | StepFun: Step 3.7 Flash stepfun | not evaluated | — | — | — | — | — | — | — | $0.200 | 256K |
| — | IBM: Granite 4.1 8B ibm-granite | not evaluated | — | — | — | — | — | — | — | $0.050 | 131K |
| — | NVIDIA: Nemotron 3 Nano Omni (free) nvidia | not evaluated | — | — | — | — | — | — | — | Free | 256K |
| — | Poolside: Laguna M.1 poolside | not evaluated | — | — | — | — | — | — | — | $0.200 | 262K |
| — | Poolside: Laguna M.1 (free) poolside | not evaluated | — | — | — | — | — | — | — | Free | 262K |
| — | Qwen: Qwen3.6 35B A3B qwen | not evaluated | — | — | — | — | — | — | — | $0.140 | 262K |
| — | Tencent: Hy3 preview tencent | not evaluated | — | — | — | — | — | — | — | $0.063 | 262K |
| — | Xiaomi: MiMo-V2.5-Pro xiaomi | not evaluated | — | — | — | — | — | — | — | $0.435 | 1.0M |
| — | Xiaomi: MiMo-V2.5 xiaomi | not evaluated | — | — | — | — | — | — | — | $0.140 | 1.0M |
| — | Z.ai: GLM 5.1 z-ai | not evaluated | — | — | — | — | — | — | — | $0.966 | 203K |
| — | Google: Gemma 4 26B A4B google | not evaluated | — | — | — | — | — | — | — | $0.070 | 262K |
| — | Google: Gemma 4 26B A4B (free) google | not evaluated | — | — | — | — | — | — | — | Free | 262K |
| — | Google: Gemma 4 31B google | not evaluated | — | — | — | — | — | — | — | $0.120 | 262K |
| — | Google: Gemma 4 31B (free) google | not evaluated | — | — | — | — | — | — | — | Free | 262K |
| — | Arcee AI: Trinity Large Thinking arcee-ai | not evaluated | — | — | — | — | — | — | — | $0.250 | 262K |
| — | Reka Edge rekaai | not evaluated | — | — | — | — | — | — | — | $0.100 | 16K |
| — | MiniMax: MiniMax M2.7 minimax | not evaluated | — | — | — | — | — | — | — | $0.250 | 205K |
| — | Mistral: Mistral Small 4 mistralai | not evaluated | — | — | — | — | — | — | — | $0.150 | 262K |
| — | NVIDIA: Nemotron 3 Super nvidia | not evaluated | — | — | — | — | — | — | — | $0.080 | 1M |
| — | NVIDIA: Nemotron 3 Super (free) nvidia | not evaluated | — | — | — | — | — | — | — | Free | 1M |
| — | Qwen: Qwen3.5-9B qwen | not evaluated | — | — | — | — | — | — | — | $0.100 | 262K |
| — | Qwen: Qwen3.5-35B-A3B qwen | not evaluated | — | — | — | — | — | — | — | $0.140 | 262K |
| — | Qwen: Qwen3.5-27B qwen | not evaluated | — | — | — | — | — | — | — | $0.260 | 262K |
| — | Qwen: Qwen3.5-122B-A10B qwen | not evaluated | — | — | — | — | — | — | — | $0.260 | 262K |
| — | Qwen: Qwen3.5 397B A17B qwen | not evaluated | — | — | — | — | — | — | — | $0.390 | 262K |
| — | MiniMax: MiniMax M2.5 minimax | not evaluated | — | — | — | — | — | — | — | $0.150 | 205K |
| — | Z.ai: GLM 5 z-ai | not evaluated | — | — | — | — | — | — | — | $0.950 | 205K |
| — | Qwen: Qwen3 Coder Next qwen | not evaluated | — | — | — | — | — | — | — | $0.110 | 262K |
| — | StepFun: Step 3.5 Flash stepfun | not evaluated | — | — | — | — | — | — | — | $0.100 | 262K |
| — | MoonshotAI: Kimi K2.5 moonshotai | not evaluated | — | — | — | — | — | — | — | $0.570 | 262K |
| — | Z.ai: GLM 4.7 Flash z-ai | not evaluated | — | — | — | — | — | — | — | $0.061 | 200K |
| — | MiniMax: MiniMax M2.1 minimax | not evaluated | — | — | — | — | — | — | — | $0.300 | 205K |
| — | Z.ai: GLM 4.7 z-ai | not evaluated | — | — | — | — | — | — | — | $0.400 | 203K |
| — | NVIDIA: Nemotron 3 Nano 30B A3B nvidia | not evaluated | — | — | — | — | — | — | — | $0.050 | 262K |
| — | NVIDIA: Nemotron 3 Nano 30B A3B (free) nvidia | not evaluated | — | — | — | — | — | — | — | Free | 256K |
| — | Mistral: Devstral 2 2512 mistralai | not evaluated | — | — | — | — | — | — | — | $0.400 | 262K |
| — | Z.ai: GLM 4.6V z-ai | not evaluated | — | — | — | — | — | — | — | $0.300 | 131K |
| — | Mistral: Ministral 3 14B 2512 mistralai | not evaluated | — | — | — | — | — | — | — | $0.200 | 262K |
| — | Mistral: Ministral 3 8B 2512 mistralai | not evaluated | — | — | — | — | — | — | — | $0.150 | 262K |
| — | Mistral: Ministral 3 3B 2512 mistralai | not evaluated | — | — | — | — | — | — | — | $0.100 | 131K |
| — | DeepSeek: DeepSeek V3.2 deepseek | not evaluated | — | — | — | — | — | — | — | $0.269 | 164K |
| — | AllenAI: Olmo 3 32B Think allenai | not evaluated | — | — | — | — | — | — | — | $0.150 | 66K |
| — | MoonshotAI: Kimi K2 Thinking moonshotai | not evaluated | — | — | — | — | — | — | — | $0.600 | 262K |
| — | Mistral: Voxtral Small 24B 2507 mistralai | not evaluated | — | — | — | — | — | — | — | $0.100 | 32K |
| — | OpenAI: gpt-oss-safeguard-20b openai | not evaluated | — | — | — | — | — | — | — | $0.075 | 131K |
| — | NVIDIA: Nemotron Nano 12B 2 VL (free) nvidia | not evaluated | — | — | — | — | — | — | — | Free | 128K |
| — | MiniMax: MiniMax M2 minimax | not evaluated | — | — | — | — | — | — | — | $0.300 | 205K |
| — | Qwen: Qwen3 VL 32B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.104 | 262K |
| — | IBM: Granite 4.0 Micro ibm-granite | not evaluated | — | — | — | — | — | — | — | $0.017 | 131K |
| — | Qwen: Qwen3 VL 8B Thinking qwen | not evaluated | — | — | — | — | — | — | — | $0.117 | 256K |
| — | Qwen: Qwen3 VL 8B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.117 | 256K |
| — | Qwen: Qwen3 VL 30B A3B Thinking qwen | not evaluated | — | — | — | — | — | — | — | $0.130 | 131K |
| — | Qwen: Qwen3 VL 30B A3B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.130 | 262K |
| — | Z.ai: GLM 4.6 z-ai | not evaluated | — | — | — | — | — | — | — | $0.500 | 203K |
| — | DeepSeek: DeepSeek V3.2 Exp deepseek | not evaluated | — | — | — | — | — | — | — | $0.270 | 164K |
| — | TheDrummer: Cydonia 24B V4.1 thedrummer | not evaluated | — | — | — | — | — | — | — | $0.300 | 131K |
| — | Qwen: Qwen3 VL 235B A22B Thinking qwen | not evaluated | — | — | — | — | — | — | — | $0.260 | 131K |
| — | Qwen: Qwen3 VL 235B A22B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.210 | 131K |
| — | DeepSeek: DeepSeek V3.1 Terminus deepseek | not evaluated | — | — | — | — | — | — | — | $0.270 | 131K |
| — | Qwen: Qwen3 Next 80B A3B Thinking qwen | not evaluated | — | — | — | — | — | — | — | $0.098 | 262K |
| — | Qwen: Qwen3 Next 80B A3B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.098 | 262K |
| — | NVIDIA: Nemotron Nano 9B V2 (free) nvidia | not evaluated | — | — | — | — | — | — | — | Free | 128K |
| — | MoonshotAI: Kimi K2 0905 moonshotai | not evaluated | — | — | — | — | — | — | — | $0.600 | 262K |
| — | Qwen: Qwen3 30B A3B Thinking 2507 qwen | not evaluated | — | — | — | — | — | — | — | $0.130 | 131K |
| — | Nous: Hermes 4 70B nousresearch | not evaluated | — | — | — | — | — | — | — | $0.130 | 131K |
| — | Nous: Hermes 4 405B nousresearch | not evaluated | — | — | — | — | — | — | — | $1.00 | 131K |
| — | DeepSeek: DeepSeek V3.1 deepseek | not evaluated | — | — | — | — | — | — | — | $0.250 | 164K |
| — | Z.ai: GLM 4.5V z-ai | not evaluated | — | — | — | — | — | — | — | $0.600 | 66K |
| — | AI21: Jamba Large 1.7 ai21 | not evaluated | — | — | — | — | — | — | — | $2.00 | 256K |
| — | OpenAI: gpt-oss-120b openai | not evaluated | — | — | — | — | — | — | — | $0.037 | 131K |
| — | OpenAI: gpt-oss-20b openai | not evaluated | — | — | — | — | — | — | — | $0.030 | 131K |
| — | OpenAI: gpt-oss-20b (free) openai | not evaluated | — | — | — | — | — | — | — | Free | 131K |
| — | Qwen: Qwen3 Coder 30B A3B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.070 | 160K |
| — | Qwen: Qwen3 30B A3B Instruct 2507 qwen | not evaluated | — | — | — | — | — | — | — | $0.100 | 262K |
| — | Z.ai: GLM 4.5 z-ai | not evaluated | — | — | — | — | — | — | — | $0.600 | 131K |
| — | Z.ai: GLM 4.5 Air z-ai | not evaluated | — | — | — | — | — | — | — | $0.130 | 131K |
| — | Qwen: Qwen3 235B A22B Thinking 2507 qwen | not evaluated | — | — | — | — | — | — | — | $0.300 | 262K |
| — | Qwen: Qwen3 Coder 480B A35B qwen | not evaluated | — | — | — | — | — | — | — | $0.300 | 1.0M |
| — | ByteDance: UI-TARS 7B bytedance | not evaluated | — | — | — | — | — | — | — | $0.100 | 128K |
| — | Qwen: Qwen3 235B A22B Instruct 2507 qwen | not evaluated | — | — | — | — | — | — | — | $0.090 | 262K |
| — | MoonshotAI: Kimi K2 0711 moonshotai | not evaluated | — | — | — | — | — | — | — | $0.570 | 131K |
| — | Venice: Uncensored cognitivecomputations | not evaluated | — | — | — | — | — | — | — | $0.200 | 128K |
| — | Tencent: Hunyuan A13B Instruct tencent | not evaluated | — | — | — | — | — | — | — | $0.140 | 131K |
| — | Baidu: ERNIE 4.5 VL 424B A47B baidu | not evaluated | — | — | — | — | — | — | — | $0.420 | 131K |
| — | Mistral: Mistral Small 3.2 24B mistralai | not evaluated | — | — | — | — | — | — | — | $0.100 | 131K |
| — | DeepSeek: R1 0528 deepseek | not evaluated | — | — | — | — | — | — | — | $0.500 | 164K |
| — | Google: Gemma 3n 4B google | not evaluated | — | — | — | — | — | — | — | $0.060 | 33K |
| — | Meta: Llama Guard 4 12B meta-llama | not evaluated | — | — | — | — | — | — | — | $0.180 | 164K |
| — | Qwen: Qwen3 30B A3B qwen | not evaluated | — | — | — | — | — | — | — | $0.130 | 131K |
| — | Qwen: Qwen3 8B qwen | not evaluated | — | — | — | — | — | — | — | $0.117 | 131K |
| — | Qwen: Qwen3 14B qwen | not evaluated | — | — | — | — | — | — | — | $0.120 | 132K |
| — | Qwen: Qwen3 32B qwen | not evaluated | — | — | — | — | — | — | — | $0.080 | 131K |
| — | Qwen: Qwen3 235B A22B qwen | not evaluated | — | — | — | — | — | — | — | $0.455 | 131K |
| — | Meta: Llama 4 Maverick meta-llama | not evaluated | — | — | — | — | — | — | — | $0.200 | 1.0M |
| — | Meta: Llama 4 Scout meta-llama | not evaluated | — | — | — | — | — | — | — | $0.100 | 10M |
| — | DeepSeek: DeepSeek V3 0324 deepseek | not evaluated | — | — | — | — | — | — | — | $0.270 | 164K |
| — | Mistral: Mistral Small 3.1 24B mistralai | not evaluated | — | — | — | — | — | — | — | $0.351 | 128K |
| — | Google: Gemma 3 4B google | not evaluated | — | — | — | — | — | — | — | $0.050 | 131K |
| — | Google: Gemma 3 12B google | not evaluated | — | — | — | — | — | — | — | $0.050 | 131K |
| — | Cohere: Command A cohere | not evaluated | — | — | — | — | — | — | — | $2.50 | 256K |
| — | Reka Flash 3 rekaai | not evaluated | — | — | — | — | — | — | — | $0.100 | 66K |
| — | Google: Gemma 3 27B google | not evaluated | — | — | — | — | — | — | — | $0.100 | 131K |
| — | TheDrummer: Skyfall 36B V2 thedrummer | not evaluated | — | — | — | — | — | — | — | $0.550 | 33K |
| — | Qwen: Qwen2.5 VL 72B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.800 | 131K |
| — | Mistral: Mistral Small 3 mistralai | not evaluated | — | — | — | — | — | — | — | $0.050 | 33K |
| — | DeepSeek: R1 Distill Llama 70B deepseek | not evaluated | — | — | — | — | — | — | — | $0.800 | 128K |
| — | DeepSeek: R1 deepseek | not evaluated | — | — | — | — | — | — | — | $0.700 | 164K |
| — | MiniMax: MiniMax-01 minimax | not evaluated | — | — | — | — | — | — | — | $0.200 | 1.0M |
| — | Microsoft: Phi 4 microsoft | not evaluated | — | — | — | — | — | — | — | $0.070 | 16K |
| — | DeepSeek: DeepSeek V3 deepseek | not evaluated | — | — | — | — | — | — | — | $0.200 | 131K |
| — | Sao10K: Llama 3.3 Euryale 70B sao10k | not evaluated | — | — | — | — | — | — | — | $0.650 | 131K |
| — | Meta: Llama 3.3 70B Instruct meta-llama | not evaluated | — | — | — | — | — | — | — | $0.130 | 131K |
| — | Qwen2.5 Coder 32B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.660 | 128K |
| — | TheDrummer: UnslopNemo 12B thedrummer | not evaluated | — | — | — | — | — | — | — | $0.400 | 33K |
| — | Magnum v4 72B anthracite-org | not evaluated | — | — | — | — | — | — | — | $3.00 | 33K |
| — | Qwen: Qwen2.5 7B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.040 | 131K |
| — | TheDrummer: Rocinante 12B thedrummer | not evaluated | — | — | — | — | — | — | — | $0.250 | 66K |
| — | Meta: Llama 3.2 1B Instruct meta-llama | not evaluated | — | — | — | — | — | — | — | $0.027 | 131K |
| — | Meta: Llama 3.2 3B Instruct meta-llama | not evaluated | — | — | — | — | — | — | — | $0.051 | 131K |
| — | Qwen2.5 72B Instruct qwen | not evaluated | — | — | — | — | — | — | — | $0.360 | 131K |
| — | Sao10K: Llama 3.1 Euryale 70B v2.2 sao10k | not evaluated | — | — | — | — | — | — | — | $0.850 | 131K |
| — | Nous: Hermes 3 70B Instruct nousresearch | not evaluated | — | — | — | — | — | — | — | $0.700 | 131K |
| — | Nous: Hermes 3 405B Instruct nousresearch | not evaluated | — | — | — | — | — | — | — | $1.00 | 131K |
| — | Sao10K: Llama 3 8B Lunaris sao10k | not evaluated | — | — | — | — | — | — | — | $0.040 | 8K |
| — | Meta: Llama 3.1 70B Instruct meta-llama | not evaluated | — | — | — | — | — | — | — | $0.400 | 131K |
| — | Meta: Llama 3.1 8B Instruct meta-llama | not evaluated | — | — | — | — | — | — | — | $0.050 | 131K |
| — | Mistral: Mistral Nemo mistralai | not evaluated | — | — | — | — | — | — | — | $0.019 | 131K |
| — | Google: Gemma 2 27B google | not evaluated | — | — | — | — | — | — | — | $0.650 | 8K |
| — | Mistral: Mixtral 8x22B Instruct mistralai | not evaluated | — | — | — | — | — | — | — | $2.00 | 66K |
| — | WizardLM-2 8x22B microsoft | not evaluated | — | — | — | — | — | — | — | $0.620 | 66K |
| — | ReMM SLERP 13B undi95 | not evaluated | — | — | — | — | — | — | — | $0.450 | 6K |
| — | MythoMax 13B gryphe | not evaluated | — | — | — | — | — | — | — | $0.060 | 4K |
What "open weights" means here
A model is listed here when its weights are published and downloadable — you can run it on your own hardware, fine-tune it, and keep your prompts off a third party's servers. That is a narrower claim than "open source": most of these models publish weights without publishing training data or training code, and several carry licences with commercial restrictions. Check the licence on the model card before you build a business on one.
Why the ranking is the same score
Open-weight models are scored on exactly the same LiveBench tasks and the same published weights as proprietary ones — there is no separate, easier scale. That is deliberate: an open-model ranking is only useful if it tells you what you give up, if anything, by self-hosting. Where an open model reaches the top of the overall index, that is a like-for-like result, not a handicap category.
The price column still applies: most open-weight models here are also served by hosted providers, and the figure shown is that hosted price. If you self-host, your real cost is your own compute instead, which is usually cheaper at sustained volume and more expensive at low or spiky volume. Context length, by contrast, is a property of the model and holds wherever you run it.
Models with no score are absent from the current benchmark snapshot. They keep their factual data — price, context, modality — and make no capability claim, rather than being quietly ranked last. See methodology.