LLM leaderboard

LLM Leaderboard — Model Rankings by Price & Performance

Independent ranking of 130+ frontier and open-weight language models, scored by Value Score — quality benchmark divided by blended $/1M token cost, normalized across every tracked model. Token usage, list pricing, context window, and 30-day price moves are refreshed daily so you can see which models deliver the most performance per dollar this week.

Last updated: 2026-07-17 · 55 models · sorted by Value Score (highest first)

Value Score formula: Value Score = quality benchmark ÷ blended token cost, normalized across all models. Higher = more performance per dollar.

#ModelProviderCategoryInput $/MOutput $/MContextBenchmarkValue ScoreTokensDoDPrice Δ 30d
01 Ministral 8B
ministral-8b
Mistral Budget $0.100 $0.100 128K 58 100 High Value 155B +12% 0%
02 MiMo V2.5 Lite
mimo-v2-5-lite
Xiaomi Budget $0.050 $0.200 64K 58 71 Fair Value 209B +20% 0%
03 DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek Budget $0.070 $0.280 128K 78 69 Fair Value 5.42T +64% -30%
04 DeepSeek V3.2
deepseek-v3-2
DeepSeek Budget $0.140 $0.280 128K 75 58 Fair Value 1.05T -17% -50%
05 abab6.5s-chat
minimax-abab-6-5s-chat
MiniMax Budget $0.100 $0.300 128K 60 47 Fair Value 123B +4% 0%
06 Gemini 3 Flash
gemini-3-flash
Google Budget $0.100 $0.400 1M 74 46 Fair Value 441B +13% -33%
07 Hy3 Lite
hy3-lite
Tencent Budget $0.100 $0.400 128K 64 39 Low Value 169B +38% 0%
08 Kimi K2 Mini
kimi-k2-mini
Moonshot Budget $0.100 $0.400 128K 60 37 Low Value 166B +13% 0%
09 Gemini 2.5 Flash
gemini-2-5-flash
Google Budget $0.150 $0.600 1M 70 29 Low Value 591B -26% 0%
10 MiniMax M2
minimax-m2
MiniMax Coding $0.200 $0.800 200K 74 23 Low Value 235B +65% 0%
11 MiMo V2.5
mimo-v2-5
Xiaomi Open $0.200 $0.800 128K 71 22 Low Value 2.84T +14% 0%
12 MiniMax M1
minimax-m1
MiniMax Reasoning $0.200 $0.800 128K 68 21 Low Value 195B +2% 0%
13 Llama 4 Maverick
llama-4-maverick
Meta Open $0.270 $0.850 1M 73 20 Low Value 525B -15% 0%
14 DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek Flagship $0.270 $1.10 128K 82 18 Low Value 1.76T -19% -10%
15 Grok 4.1 Fast
grok-4-1-fast
xAI Long-context $0.200 $1.00 256K 72 18 Low Value 745B -5% -20%
16 Grok Code Fast 1
grok-code-fast-1
xAI Coding $0.200 $1.00 128K 71 18 Low Value 615B +99% 0%
17 Codestral 25.10
codestral-25-10
Mistral Coding $0.300 $0.900 32K 70 18 Low Value 368B +22% 0%
18 MiniMax M3
minimax-m3
MiniMax Reasoning $0.300 $1.20 1M 80 16 Low Value 3.47T +125% 0%
19 Grok 3 Mini
grok-3-mini
xAI Budget $0.200 $1.00 64K 62 16 Low Value 220B -9% 0%
20 Kimi K2 Instruct
kimi-k2-instruct
Moonshot Open $0.300 $1.20 200K 72 15 Low Value 732B +60% 0%
21 GPT-5.4 Mini
gpt-5-4-mini
OpenAI Budget $0.300 $1.20 400K 73 15 Low Value 510B +8% -10%
22 MiMo Reasoner
mimo-reasoner
Xiaomi Reasoning $0.300 $1.20 128K 72 15 Low Value 417B +50% 0%
23 MiMo Coder
mimo-coder
Xiaomi Coding $0.300 $1.20 128K 69 14 Low Value 316B +41% 0%
24 Kimi K1.5 Long
kimi-k1-5-long
Moonshot Long-context $0.300 $1.20 1M 66 14 Low Value 311B +21% 0%
25 Hunyuan-Large
hunyuan-large
Tencent Open $0.300 $1.20 128K 67 14 Low Value 169B -3% 0%
26 Kimi K2.6
kimi-k2-6
Moonshot Coding $0.400 $1.60 256K 76 12 Low Value 972B +10% 0%
27 Kimi Coder
kimi-coder
Moonshot Coding $0.400 $1.60 256K 77 12 Low Value 528B +46% 0%
28 Grok 4 Mini
grok-4-mini
xAI Budget $0.300 $1.50 128K 70 12 Low Value 461B +15% 0%
29 Mixtral 8x22B
mixtral-8x22b
Mistral Open $0.900 $0.900 64K 64 12 Low Value 132B -21% 0%
30 Hunyuan-Turbo
hunyuan-turbo
Tencent Flagship $0.400 $1.60 128K 74 11 Low Value 424B +13% 0%
31 Devstral Medium
devstral-medium
Mistral Coding $0.500 $1.50 128K 73 11 Low Value 300B +75% 0%
32 Hy3 Preview
hy3-preview
Tencent Flagship $0.500 $2.00 256K 79 10 Low Value 3.30T +13% 0%
33 MiMo V2.5 Pro
mimo-v2-5-pro
Xiaomi Flagship $0.500 $2.00 128K 78 10 Low Value 790B +33% 0%
34 Hunyuan-T1
hunyuan-t1
Tencent Reasoning $0.500 $2.00 128K 77 9 Low Value 276B +71% 0%
35 DeepSeek R1
deepseek-r1
DeepSeek Reasoning $0.550 $2.19 64K 79 9 Low Value 258B -42% 0%
36 Owl Alpha
owl-alpha
OpenRouter Reasoning $0.600 $2.40 200K 76 8 Low Value 2.07T +14% 0%
37 Kimi K2 Thinking
kimi-k2-thinking
Moonshot Reasoning $0.600 $2.50 256K 81 8 Low Value 864B +117% 0%
38 Kimi Researcher
kimi-researcher
Moonshot Reasoning $0.600 $2.50 256K 76 8 Low Value 183B +102% 0%
39 Claude Haiku 4.5
claude-haiku-4-5
Anthropic Budget $0.800 $4.00 200K 71 5 Low Value 705B +14% 0%
40 Magistral Medium 2
magistral-medium-2
Mistral Reasoning $2.00 $5.00 128K 73 3 Low Value 458B +80% 0%
41 Pixtral Large
pixtral-large
Mistral Flagship $2.00 $6.00 128K 72 3 Low Value 217B -2% 0%
42 Gemini 3.1 Pro
gemini-3-1-pro
Google Long-context $2.50 $10.00 2M 86 2 Low Value 1.22T +29% -5%
43 Mistral Large 3
mistral-large-3
Mistral Flagship $3.00 $9.00 128K 78 2 Low Value 405B -8% 0%
44 Grok 3 Reasoning
grok-3-reasoning
xAI Reasoning $2.00 $10.00 128K 75 2 Low Value 294B -19% 0%
45 Claude Sonnet 4.6
claude-sonnet-4-6
Anthropic Coding $3.00 $15.00 200K 82 1 Low Value 2.04T +24% 0%
46 GPT-5.6
gpt-5-6
OpenAI Flagship $6.00 $36.00 400K 91 1 Low Value 1.69T +50% 0%
47 GPT-5.5
gpt-5-5
OpenAI Flagship $5.00 $20.00 400K 88 1 Low Value 1.31T +38% -12%
48 Grok 4 Heavy
grok-4-heavy
xAI Reasoning $5.00 $15.00 256K 84 1 Low Value 894B +87% 0%
49 Grok 4.1
grok-4.1
xAI Long-context $3.00 $15.00 256K 82 1 Low Value 854B +63% 0%
50 GPT-5.4
gpt-5-4
OpenAI Flagship $5.00 $20.00 400K 85 1 Low Value 846B +12% -10%
51 Grok 4
grok-4
xAI Reasoning $3.00 $15.00 256K 80 1 Low Value 319B -3% 0%
52 Sonar Pro
sonar-pro
Perplexity Long-context $3.00 $15.00 200K 70 1 Low Value 283B +17% 0%
53 Claude Fable 5
claude-fable-5
Anthropic Flagship $12.00 $60.00 200K 87 0 Low Value 2.80T +76% -8%
54 Claude Opus 4.8
claude-opus-4-8
Anthropic Flagship $15.00 $75.00 200K 89 0 Low Value 1.68T +20% 0%
55 Claude Opus 4.7
claude-opus-4-7
Anthropic Flagship $15.00 $75.00 200K 86 0 Low Value 1.33T -15% 0%

Token usage aggregated daily from Tokenscost calculator runs, public API mirrors, and OpenRouter ranking telemetry. Trailing 7-day window; figures refresh automatically each UTC day.