AI Calculator Pro

Intelligence per dollar

The smartest model is rarely the one you should ship. This board ranks every major LLM by community Arena score and by intelligence per dollar, and flags the models on the price/quality frontier — the ones nothing else beats on both quality and cost. Pick an arena, then sort by what matters for your task.

Decision-ready picks

Skip the leaderboard — here’s the one model that best answers each common question (overall arena).

Best overall
Kimi K3
Moonshot AI · Arena 1473

The highest-scoring model we track for this arena — the quality ceiling.

Best value
Amazon Nova Micro
AWS Bedrock · 20702 Elo per $/1M

The most Arena score per dollar — punches far above its price.

Cheapest that clears the bar
Amazon Nova Micro
AWS Bedrock · $0.061/1M blended

The lowest price among models still within 15% of the top Arena score.

Largest useful context
Llama 4 Scout
Meta · 10M tokens

The biggest context window among genuinely capable models — for long docs and codebases.

Best for coding
Kimi K3
Moonshot AI · Coding Arena 1679

Top of the coding Arena — the safest pick for agents and multi-file edits.

Price vs performance

Every scored model plotted by price and Arena score. The best value sits toward the top-left.

Price vs quality — lower and higher is better. Highlighted points are on the price/quality frontier (nothing we track beats them on both). Price uses a 3:1 blended $/1M on a log scale.
1189128413781473$0.100$0.250$0.500$1.00$2.00$5.00$10$25Blended price $/1M (log scale) →Arena score →Amazon Nova LiteAmazon Nova MicroDeepSeek V4 FlashKimi K3GLM-5.2MiMo-V2.5-ProHunyuan Hy3GLM-4.6GLM-4.7-FlashGPT-5 NanoStep 3.5 Flash

11 of 127 models are on the price/quality frontier (marked Best value). An Est. tag means the score is inherited from the base model (a regional/creator-prefixed hosting duplicate). Click a column to sort; click again to flip.

1Kimi K3Best value1473$6.00246Moonshot AI
2GPT-5.51465$11.25130OpenAI
3GPT-5.5Est.1465$12.38118AWS Bedrock
4Gemini 2.5 Pro1463$3.44426Google
5GLM-5.2Best value1463$2.15680Z.ai (Zhipu AI)
6GPT-5.41461$5.63260OpenAI
7GPT-5.4Est.1461$6.19236AWS Bedrock
8GPT-5.6 Sol1455$11.25129OpenAI
9GPT-5.6 SolEst.1455$11.25129AWS Bedrock
10Claude Fable 51454$20.0073Anthropic
11Claude Fable 5Est.1454$20.0073AWS Bedrock
12Claude Opus 4.81453$10.00145Anthropic
13Claude Opus 4.8Est.1453$10.00145AWS Bedrock
14Grok 4.51451$3.00484xAI
15Claude Opus 4.61449$10.00145Anthropic
16Claude Opus 4.5 (latest)1449$10.00145Anthropic
17Claude Opus 5 (JP)Est.1449$10.00145AWS Bedrock
18Claude Opus 51449$10.00145Anthropic
19GLM-4.6Best value1448$1.001448Z.ai (Zhipu AI)
20Claude Sonnet 4.61445$6.00241Anthropic
21MiMo-V2.5-ProBest value1445$0.5442657Xiaomi
22AU Anthropic Claude Sonnet 4.6Est.1445$6.60219AWS Bedrock
23Hunyuan Hy3Best value1444$0.3504126Tencent
24Qwen3.7 Max1444$3.75385Alibaba
25Claude Sonnet 51440$4.00360Anthropic
26Claude Opus 4.7 (US)Est.1440$10.00144AWS Bedrock
27Claude Sonnet 5 (Global)Est.1440$4.00360AWS Bedrock
28Claude Opus 4.71440$10.00144Anthropic
29Qwen3.5 397B-A17B1439$1.351066Alibaba
30Gemini 3.5 Flash Lite1439$0.8501693Google
31Gemini 3.6 Flash1439$3.00480Google
32GLM-51438$1.55928Z.ai (Zhipu AI)
33Claude Sonnet 4.5 (latest)1438$6.00240Anthropic
34GLM-5Est.1438$1.55928AWS Bedrock
35Kimi K2.61435$1.71838Moonshot AI
36MiMo-V2-Pro1435$0.5442639Xiaomi
37DeepSeek V4 Pro1434$0.5442637DeepSeek
38Gemini 3.5 Flash1434$3.38425Google
39Qwen3.7 Plus1433$1.131274Alibaba
40GLM-4.71432$1.001432Z.ai (Zhipu AI)
41GLM-4.7Est.1432$1.001432AWS Bedrock
42DeepSeek V4 FlashBest value1430$0.1758171DeepSeek
43GLM-4.51429$1.001429Z.ai (Zhipu AI)
44GLM-5.11429$2.15665Z.ai (Zhipu AI)
45MiMo-V2.51427$0.1758154Xiaomi
46GLM-5V-Turbo1424$1.90749Z.ai (Zhipu AI)
47Mistral Large 31423$0.7501897Mistral AI
48o31422$3.50406OpenAI
49GPT-5.11422$3.44414OpenAI
50MiMo-V2-Omni1422$0.1758126Xiaomi
51MiniMax-M31421$0.5252707MiniMax
52Kimi K2 Thinking Turbo1421$2.86496Moonshot AI
53DeepSeek V3.21420$0.3154508DeepSeek
54Kimi K2.51420$1.201183Moonshot AI
55Kimi K2.5Est.1420$1.201183AWS Bedrock
56Qwen3.6 Plus1418$1.131260Alibaba
57Qwen3.5 122B-A10B1418$1.101289Alibaba
58Gemini 2.5 Flash1417$0.8501667Google
59Claude Opus 4.1 (latest)1417$30.0047Anthropic
60Gemini 3.1 Flash Lite1415$0.5632516Google
61GPT-5.4 mini1413$1.69837OpenAI
62GPT-5.21412$4.81293OpenAI
63Muse Spark 1.11410$2.00705Meta
64Qwen3.5 27B1408$0.8251707Alibaba
65GPT-51405$3.44409OpenAI
66MiniMax-M2.71405$0.5252676MiniMax
67Qwen3 Max1404$2.40585Alibaba
68Step 3.5 FlashBest value1404$0.1509360StepFun
69Qwen/Qwen3-VL-235B-A22B-InstructEst.1401$0.6002335AWS Bedrock
70Qwen3-VL 235B-A22B1401$1.221144Alibaba
71Amazon Nova Premier1400$5.00280AWS Bedrock
72Grok 4.31400$1.56896xAI
73Qwen3.5 35B-A3B1396$0.6882031Alibaba
74MiMo-V2-Flash1395$0.1757971Xiaomi
75Claude Haiku 4.51394$2.00697Anthropic
76Qwen3-Next 80B-A3B Instruct1392$0.8751591Alibaba
77MiniMax-M2.11391$0.5252650MiniMax
78MiniMax M2.1Est.1391$0.5252650AWS Bedrock
79GLM-4.5-Air1383$0.4253254Z.ai (Zhipu AI)
80o4-mini1382$1.93718OpenAI
81GLM-4.6V1375$0.4503056Z.ai (Zhipu AI)
82GPT-5 Mini1374$0.6881999OpenAI
83GPT-5.4 nano1373$0.4632969OpenAI
84GPT-4.11372$3.50392OpenAI
85Kimi K2 Thinking1371$1.081275Moonshot AI
86Kimi K2 ThinkingEst.1371$1.081275AWS Bedrock
87Qwen3-Next 80B-A3B (Thinking)1368$1.88730Alibaba
88Qwen3 235B-A22B1366$1.221115Alibaba
89Command A1365$4.38312Cohere
90Llama 4 Maverick1360$0.3783603Meta
91MiniMax-M2.51359$0.5252589MiniMax
92MiniMax M2.5Est.1359$0.5252589AWS Bedrock
93Qwen3-Coder 480B-A35B Instruct1356$3.00452Alibaba
94GPT-4o1355$4.38310OpenAI
95Gemini 2.0 Flash1354$0.1757737Google
96GLM-4.7-FlashBest value1353$0.1459331Z.ai (Zhipu AI)
97o11353$26.2552OpenAI
98GLM-4.7-FlashEst.1353$0.1538872AWS Bedrock
99Mistral Medium 3.11350$0.8001688Mistral AI
100Amazon Nova Pro1348$1.40963AWS Bedrock
101GPT-4.1 Mini1345$0.7001921OpenAI
102MiniMax-M21342$0.5252556MiniMax
103MiniMax M2Est.1342$0.5252556AWS Bedrock
104Devstral 21340$0.5252552Mistral AI
105Qwen3 32B1340$1.221094Alibaba
106GLM-4.5V1334$0.9001482Z.ai (Zhipu AI)
107Llama 4 Scout1332$0.1687952Meta
108Gemini 2.5 Flash-Lite1330$0.1757600Google
109Qwen Plus1327$0.6002212Alibaba
110Mistral Small 41325$0.2625048Mistral AI
111Step 2 (16K)1321$8.02165StepFun
112GPT-5 NanoBest value1320$0.1389600OpenAI
113o3-mini1319$1.93685OpenAI
114Llama 3.3 70B1315$0.6002192Meta
115GPT-4o mini1310$0.2624990OpenAI
116GPT-4.1 Nano1300$0.1757429OpenAI
117Amazon Nova LiteBest value1298$0.10512362AWS Bedrock
118Qwen Max1282$2.80458Alibaba
119Llama-3.3-70B-Instruct1275$0.1986456Meta
120Qwen2.5 72B Instruct1269$2.45518Alibaba
121Amazon Nova MicroBest value1268$0.06120702AWS Bedrock
122Command R7B1262$0.06619230Cohere
123Phi 41217$0.08813909Microsoft
124Command R+1204$4.38275Cohere
125GPT-41186$37.5032OpenAI
126Command R1163$0.2624430Cohere
127GPT-3.5-turbo1094$0.7501459OpenAI

Intelligence scores are community Arena ratings from LMArena / arena.ai, used under CC BY 4.0. Snapshot last refreshed 28 July 2026. Scores are a relative signal, not an absolute measure of capability. A Measured score was voted on directly; an Estimated score is inherited from an identical base model (a regional/creator-prefixed hosting duplicate). See our methodology.

Download / cite this dataset

The full ranking is available as open data — reuse it with attribution and a link back.

Cite as: AI Calculator Pro, “LLM intelligence per dollar”, https://aicalculatorpro.com/intelligence/ (snapshot 2026-07-28). Underlying scores LMArena / arena.ai, CC BY 4.0.

FAQ

What does the Arena score mean?

It is a community Elo rating from blind, head-to-head votes: users compare two anonymous model answers and pick the better one. A higher score means the model is preferred more often. It measures human preference, not correctness on any single benchmark.

What is 'intelligence per dollar'?

We divide a model's Arena score by its blended price (a 3:1 input:output blend, $ per 1M tokens). It surfaces models that punch above their price. Use it as a starting point, then confirm the model clears the quality bar your task actually needs.

Why isn't every model listed?

We only show a score when we can confidently map an Arena entry to a model we track. Models without a confident match carry no score rather than a wrong one. Scores come from LMArena / arena.ai and were last refreshed 2026-07-28.

What do the “Measured” and “Estimated” labels mean?

Measured means the Arena leaderboard voted on that exact model. Estimated means the model is a regional or creator-prefixed hosting listing of a base model (for example an AWS Bedrock “jp.anthropic.…” profile) — the weights are identical, so we inherit the base model's score, but Arena never voted on that specific listing. We label it rather than hide it, so you always know whether a score is direct or inherited.

Cite / link to this page

You’re welcome to reference this page and its figures with attribution and a link back.

https://aicalculatorpro.com/intelligence/
<a href="https://aicalculatorpro.com/intelligence/">Intelligence per dollar — LLM rankings — AI Calculator Pro</a>