Intelligence Index benchmark
453 AI models ranked on Intelligence Index. Each row shows the model’s score next to its cheapest API price per 1M tokens where it has one, so you can weigh quality against cost.
Top model today
Claude Opus 5 (Anthropic) leads at 63.1, from $5.00 per 1M input.
Data via Artificial Analysis methodology →Models ranked
453
on this benchmark
Top score
63.1
current leader
Type
Composite
index
Updated
30 Aug 2026
last source fetch
Cost vs Intelligence Index
263 priced modelsEach labelled name links to that model’s page. Curated to the cost-efficiency frontier, top scorers and cheapest so labels stay readable; the full field is in the leaderboard table below (and via “select all”). Reasoning models (extended “thinking” before answering) carry a ring around the dot.
Models ranked on Intelligence Index
highest score first · 453 models| # | ||||||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1 | Claude Opus 5 | 63.1 | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| 2 | Claude Fable 5 | 62.1 | $10.00 | $50.00 | ||||||||||||||||||||||||||||||
| 3 | GPT-5.6 Sol | 60.9 | $2.00 | $10.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 4 | Grok 4.6 | 60.9 | $2.00 | $6.00 | ||||||||||||||||||||||||||||||
| 5 | Kimi K3 | 59.7 | $2.85 | $14.25 | ||||||||||||||||||||||||||||||
| 6 | GLM 5.3 | 59.5 | $1.40 | $4.40 | ||||||||||||||||||||||||||||||
| 7 | Qwen3.8 Max | 58.1 | $1.65 | $4.95 | ||||||||||||||||||||||||||||||
| 8 | Qwen3.8 2.4T A95B | 57.7 | $2.00 | $6.00 | ||||||||||||||||||||||||||||||
| 9 | GLM 5.3 Flash | 57.5 | $0.07 | $0.25 | ||||||||||||||||||||||||||||||
| 10 | Claude Opus 4.8 | 57.3 | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| 11 | Muse Spark 1.2 | Other | 56.8 | $1.25 | $4.25 | |||||||||||||||||||||||||||||
| 12 | GPT-5.6 Terra | 56.6 | $2.00 | $12.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 13 | GPT-5.5 | 56.3 | $5.00 | $30.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 14 | Gemini 3.7 Flash | 56.0 | $0.75 | $3.75 | ||||||||||||||||||||||||||||||
| 15 | Qwen3.8-Flash-Next | 55.8 | n/a | n/a | ||||||||||||||||||||||||||||||
| 16 | Grok 4.5 | 55.8 | $2.00 | $6.00 | ||||||||||||||||||||||||||||||
| 17 | Claude Sonnet 5 | 55.3 | $2.00 | $10.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 18 | Claude Opus 4.7 | 55.0 | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 19 | Muse Spark 1.1 | Other | 53.2 | $1.25 | $4.25 | |||||||||||||||||||||||||||||
| 20 | DeepSeek V4 Pro 0423 | 53.2 | $0.43 | $0.87 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 21 | GPT-5.4 | 53.1 | $2.50 | $15.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 22 | GLM 5.2 | 52.6 | $0.61 | $1.98 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 23 | GPT-5.6 Luna | 52.3 | $0.20 | $1.20 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 24 | Qwen3.8 27B | 52.0 | $0.40 | $2.55 | ||||||||||||||||||||||||||||||
| 25 | Gemini 3.5 Flash | 52.0 | $1.50 | $9.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 26 | DeepSeek V4 Flash 0423 | 51.8 | $0.08 | $0.17 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 27 | Gemini 3.6 Flash | 51.6 | $0.75 | $3.75 | ||||||||||||||||||||||||||||||
| 28 | DeepSeek V4 Flash Vision (Reasoning, Max Effort) | 51.5 | n/a | n/a | ||||||||||||||||||||||||||||||
| 29 | Agnes 2.5 Pro Beta | Sapiens-ai | 49.1 | n/a | n/a | |||||||||||||||||||||||||||||
| 30 | Claude Sonnet 4.6 | 48.4reasoning: adaptive | $3.00 | $15.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 31 | Gemini 3.1 Pro Preview | 47.7 | $2.00 | $12.00 | ||||||||||||||||||||||||||||||
| 32 | Motif 3 | Motif-technologies | 47.4 | n/a | n/a | |||||||||||||||||||||||||||||
| 33 | Qwen3.7 Max | 46.7 | $1.25 | $3.75 | ||||||||||||||||||||||||||||||
| 34 | GPT-5.3-Codex | 45.5 | $1.75 | $14.00 | ||||||||||||||||||||||||||||||
| 35 | MiniMax M3 | 45.4 | $0.23 | $0.96 | ||||||||||||||||||||||||||||||
| 36 | Motif 3 (Beta) | Motif-technologies | 45.3 | n/a | n/a | |||||||||||||||||||||||||||||
| 37 | DeepSeek V4 Pro (max) | 45.3 | n/a | n/a | ||||||||||||||||||||||||||||||
| 38 | Kimi K2.6 | 45.1 | $0.65 | $3.40 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 39 | Claude Opus 4.6 | 44.9reasoning: adaptive | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 40 | Muse Spark | 44.3 | n/a | n/a | ||||||||||||||||||||||||||||||
| 41 | GPT-5.2 | 43.3 | $1.75 | $14.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 42 | Kimi K2.7 Code | 43.0 | $0.66 | $3.40 | ||||||||||||||||||||||||||||||
| 43 | MiMo-V2.5-Pro | 42.9 | $0.43 | $0.87 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 44 | Inkling | Other | 42.3 | $0.95 | $4.05 | |||||||||||||||||||||||||||||
| 45 | Hy3 | Other | 42.2 | $0.13 | $0.53 | |||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 46 | DeepSeek V4 Flash (max) | 42.1 | n/a | n/a | ||||||||||||||||||||||||||||||
| 47 | Claude Opus 4.5 | 41.9reasoning: true | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 48 | Nex-N2-Pro | Other | 41.7 | $0.25 | $1.00 | |||||||||||||||||||||||||||||
| 49 | Solar Pro 4 | Other | 41.6 | $0.03 | $0.12 | |||||||||||||||||||||||||||||
| 50 | MiMo-V2-Pro | 41.4 | n/a | n/a | ||||||||||||||||||||||||||||||
Frequently asked questions
What is the Intelligence Index benchmark?
Artificial Analysis's headline composite score (0 to 100) that combines many independent evaluations across reasoning, knowledge, science, coding and agentic tasks into a single number, each run independently by Artificial Analysis. The exact component evals and their weights are versioned by Artificial Analysis; see their methodology for the current makeup.
Which AI model scores highest on Intelligence Index?
Claude Opus 5 (Anthropic) leads with 63.1, from $5.00 per 1M input tokens.
How many models are ranked on Intelligence Index?
453 models carry a Intelligence Index score. Scores come from Artificial Analysis; where a model has a live API price it is verified against LiteLLM and OpenRouter, and models with no current price are listed without one.
Other benchmarks
compare the same models on a different evalEvery weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI