Coding Index benchmark
184 AI models ranked on Coding Index. Each row shows the model’s score next to its cheapest API price per 1M tokens where it has one, so you can weigh quality against cost.
Top model today
GPT-5.6 Sol (OpenAI) leads at 78.3, from $2.00 per 1M input.
Data via Artificial Analysis →Models ranked
184
on this benchmark
Top score
78.3
current leader
Type
Composite
index
Updated
29 Aug 2026
last source fetch
Cost vs Coding Index
120 priced modelsEach labelled name links to that model’s page. Curated to the cost-efficiency frontier, top scorers and cheapest so labels stay readable; the full field is in the leaderboard table below (and via “select all”). Reasoning models (extended “thinking” before answering) carry a ring around the dot.
Models ranked on Coding Index
highest score first · 184 models| # | ||||||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1 | GPT-5.6 Sol | 78.3xhigh | $2.00 | $10.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 2 | Claude Opus 5 | 78.0 | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| 3 | Grok 4.6 | 76.8 | $2.00 | $6.00 | ||||||||||||||||||||||||||||||
| 4 | GPT-5.6 Terra | 76.7 | $2.00 | $12.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 5 | Claude Fable 5 | 76.5 | $10.00 | $50.00 | ||||||||||||||||||||||||||||||
| 6 | Kimi K3 | 76.2 | $2.85 | $14.25 | ||||||||||||||||||||||||||||||
| 7 | Gemini 3.7 Flash | 76.1 | $0.38 | $1.88 | ||||||||||||||||||||||||||||||
| 8 | GPT-5.5 | 74.9 | $5.00 | $30.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 9 | GLM 5.3 | 74.8 | $1.40 | $4.40 | ||||||||||||||||||||||||||||||
| 10 | Claude Opus 4.8 | 74.3 | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| 11 | Claude Opus 4.7 | 73.6 | $5.00 | $25.00 | ||||||||||||||||||||||||||||||
| 12 | Qwen3.8-Flash-Next | 73.1 | n/a | n/a | ||||||||||||||||||||||||||||||
| 13 | Grok 4.5 | 72.4 | $2.00 | $6.00 | ||||||||||||||||||||||||||||||
| 14 | Muse Spark 1.2 | Other | 72.2 | $1.25 | $4.25 | |||||||||||||||||||||||||||||
| 15 | Qwen3.8 2.4T A95B | 71.9 | $2.00 | $6.00 | ||||||||||||||||||||||||||||||
| 16 | Qwen3.8 Max | 71.8 | $1.65 | $4.95 | ||||||||||||||||||||||||||||||
| 17 | Claude Sonnet 5 | 71.5 | $2.00 | $10.00 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 18 | GLM 5.3 Flash | 71.5 | $0.07 | $0.25 | ||||||||||||||||||||||||||||||
| 19 | GPT-5.6 Luna | 71.4 | $0.20 | $1.20 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 20 | Muse Spark 1.1 | Other | 71.3 | $1.25 | $4.25 | |||||||||||||||||||||||||||||
| 21 | GPT-5.4 | 71.1 | $2.50 | $15.00 | ||||||||||||||||||||||||||||||
| 22 | Gemini 3.5 Flash | 70.1 | $1.50 | $9.00 | ||||||||||||||||||||||||||||||
| 23 | Gemini 3.6 Flash | 69.2 | $0.75 | $3.75 | ||||||||||||||||||||||||||||||
| 24 | DeepSeek V4 Flash 0423 | 69.1 | $0.09 | $0.17 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 25 | DeepSeek V4 Pro 0423 | 68.8 | $0.43 | $0.87 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 26 | Gemini 3.1 Pro Preview | 68.8 | $2.00 | $12.00 | ||||||||||||||||||||||||||||||
| 27 | GLM 5.2 | 68.8 | $0.61 | $1.98 | ||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||
| 28 | Qwen3.8 27B | 68.1 | $0.40 | $2.55 | ||||||||||||||||||||||||||||||
| 29 | Qwen3.7 Max | 66.0 | $1.25 | $3.75 | ||||||||||||||||||||||||||||||
| 30 | DeepSeek V4 Flash Vision (Reasoning, Max Effort) | 65.0 | n/a | n/a | ||||||||||||||||||||||||||||||
| 31 | Motif 3 | Motif-technologies | 63.5 | n/a | n/a | |||||||||||||||||||||||||||||
| 32 | Claude Sonnet 4.6 | 63.0reasoning: adaptive | $3.00 | $15.00 | ||||||||||||||||||||||||||||||
| 33 | Agnes 2.5 Pro Beta | Sapiens-ai | 62.3 | n/a | n/a | |||||||||||||||||||||||||||||
| 34 | Motif 3 (Beta) | Motif-technologies | 62.0 | n/a | n/a | |||||||||||||||||||||||||||||
| 35 | Kimi K2.6 | 61.8 | $0.65 | $3.40 | ||||||||||||||||||||||||||||||
| 36 | Kimi K2.7 Code | 60.8 | $0.66 | $3.40 | ||||||||||||||||||||||||||||||
| 37 | MiMo-V2.5-Pro | 60.2 | $0.43 | $0.87 | ||||||||||||||||||||||||||||||
| 38 | KAT-Coder-Pro V2 | Other | 59.5 | $0.30 | $1.20 | |||||||||||||||||||||||||||||
| 39 | DeepSeek V4 Pro (max) | 59.4 | n/a | n/a | ||||||||||||||||||||||||||||||
| 40 | Nex-N2-Pro | Other | 59.1 | $0.25 | $1.00 | |||||||||||||||||||||||||||||
| 41 | Hy3 | Other | 58.8 | $0.13 | $0.53 | |||||||||||||||||||||||||||||
| 42 | Agnes 2.5 Pro Alpha | Sapiens-ai | 58.8 | n/a | n/a | |||||||||||||||||||||||||||||
| 43 | Muse Spark | 58.6 | n/a | n/a | ||||||||||||||||||||||||||||||
| 44 | MiniMax M3 | 58.6 | $0.23 | $0.96 | ||||||||||||||||||||||||||||||
| 45 | MiMo-V2.5 | 56.8 | n/a | n/a | ||||||||||||||||||||||||||||||
| 46 | DeepSeek V4 Flash (max) | 56.2 | n/a | n/a | ||||||||||||||||||||||||||||||
| 47 | GPT-5.4 Mini | 56.1 | $0.75 | $4.50 | ||||||||||||||||||||||||||||||
| 48 | GPT-5.4 Nano | 56.1 | $0.20 | $1.25 | ||||||||||||||||||||||||||||||
| 49 | Qwen3.7 Plus | 55.9 | $0.32 | $1.28 | ||||||||||||||||||||||||||||||
| 50 | GLM 5.1 | 55.8 | $1.05 | $3.50 | ||||||||||||||||||||||||||||||
Frequently asked questions
What is the Coding Index benchmark?
Artificial Analysis's composite coding score (0 to 100), averaging coding evaluations such as Terminal-Bench (terminal-based software-engineering and sysadmin tasks) and SciCode (scientific-computing Python problems) and normalising to a 0 to 100 scale. Higher means stronger overall coding performance.
Which AI model scores highest on Coding Index?
GPT-5.6 Sol (OpenAI) leads with 78.3, from $2.00 per 1M input tokens.
How many models are ranked on Coding Index?
184 models carry a Coding Index score. Scores come from Artificial Analysis; where a model has a live API price it is verified against LiteLLM and OpenRouter, and models with no current price are listed without one.
Other benchmarks
compare the same models on a different evalEvery weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI