AI Coding Benchmarks & Rankings Leaderboard
Empirical evaluations measuring coding proficiency, reasoning benchmarks (GPQA, SWE-bench Verified, HLE), and real-time value-for-money ($/point) metrics.
Highlights
Updated Daily • Empirical TelemetryIntelligence
Artificial Analysis Intelligence Index • Higher is better
Speed
Output tokens per second • Higher is better
Cost per 1M Tokens
Weighted average blended price (USD) • Lower is better
Artificial Analysis Intelligence Index
↗Intelligence Index v4.3 incorporates evaluations: SWE-bench Verified (40%), Terminal-Bench 2.0 (25%), Aider Polyglot (20%), LiveCodeBench (15%). Every figure is verified against primary source logs without prompt leakage.
Intelligence Index vs. Cost per 1M Tokens
Code Intelligence Index vs. Weighted average cost (USD, Log Scale)
Frontier Language Model Intelligence, Over Time
Progressive evolution across Anthropic, OpenAI, DeepSeek, Google, and Meta from 2023 to 2026.
Canonical Benchmark Suite
Breakdown across isolated execution harnesses matching Artificial Analysis methodologies.
Value-for-Money Intelligence vs. Cost Matrix
| RANK | MODEL & CREATOR | |||||||
|---|---|---|---|---|---|---|---|---|
| 01 | GPT-6 Astra (max) OpenAI • API • ▲ +29 | —% | —% | 1M | 57 t/s TTFT: 283660ms | $12.0000 | — | |
| 02 | Claude Fable 5.1 (max with fallback) Anthropic • API • ▼ -1 | —% | —% | 1M | 68 t/s TTFT: 262210ms | $30.0000 | — | |
| 03 | Claude Opus 5 (max) Anthropic • API • ▲ +27 | —% | —% | 1M | 51 t/s TTFT: 57030ms | $20.0000 | — | |
| 04 | Muse Spark 1.3 (max) Meta • API • ▲ +26 | —% | —% | 1M | 214 t/s TTFT: 24160ms | $2.2500 | — | |
| 05 | GPT-5.6 Sol (max) OpenAI • API • ▲ +25 | —% | —% | 1M | 61 t/s TTFT: 129500ms | $6.1200 | — | |
| 06 | GLM-5.3 (max) Z AI • API • ▲ +24 | —% | —% | 1M | 72 t/s TTFT: 2960ms | $1.5000 | — | |
| 07 | Qwen3.8 Max (0902) Alibaba • API • ▼ -1 | —% | —% | 984K | 37 t/s TTFT: 2720ms | $3.0000 | — | |
| 08 | Step 5 Preview StepFun • API • ▲ +22 | —% | —% | 1M | 100 t/s TTFT: 2960ms | $1.2000 | — | |
| 09 | Grok 4.6 (high) SpaceXAI • API • ▼ -1 | —% | —% | 500K | 60 t/s TTFT: 44250ms | $4.0000 | — | |
| 10 | Kimi K3 (max) Kimi • API • ▼ -1 | —% | —% | 1.05M | 38 t/s TTFT: 3990ms | $1.8000 | — | |
| 11 | GPT-5.6 Terra (max) OpenAI • API • ▲ +19 | —% | —% | 1M | 84 t/s TTFT: 220860ms | $3.1500 | — | |
| 12 | GLM-5.3-Flash Z AI • API • ▲ +18 | —% | —% | 1M | 95 t/s TTFT: 2500ms | $0.3500 | — | |
| 13 | Gemini 3.8 Flash (high) Google • API • ▲ +17 | —% | —% | 1M | 298 t/s TTFT: 15470ms | $1.3100 | — | |
| 14 | DeepSeek V4.1 Flash (max) DeepSeek • API • ▲ +16 | —% | —% | 1M | 208 t/s TTFT: 1140ms | $0.3700 | — | |
| 15 | Qwen3.8-Flash-Next Alibaba • API • ▼ -1 | —% | —% | 256K | 55 t/s TTFT: 2890ms | $0.4500 | — | |
| 16 | Qwen3.8 2.4T A95B Alibaba • API • ▼ -1 | —% | —% | 984K | 38 t/s TTFT: 2690ms | $2.2500 | — | |
| 17 | Claude Sonnet 5 (max) Anthropic • API • ▲ +13 | —% | —% | 1M | 74 t/s TTFT: 133400ms | $8.0000 | — | |
| 18 | GPT-5.6 Luna (max) OpenAI • API • ▼ -1 | —% | —% | 1M | 133 t/s TTFT: 113240ms | $0.5200 | — | |
| 19 | DeepSeek V4 Pro 0813 (max) DeepSeek • API • ▲ +11 | —% | —% | 1M | 88 t/s TTFT: 1700ms | $0.9600 | — | |
| 20 | DeepSeek V4 Flash Vision (max) DeepSeek • API • ▲ +10 | —% | —% | 1M | 214 t/s TTFT: 1050ms | $0.4700 | — | |
| 21 | Qwen3.8 27B (xhigh) Alibaba • API • ▲ +9 | —% | —% | 256K | 43 t/s TTFT: 3840ms | $0.9000 | — | |
| 22 | K2 Horizon 375B A23B Institute of Foundation Models • API • ▲ +8 | —% | —% | 524K | 80 t/s TTFT: 3500ms | $1.2000 | — | |
| 23 | MiniMax-M3 MiniMax • API • ▲ +7 | —% | —% | 1M | 186 t/s TTFT: 1000ms | $0.6000 | — | |
| 24 | Gemini 3.1 Pro Preview Google • API • ▲ +6 | —% | —% | 1M | 119 t/s TTFT: 33160ms | $1.4000 | — | |
| 25 | Inkling Thinking Machines • API • ▲ +5 | —% | —% | 1M | 81 t/s TTFT: 2480ms | $0.7500 | — | |
| 26 | Nemotron 3 Ultra NVIDIA • API • ▲ +4 | —% | —% | 262K | 164 t/s TTFT: 2400ms | $0.7500 | — | |
| 27 | Gemini 3.5 Flash-Lite Google • API • ▲ +3 | —% | —% | 1M | 358 t/s TTFT: 9780ms | $0.1700 | — | |
| 28 | Muse Glimmer (high) Meta • API • ▲ +2 | —% | —% | 131K | 93 t/s TTFT: 990ms | $0.1200 | — | |
| 29 | Mistral Medium 3.5 Mistral • API • ▲ +1 | —% | —% | 256K | 137 t/s TTFT: 2290ms | $0.6000 | — | |
| 30 | gpt-oss-120b (high) OpenAI • API • ▼ -1 | —% | —% | 131K | 181 t/s TTFT: 850ms | $0.1800 | — |