Best AI Models for Coding
Objective rankings based on coding benchmarks, pricing, and real-world performance. Updated weekly.
30
Models Ranked
Claude Fable 5
Top Model
100
Highest Score
Coding Leaderboard
Compare models →| # | Model | Score | Price (1M tokens) | Context |
|---|---|---|---|---|
| 1 | Claude Fable 5 anthropic | 100 | $10 / $50 per 1M | 1M |
| 2 | Kimi K3 moonshotai | 99.62 | $3 / $15 per 1M | 1M |
| 3 | GPT-5.6 terra openai | 99.5 | $10 / $60 per 1M | 1M |
| 4 | GPT-5.6 sol openai | 98.8 | $5 / $30 per 1M | 1M |
| 5 | GPT-5.6 Luna openai | 98.1 | $1.5 / $9 per 1M | 400K |
| 6 | GPT-5.5 openai | 97.91 | $5 / $30 per 1M | 1M |
| 7 | Claude Opus 4.8 anthropic | 97.07 | $5 / $25 per 1M | 1M |
| 8 | Qwen3.8 Max Preview qwen | 95 | $2.5 / $10 per 1M | 128K |
| 9 | Grok 4.5 x-ai | 94.65 | $2 / $6 per 1M | 500K |
| 10 | GLM-5.2 z-ai | 89.89 | $0.93 / $3 per 1M | 1M |
| 11 | DeepSeek V4 Flash (0731) deepseek | 88.12 | $0.09 / $0.18 per 1M | 1M |
| 12 | Qwen3.7 Max qwen | 86.24 | $1.25 / $3.75 per 1M | 1M |
| 13 | Claude Sonnet 4.6 anthropic | 82.4 | $3 / $15 per 1M | 1M |
| 14 | Kimi K2.7 Code moonshotai | 79.44 | $0.74 / $3.5 per 1M | 262K |
| 15 | MiMo V2.5 Pro mimo | 78.69 | $0.44 / $0.87 per 1M | 1M |
| 16 | DeepSeek V4 Pro deepseek | 77.61 | $0.44 / $0.87 per 1M | 1M |
| 17 | Muse Spark meta | 76.63 | $0 / $0 per 1M | 128K |
| 18 | MiniMax M3 minimax | 76.57 | $0.3 / $1.2 per 1M | 1M |
| 19 | HY3 tencent | 73.5 | $0.13 / $0.53 per 1M | 262K |
| 20 | GPT-5.4 mini openai | 73.31 | $0.75 / $4.5 per 1M | 400K |
| 21 | GLM-5.1 z-ai | 72.93 | $0.98 / $4.3 per 1M | 203K |
| 22 | Kimi K2.6 moonshotai | 72.5 | $0.55 / $3.2 per 1M | 262K |
| 23 | Gemini 3.5 Flash google | 70 | $1.5 / $9 per 1M | 1M |
| 24 | Gemini 3.1 Pro google | 65 | $2 / $12 per 1M | 1M |
| 25 | Nemotron 3 Ultra nvidia | 64.41 | $0.5 / $2.2 per 1M | 1M |
| 26 | Qwen3.5 Coder qwen | 63.03 | $0.11 / $0.8 per 1M | 262K |
| 27 | Mistral Medium 3.5 mistralai | 61.32 | $1.5 / $7.5 per 1M | 262K |
| 28 | Claude Haiku 4.5 anthropic | 57.38 | $1 / $5 per 1M | 200K |
| 29 | Gemma 4 31B google | 56.78 | $0 / $0 per 1M | 262K |
| 30 | Grok 4.3 x-ai | 55.23 | $1.25 / $2.5 per 1M | 1M |
Methodology
Our coding score aggregates multiple benchmarks including HumanEval, MBPP, SWE-bench, and LiveCodeBench. Scores are weighted by benchmark relevance to real-world coding tasks. Pricing reflects standard API rates as of June 2026.