Benchmarks / Artificial Analysis

Reported by Artificial Analysis

Artificial Analysis

Artificial Analysis's blend of its coding evaluations.

Last updated 8 Oct 2026

Results dated
8 Oct 2026
Results
188 configurations of 131 models
Unit
index score
Licence
Artificial Analysis commercial data licence

Artificial Analysis Coding Index: GPT-5.5

Top 15 of 131 results · index score, higher is better. Choose a model to highlight it.Clear highlight

  1. 1 Claude Fable 5.1 (max reasoning)Anthropic 81.6
  2. 2 GPT-5.6 Sol (extra-high reasoning)OpenAI 78.3
  3. 3 Claude Opus 5 (max reasoning)Anthropic 78.0
  4. 4 GPT-6 Astra (high reasoning)OpenAI 77.1
  5. 5 Grok 4.6 (high reasoning)xAI 76.8
  6. 6 GPT-5.6 Terra (max reasoning)OpenAI 76.7
  7. 7 Claude Fable 5 (max reasoning)Anthropic 76.5
  8. 7 Muse Spark 1.3 (extra-high reasoning)Meta 76.5
  9. 9 Gemini 3.8 Flash (high reasoning)Google 76.3
  10. 10 Qwen3.8-Max (0902)Alibaba 76.2
  11. 10 Kimi K3 (max reasoning)Moonshot AI 76.2
  12. 12 Gemini 3.7 Flash (high reasoning)Google 76.1
  13. 13 GPT-5.5 (extra-high reasoning)OpenAI 74.9
  14. 14 GLM-5.3 (max reasoning)Z.ai 74.8
  15. 15 Claude Opus 4.8 (max reasoning)Anthropic 74.3

Full results

Artificial Analysis: Artificial Analysis Coding Index, index score, higher is better
#ModelArtificial Analysis Coding Index
index score, higher is better
Price
$ per million tokens, in / out
1 Claude Fable 5.1 (max reasoning)Anthropic · best of 5 settings
81.6
$10 / $50
2 GPT-5.6 Sol (extra-high reasoning)OpenAI · best of 6 settings
78.3
$4 / $20
3 Claude Opus 5 (max reasoning)Anthropic · best of 5 settings
78.0
$5 / $25
4 GPT-6 Astra (high reasoning)OpenAI · best of 5 settings
77.1
$10 / $50
5 Grok 4.6 (high reasoning)xAI · best of 4 settings
76.8
$2 / $6
6 GPT-5.6 Terra (max reasoning)OpenAI · best of 6 settings
76.7
$2 / $12
7 Claude Fable 5 (max reasoning)Anthropic
76.5
$10 / $50
7 Muse Spark 1.3 (extra-high reasoning)Meta · best of 2 settings
76.5
$1.25 / $4.25
9 Gemini 3.8 Flash (high reasoning)Google · best of 3 settings
76.3
$1.50 / $7.50
10 Qwen3.8-Max (0902)Alibaba
76.2
$2 / $6
10 Kimi K3 (max reasoning)Moonshot AI · best of 2 settings
76.2
$3 / $15
12 Gemini 3.7 Flash (high reasoning)Google · best of 3 settings
76.1
$1.50 / $7.50
13 GPT-5.5 (extra-high reasoning)OpenAI · best of 5 settings
74.9
$5 / $30
14 GLM-5.3 (max reasoning)Z.ai
74.8
$1.40 / $4.40
15 Claude Opus 4.8 (max reasoning)Anthropic
74.3
$5 / $25
16 Claude Opus 4.7 (max reasoning)Anthropic
73.6
$5 / $25
17 Qwen3.8-Flash-NextAlibaba
73.1
–
18 Grok 4.5 (high reasoning)xAI
72.4
$2 / $6
19 Muse Spark 1.2 (extra-high reasoning)Meta
72.2
$1.25 / $4.25
20 Qwen3.8-2.4T-A95BAlibaba
71.9
$2 / $6
21 Qwen3.8-Max (0803)Alibaba
71.8
–
22 Claude Sonnet 5 (max reasoning)Anthropic · best of 2 settings
71.5
$2 / $10
22 GLM-5.3-FlashZ.ai
71.5
$0.15 / $0.50
24 GPT-5.6 Luna (max reasoning)OpenAI · best of 6 settings
71.4
$0.20 / $1.20
25 Muse Spark 1.1 (extra-high reasoning)Meta
71.3
$1.25 / $4.25
26 GPT-5.4 (extra-high reasoning)OpenAI
71.1
$2.50 / $15
27 Gemini 3.5 Flash (high reasoning)Google
70.1
$1.50 / $9
28 Gemini 3.6 Flash (high reasoning)Google
69.2
$1.50 / $7.50
29 DeepSeek-V4-Flash (0731, max reasoning)DeepSeek
69.1
$0.14 / $0.28
30 DeepSeek-V4-Pro (0813, max reasoning)DeepSeek
68.8
$1.32 / $3.96
30 Gemini 3.1 Pro PreviewGoogle
68.8
$2 / $12
30 GLM-5.2 (max reasoning)Z.ai · best of 2 settings
68.8
$1.40 / $4.40
33 Qwen3.8-27B (extra-high reasoning)Alibaba · best of 4 settings
68.1
$0.50 / $3
34 Qwen3.7-MaxAlibaba
66.0
$1.48 / $4.42
35 DeepSeek-V4-Flash-Vision-Exp (max reasoning)DeepSeek
65.0
$0.44 / $1.32
36 Claude Sonnet 4.6 (max reasoning)Anthropic
63.0
$3 / $15
37 Kimi K2.6Moonshot AI
61.8
$0.95 / $4
38 Kimi K2.7 CodeMoonshot AI
60.8
$0.95 / $4
39 MiMo-V2.5-Pro (reasoning on)Xiaomi
60.2
$0.43 / $0.87
40 DeepSeek-V4-Pro (0423, max reasoning)DeepSeek · best of 2 settings
59.4
$1.42 / $2.83
41 Hy3Tencent
58.8
$0.14 / $0.58
42 Muse SparkMeta
58.6
–
42 MiniMax-M3MiniMax
58.6
$0.30 / $1.20
44 MiMo-V2.5Xiaomi
56.8
$0.17 / $0.34
45 DeepSeek-V4-Flash (0423, max reasoning)DeepSeek · best of 2 settings
56.2
$0.14 / $0.28
46 GPT-5.4 mini (extra-high reasoning)OpenAI
56.1
$0.75 / $4.50
46 GPT-5.4 nano (extra-high reasoning)OpenAI
56.1
$0.20 / $1.25
48 Qwen3.7-PlusAlibaba
55.9
$0.32 / $1.28
49 GLM-5.1Z.ai
55.8
$1.38 / $4.40
50 Qwen3.6-PlusAlibaba
54.5
$0.33 / $1.95
51 Qwen3.6-27BAlibaba · best of 2 settings
53.7
$0.30 / $3.20
52 Inkling SmallThinking Machines
52.9
$0.45 / $1.20
53 Solar Pro 4Upstage
52.7
$0.09 / $0.36
54 MiniMax-M2.7MiniMax
52.6
$0.30 / $1.20
55 Claude Sonnet 4.5 (reasoning on)Anthropic
52.1
$3 / $15
55 Inkling (extra-high reasoning)Thinking Machines
52.1
$0.95 / $4.05
57 Grok Build 0.1xAI
51.5
$1 / $2
58 GPT-5.1 (high reasoning)OpenAI
49.4
$1.25 / $10
59 Gemini 3.5 Flash-LiteGoogle
49.3
$0.30 / $2.50
60 Muse Glimmer (high reasoning)Meta
49.0
–
61 Qwen3.5-397B-A17BAlibaba
48.2
$0.55 / $3.50
62 Mistral Medium 3.5Mistral AI
46.9
$1.50 / $7.50
63 Kimi K2.5Moonshot AI
46.8
$0.57 / $2.85
64 GLM-4.6 (reasoning on)Z.ai
45.8
$0.50 / $2
65 Qwen3.5-122B-A10BAlibaba · best of 2 settings
45.7
$0.26 / $2.08
66 GLM-4.7Z.ai
45.3
$0.54 / $1.98
67 DeepSeek-V3.2 (reasoning on)DeepSeek
44.2
$0.30 / $0.96
68 Claude Haiku 4.5 (reasoning on)Anthropic
43.9
$1 / $5
69 DeepSeek-V3.1-Terminus (reasoning on)DeepSeek
43.5
$0.27 / $1
70 Gemma 4 31BGoogle · best of 2 settings
43.4
$0.14 / $0.40
71 Grok 4.3 (high reasoning)xAI · best of 2 settings
42.2
$1.25 / $2.50
72 Qwen3.6-35B-A3BAlibaba · best of 2 settings
41.9
$0.10 / $1
73 o1OpenAI
39.7
$15 / $60
74 GPT-5.5 Instant (2026-06-26)OpenAI
39.4
–
75 Gemma 4 26B A4BGoogle
39.3
$0.10 / $0.30
76 GPT-5 (high reasoning)OpenAI
37.8
$1.25 / $10
77 Nemotron 3 Super 120B A12BNVIDIA
37.7
$0.085 / $0.40
78 Claude Sonnet 4 (reasoning on)Anthropic
37.6
$3 / $15
79 Qwen3.5-35B-A3B (no reasoning)Alibaba
37.0
$0.16 / $1.30
80 Qwen3-Coder-NextAlibaba
36.2
$0.18 / $0.90
81 Gemini 3.1 Flash-Lite PreviewGoogle
34.7
$0.25 / $1.50
82 Gemini 2.5 ProGoogle
33.3
$1.25 / $10
83 Devstral 2Mistral AI
31.3
$0.40 / $2
84 Mercury 2Inception
31.1
$0.25 / $0.75
85 Gemma 4 12BGoogle
31.0
–
86 gpt-oss-120b (high reasoning)OpenAI · best of 2 settings
30.4
$0.15 / $0.60
87 Devstral Small 2Mistral AI
29.3
–
88 Qwen3.5-9BAlibaba · best of 2 settings
28.7
$0.10 / $0.15
89 Mistral Small 4Mistral AI
26.6
$0.15 / $0.60
90 Mistral Small 3.1 24BMistral AI
26.3
$0.35 / $0.56
91 Trinity Large ThinkingArcee AI
25.8
$0.25 / $0.80
92 DeepSeek-R1DeepSeek
24.6
$0.70 / $2.50
93 GPT-4o (2024-05-13)OpenAI
24.2
$5 / $15
94 Nova 2 Lite (high reasoning)Amazon
23.0
$0.30 / $2.50
94 DeepSeek-V3DeepSeek
23.0
$0.26 / $1.03
96 Qwen3.5-4B (reasoning on)Alibaba · best of 2 settings
22.6
–
97 Granite 4.2 8BIBM
22.4
$0.06 / $0.25
98 Qwen3-235B-A22B-Thinking-2507Alibaba
22.1
$0.30 / $3
99 Magistral Medium 1.2Mistral AI
21.3
–
100 DeepSeek-V3-0324DeepSeek
21.2
$0.25 / $1
101 gpt-oss-20b (high reasoning)OpenAI
20.7
$0.03 / $0.15
102 Mistral Medium 3.1Mistral AI
20.5
$0.40 / $2
103 GPT-4.1 miniOpenAI
20.2
$0.40 / $1.60
104 Mistral Large 3Mistral AI
20.1
$0.50 / $1.50
105 DiffusionGemma 26B A4BGoogle
19.7
–
106 Qwen3-Next-80B-A3B-ThinkingAlibaba
17.4
$0.15 / $1.20
107 Llama 4 MaverickMeta
16.3
$0.27 / $0.85
107 o3-mini (high reasoning)OpenAI
16.3
$1.10 / $4.40
109 Solar Pro 3Upstage
16.2
$0.15 / $0.60
110 GPT-5 mini (high reasoning)OpenAI
15.6
$0.25 / $2
111 Qwen3-32B (reasoning on)Alibaba
15.3
$0.14 / $0.40
112 Magistral Small 1.2Mistral AI
14.7
–
113 Ministral 3 14BMistral AI
14.4
$0.20 / $0.20
113 Nemotron 3 Nano 30B A3B (reasoning on)NVIDIA
14.4
$0.05 / $0.20
115 Qwen3-14B (reasoning on)Alibaba
13.8
$0.12 / $0.24
116 Mistral Small 3.2 24BMistral AI
12.5
$0.094 / $0.25
117 Qwen3-30B-A3B-Thinking-2507Alibaba
12.1
$0.20 / $2.40
118 Llama 3.3 70B InstructMeta
11.9
$0.59 / $0.79
119 GPT-4o mini (2024-07-18)OpenAI
11.4
$0.15 / $0.60
120 GPT-4.1 nanoOpenAI
11.1
$0.10 / $0.40
121 Gemma 3 27BGoogle
10.1
$0.12 / $0.20
122 Ministral 3 8BMistral AI
9.7
$0.15 / $0.15
123 Qwen3-8B (reasoning on)Alibaba
9.0
$0.12 / $0.46
124 Llama 4 ScoutMeta
8.2
$0.18 / $0.59
125 Gemma 4 E2BGoogle
7.2
–
126 Gemma 3 12BGoogle
5.8
$0.05 / $0.15
127 Llama 3.1 8B InstructMeta
5.4
$0.05 / $0.08
128 Ministral 3 3BMistral AI
4.8
$0.10 / $0.10
129 Qwen3.5-2B (reasoning on)Alibaba · best of 2 settings
2.9
–
130 Gemma 3 4BGoogle
2.7
$0.05 / $0.10
131 Qwen3.5-0.8B (no reasoning)Alibaba · best of 2 settings
1.2
–

Swipe the table sideways for more columns.

Ranks follow the score as shown, so equal numbers share a rank. Each model is shown at its best setting; show every setting. Results as published by Artificial Analysis; we do not re-run them.

What it measures

Artificial Analysis's blend of its coding evaluations.

What it does not measure

Not your codebase or tools.

282 results from Artificial Analysis not ranked here · show why

We rank a result only when we can tie it to a specific model you can use. These are left out:

  • Not on sale through the API providers we track: 268
  • A different snapshot or variant from the model we list: 13
  • An unusual combination of settings: 1

Data sourced from Artificial Analysis. Licence: Artificial Analysis commercial data licence.