Back to AI Coding

AI Model Ranking

AI Coding Model Rankings

Top 50 from Artificial Analysis.

Snapshot: 2026-10-03 · Last reviewed: 2026-10-03

Methodology

Metrics follow a snapshot of the Artificial Analysis public LLM leaderboard: Intelligence Index, cost per task, and speed reflect benchmark performance at the snapshot date. Provider pricing, availability, and rankings change frequently, and a high benchmark score does not guarantee the best result for a specific codebase or workflow.

Source
Artificial Analysis
Last reviewed
2026-10-03

Top 50 models

01

Current reasoning model

Anthropic - Claude Opus - 5.5 (max)

Claude Opus 5.5 max leads this snapshot at Intelligence Index 58 with 1M context and cost per task $5.98. Median speed is 93 tok/s, and first chunk is 689.18s.

Best for: Highest-effort Claude Opus work when the top Intelligence Index matters more than a short first chunk.

Anthropic

Context
1M
AA Index
58
Cost per task
$5.98
Speed
93 tok/s
First chunk
689.18s
Total response
694.56s
02

Current reasoning model

Anthropic - Claude Sonnet - 5.5 (max)

Claude Sonnet 5.5 max is #2 at Intelligence Index 56 with 1M context. Cost per task is $7.67, the highest in the top 10, and first chunk is 428.36s.

Best for: Highest-effort Claude Sonnet work when Intelligence Index matters more than task cost and wait time.

Anthropic

Context
1M
AA Index
56
Cost per task
$7.67
Speed
139 tok/s
First chunk
428.36s
Total response
431.95s
03

Current reasoning model

Anthropic - Claude Opus - 5.5 (xhigh)

Claude Opus 5.5 xhigh ties Sonnet 5.5 max at Intelligence Index 56 with 1M context and cost per task $3.46. First chunk is 140.81s, shorter than Sonnet 5.5 max.

Best for: High-effort Opus 5.5 coding when you want the same Intelligence Index as Sonnet 5.5 max at a lower task cost.

Anthropic

Context
1M
AA Index
56
Cost per task
$3.46
Speed
83 tok/s
First chunk
140.81s
Total response
146.82s
04

Current reasoning model

Anthropic - Claude Opus - 5.5 (high)

Claude Opus 5.5 high is #4 at Intelligence Index 54 with 1M context. Cost per task is $1.82 and first chunk is 37.37s, far below Opus 5.5 xhigh.

Best for: Claude Opus 5.5 coding that still needs a high Intelligence Index without the xhigh first-chunk wait.

Anthropic

Context
1M
AA Index
54
Cost per task
$1.82
Speed
80 tok/s
First chunk
37.37s
Total response
43.64s
05

Current reasoning model

Anthropic - Claude Fable - 5.1 (max)

Claude Fable 5.1 max is #5 at Intelligence Index 53 with 1M context. Cost per task is $7.63, just under Sonnet 5.5 max, and first chunk is 186.52s.

Best for: Highest-effort Claude Fable work when Intelligence Index matters more than task cost and wait time.

Anthropic

Context
1M
AA Index
53
Cost per task
$7.63
Speed
67 tok/s
First chunk
186.52s
Total response
193.93s
06

Current reasoning model

Anthropic - Claude Fable - 5.1 (xhigh)

Claude Fable 5.1 xhigh stays at Intelligence Index 53, tied with Fable 5.1 max, with 1M context and cost per task $5.98. First chunk is 96.75s, shorter than Fable 5.1 max.

Best for: High-effort Fable 5.1 coding when max task cost is too high and you still want the same Intelligence Index.

Anthropic

Context
1M
AA Index
53
Cost per task
$5.98
Speed
66 tok/s
First chunk
96.75s
Total response
104.30s
07

Current reasoning model

OpenAI - GPT Astra - 6 (max)

GPT-6 Astra max is #7 at Intelligence Index 53 with 1M context. Cost per task is $3.26, below Fable 5.1 max, and first chunk is 319.00s.

Best for: OpenAI-centered frontier agents that can absorb a very long first-chunk delay.

OpenAI

Context
1M
AA Index
53
Cost per task
$3.26
Speed
54 tok/s
First chunk
319.00s
Total response
328.32s
08

Current model; partial public metrics

Google - Gemini - 4 Argon (high)

Gemini 4 Argon high is #8 at Intelligence Index 53 with 1M context and cost per task $1.99, tied with several models at 53. Output speed and latency are not published yet.

Best for: Google frontier coding when Intelligence Index matters more than published speed.

Google

Context
1M
AA Index
53
Cost per task
$1.99
Speed
— tok/s
First chunk
—
Total response
—
09

Current reasoning model

OpenAI - GPT Astra - 6 (xhigh)

GPT-6 Astra xhigh is #9 at Intelligence Index 52, 1M context, and cost per task $2.31. First chunk is 144.37s, shorter than Astra max.

Best for: OpenAI coding agents that need near-max Astra quality at a lower task cost than max.

OpenAI

Context
1M
AA Index
52
Cost per task
$2.31
Speed
49 tok/s
First chunk
144.37s
Total response
154.50s
10

Current reasoning model

Anthropic - Claude Sonnet - 5.5 (xhigh)

Claude Sonnet 5.5 xhigh is #10 at Intelligence Index 52 with 1M context. Cost per task is $2.75, median speed is 104 tok/s, and first chunk is 33.63s, far below Sonnet 5.5 max.

Best for: Claude Sonnet 5.5 coding that still needs a high Intelligence Index without the max first-chunk wait.

Anthropic

Context
1M
AA Index
52
Cost per task
$2.75
Speed
104 tok/s
First chunk
33.63s
Total response
38.45s
11

Current reasoning model

OpenAI - GPT Sol - 6.1 (max)

OpenAI GPT Sol 6.1 (max) ranks in this snapshot with Intelligence Index 52, 1M context, cost per task $0.72, median 57 tok/s, and first chunk 282.35s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
52
Cost per task
$0.72
Speed
57 tok/s
First chunk
282.35s
Total response
291.13s
12

Current reasoning model

Anthropic - Claude Opus - 5.5 (medium)

Anthropic Claude Opus 5.5 (medium) ranks in this snapshot with Intelligence Index 51, 1M context, cost per task $1.34, median 73 tok/s, and first chunk 25.24s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
51
Cost per task
$1.34
Speed
73 tok/s
First chunk
25.24s
Total response
32.06s
13

Current reasoning model

Anthropic - Claude Fable - 5.1 (high)

Anthropic Claude Fable 5.1 (high) ranks in this snapshot with Intelligence Index 51, 1M context, cost per task $3.91, median 56 tok/s, and first chunk 25.00s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
51
Cost per task
$3.91
Speed
56 tok/s
First chunk
25.00s
Total response
33.93s
14

Current reasoning model

OpenAI - GPT Sol - 6.1 (xhigh)

OpenAI GPT Sol 6.1 (xhigh) ranks in this snapshot with Intelligence Index 51, 1M context, cost per task $0.39, median 51 tok/s, and first chunk 128.47s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
51
Cost per task
$0.39
Speed
51 tok/s
First chunk
128.47s
Total response
138.37s
15

Current reasoning model

OpenAI - GPT Astra - 6 (high)

OpenAI GPT Astra 6 (high) ranks in this snapshot with Intelligence Index 51, 1M context, cost per task $1.73, median 48 tok/s, and first chunk 65.56s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
51
Cost per task
$1.73
Speed
48 tok/s
First chunk
65.56s
Total response
76.04s
16

Current reasoning model

OpenAI - GPT Sol - 6.1 (high)

OpenAI GPT Sol 6.1 (high) ranks in this snapshot with Intelligence Index 50, 1M context, cost per task $0.32, median 50 tok/s, and first chunk 57.71s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
50
Cost per task
$0.32
Speed
50 tok/s
First chunk
57.71s
Total response
67.64s
17

Current reasoning model

OpenAI - GPT Astra - 6 (medium)

OpenAI GPT Astra 6 (medium) ranks in this snapshot with Intelligence Index 50, 1M context, cost per task $1.54, median 46 tok/s, and first chunk 4.70s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
50
Cost per task
$1.54
Speed
46 tok/s
First chunk
4.70s
Total response
15.65s
18

Current reasoning model

Anthropic - Claude Fable - 5.1 (medium)

Anthropic Claude Fable 5.1 (medium) ranks in this snapshot with Intelligence Index 49, 1M context, cost per task $2.98, median 57 tok/s, and first chunk 9.11s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
49
Cost per task
$2.98
Speed
57 tok/s
First chunk
9.11s
Total response
17.85s
19

Current reasoning model

Meta - Muse Spark - 1.3 (max)

Meta Muse Spark 1.3 (max) ranks in this snapshot with Intelligence Index 48, 1M context, cost per task $1.60, median 135 tok/s, and first chunk 52.46s.

Best for: Meta coding agents, long-context work, and snapshot-based model comparison.

Meta

Context
1M
AA Index
48
Cost per task
$1.60
Speed
135 tok/s
First chunk
52.46s
Total response
71.04s
20

Current reasoning model

OpenAI - GPT Sol - 6.1 (medium)

OpenAI GPT Sol 6.1 (medium) ranks in this snapshot with Intelligence Index 48, 1M context, cost per task $0.21, median 45 tok/s, and first chunk 5.29s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
48
Cost per task
$0.21
Speed
45 tok/s
First chunk
5.29s
Total response
16.30s
21

Current reasoning model

Anthropic - Claude Fable - 5.1 (low)

Anthropic Claude Fable 5.1 (low) ranks in this snapshot with Intelligence Index 47, 1M context, cost per task $2.37, median 55 tok/s, and first chunk 4.73s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
47
Cost per task
$2.37
Speed
55 tok/s
First chunk
4.73s
Total response
13.80s
22

Current reasoning model

Anthropic - Claude Sonnet - 5.5 (high)

Anthropic Claude Sonnet 5.5 (high) ranks in this snapshot with Intelligence Index 47, 1M context, cost per task $1.12, median 102 tok/s, and first chunk 12.19s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
47
Cost per task
$1.12
Speed
102 tok/s
First chunk
12.19s
Total response
17.09s
23

Current reasoning model

SpaceXAI - Grok - 4.7 (xhigh)

SpaceXAI Grok 4.7 (xhigh) ranks in this snapshot with Intelligence Index 46, 500k context, cost per task $3.74, median 79 tok/s, and first chunk 41.26s.

Best for: SpaceXAI coding agents, long-context work, and snapshot-based model comparison.

SpaceXAI

Context
500k
AA Index
46
Cost per task
$3.74
Speed
79 tok/s
First chunk
41.26s
Total response
47.59s
24

Current reasoning model

SpaceXAI - Grok - 4.7 (high)

SpaceXAI Grok 4.7 (high) ranks in this snapshot with Intelligence Index 46, 500k context, cost per task $2.73, median 78 tok/s, and first chunk 33.12s.

Best for: SpaceXAI coding agents, long-context work, and snapshot-based model comparison.

SpaceXAI

Context
500k
AA Index
46
Cost per task
$2.73
Speed
78 tok/s
First chunk
33.12s
Total response
39.49s
25

Current reasoning model

Xiaomi - MiMo - V2.6 Pro

Xiaomi MiMo V2.6 Pro ranks in this snapshot with Intelligence Index 46, 1M context, cost per task $0.13, median 45 tok/s, and first chunk 3.82s.

Best for: Xiaomi coding agents, long-context work, and snapshot-based model comparison.

Xiaomi

Context
1M
AA Index
46
Cost per task
$0.13
Speed
45 tok/s
First chunk
3.82s
Total response
59.44s
26

Current reasoning model

OpenAI - GPT Astra - 6 (low)

OpenAI GPT Astra 6 (low) ranks in this snapshot with Intelligence Index 46, 1M context, cost per task $0.82, median 47 tok/s, and first chunk 2.75s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
46
Cost per task
$0.82
Speed
47 tok/s
First chunk
2.75s
Total response
13.34s
27

Current reasoning model

Alibaba - Qwen - 3.8 Max (0902)

Alibaba Qwen 3.8 Max (0902) ranks in this snapshot with Intelligence Index 45, 984k context, cost per task $5.41, median 38 tok/s, and first chunk 2.67s.

Best for: Alibaba coding agents, long-context work, and snapshot-based model comparison.

Alibaba

Context
984k
AA Index
45
Cost per task
$5.41
Speed
38 tok/s
First chunk
2.67s
Total response
68.62s
28

Current reasoning model

Meta - Muse Spark - 1.3 (xhigh)

Meta Muse Spark 1.3 (xhigh) ranks in this snapshot with Intelligence Index 45, 1M context, cost per task $1.37, median 146 tok/s, and first chunk 41.50s.

Best for: Meta coding agents, long-context work, and snapshot-based model comparison.

Meta

Context
1M
AA Index
45
Cost per task
$1.37
Speed
146 tok/s
First chunk
41.50s
Total response
58.68s
29

Current reasoning model

Z AI - GLM - 5.3 (max)

Z AI GLM 5.3 (max) ranks in this snapshot with Intelligence Index 45, 1M context, cost per task $2.01, median 74 tok/s, and first chunk 3.18s.

Best for: Z AI coding agents, long-context work, and snapshot-based model comparison.

Z AI

Context
1M
AA Index
45
Cost per task
$2.01
Speed
74 tok/s
First chunk
3.18s
Total response
36.90s
30

Current reasoning model

StepFun - Step - 5 Preview

StepFun Step 5 Preview ranks in this snapshot with Intelligence Index 44, 1M context, cost per task $0.72, median 85 tok/s, and first chunk 3.03s.

Best for: StepFun coding agents, long-context work, and snapshot-based model comparison.

StepFun

Context
1M
AA Index
44
Cost per task
$0.72
Speed
85 tok/s
First chunk
3.03s
Total response
32.40s
31

Current reasoning model

Kimi - Kimi - K3 (max)

Kimi Kimi K3 (max) ranks in this snapshot with Intelligence Index 44, 1.05M context, cost per task $2.00, median 36 tok/s, and first chunk 4.12s.

Best for: Kimi coding agents, long-context work, and snapshot-based model comparison.

Kimi

Context
1.05M
AA Index
44
Cost per task
$2.00
Speed
36 tok/s
First chunk
4.12s
Total response
72.72s
32

Current reasoning model

Anthropic - Claude Opus - 5.5 (low)

Anthropic Claude Opus 5.5 (low) ranks in this snapshot with Intelligence Index 42, 1M context, cost per task $0.55, median 80 tok/s, and first chunk 9.94s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
42
Cost per task
$0.55
Speed
80 tok/s
First chunk
9.94s
Total response
16.18s
33

Current reasoning model

SpaceXAI - Grok - 4.7 (low)

SpaceXAI Grok 4.7 (low) ranks in this snapshot with Intelligence Index 42, 500k context, cost per task $1.25, median 75 tok/s, and first chunk 8.62s.

Best for: SpaceXAI coding agents, long-context work, and snapshot-based model comparison.

SpaceXAI

Context
500k
AA Index
42
Cost per task
$1.25
Speed
75 tok/s
First chunk
8.62s
Total response
15.25s
34

Current reasoning model

OpenAI - GPT Sol - 6.1 (low)

OpenAI GPT Sol 6.1 (low) ranks in this snapshot with Intelligence Index 42, 1M context, cost per task $0.13, median 48 tok/s, and first chunk 3.37s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
42
Cost per task
$0.13
Speed
48 tok/s
First chunk
3.37s
Total response
13.78s
35

Current reasoning model

OpenAI - GPT Terra - 5.6 (max)

OpenAI GPT Terra 5.6 (max) ranks in this snapshot with Intelligence Index 42, 1M context, cost per task $1.40, median 102 tok/s, and first chunk 172.37s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
42
Cost per task
$1.40
Speed
102 tok/s
First chunk
172.37s
Total response
177.28s
36

Current reasoning model

Z AI - GLM - 5.3 Flash

Z AI GLM 5.3 Flash ranks in this snapshot with Intelligence Index 42, 1M context, cost per task $0.25, median 53 tok/s, and first chunk 3.29s.

Best for: Z AI coding agents, long-context work, and snapshot-based model comparison.

Z AI

Context
1M
AA Index
42
Cost per task
$0.25
Speed
53 tok/s
First chunk
3.29s
Total response
50.21s
37

Current reasoning model

Google - Gemini - 3.8 Flash (high)

Google Gemini 3.8 Flash (high) ranks in this snapshot with Intelligence Index 41, 1M context, cost per task $1.24, median 242 tok/s, and first chunk 16.13s.

Best for: Google coding agents, long-context work, and snapshot-based model comparison.

Google

Context
1M
AA Index
41
Cost per task
$1.24
Speed
242 tok/s
First chunk
16.13s
Total response
18.19s
38

Current reasoning model

Anthropic - Claude Sonnet - 5.5 (medium)

Anthropic Claude Sonnet 5.5 (medium) ranks in this snapshot with Intelligence Index 41, 1M context, cost per task $0.59, median 107 tok/s, and first chunk 1.21s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
41
Cost per task
$0.59
Speed
107 tok/s
First chunk
1.21s
Total response
5.88s
39

Current reasoning model

Alibaba - Qwen - 3.8 2.4T A95B

Alibaba Qwen 3.8 2.4T A95B ranks in this snapshot with Intelligence Index 40, 262k context, cost per task $2.16, median 39 tok/s, and first chunk 2.78s.

Best for: Alibaba coding agents, long-context work, and snapshot-based model comparison.

Alibaba

Context
262k
AA Index
40
Cost per task
$2.16
Speed
39 tok/s
First chunk
2.78s
Total response
66.94s
40

Current reasoning model

Alibaba - Qwen - 3.8 Flash Next

Alibaba Qwen 3.8 Flash Next ranks in this snapshot with Intelligence Index 40, 256k context, cost per task $0.37, median 54 tok/s, and first chunk 2.45s.

Best for: Alibaba coding agents, long-context work, and snapshot-based model comparison.

Alibaba

Context
256k
AA Index
40
Cost per task
$0.37
Speed
54 tok/s
First chunk
2.45s
Total response
48.98s
41

Current model; partial public metrics

Google - Gemini - 3.8 Flash (medium)

Google Gemini 3.8 Flash (medium) ranks in this snapshot with Intelligence Index 40, 1M context, and cost per task $0.93. Output speed and latency are not published yet.

Best for: Google coding agents, long-context work, and snapshot-based model comparison.

Google

Context
1M
AA Index
40
Cost per task
$0.93
Speed
— tok/s
First chunk
—
Total response
—
42

Current reasoning model

DeepSeek - DeepSeek - V4.1 Flash (max)

DeepSeek DeepSeek V4.1 Flash (max) ranks in this snapshot with Intelligence Index 39, 1M context, cost per task $0.27, median 224 tok/s, and first chunk 1.02s.

Best for: DeepSeek coding agents, long-context work, and snapshot-based model comparison.

DeepSeek

Context
1M
AA Index
39
Cost per task
$0.27
Speed
224 tok/s
First chunk
1.02s
Total response
12.21s
43

Current reasoning model

OpenAI - GPT Luna - 6 (max)

OpenAI GPT Luna 6 (max) ranks in this snapshot with Intelligence Index 38, 1M context, cost per task $0.07, median 129 tok/s, and first chunk 132.77s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
38
Cost per task
$0.07
Speed
129 tok/s
First chunk
132.77s
Total response
136.65s
44

Current reasoning model

OpenAI - GPT Terra - 5.6 (xhigh)

OpenAI GPT Terra 5.6 (xhigh) ranks in this snapshot with Intelligence Index 38, 1M context, cost per task $0.63, median 91 tok/s, and first chunk 28.50s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
38
Cost per task
$0.63
Speed
91 tok/s
First chunk
28.50s
Total response
33.99s
45

Current reasoning model

Xiaomi - MiMo - V2.6 Flash

Xiaomi MiMo V2.6 Flash ranks in this snapshot with Intelligence Index 38, 1M context, cost per task $0.06, median 50 tok/s, and first chunk 4.75s.

Best for: Xiaomi coding agents, long-context work, and snapshot-based model comparison.

Xiaomi

Context
1M
AA Index
38
Cost per task
$0.06
Speed
50 tok/s
First chunk
4.75s
Total response
54.95s
46

Current reasoning model

DeepSeek - DeepSeek - V4 Pro 0813 (max)

DeepSeek DeepSeek V4 Pro 0813 (max) ranks in this snapshot with Intelligence Index 36, 1M context, cost per task $0.67, median 109 tok/s, and first chunk 1.73s.

Best for: DeepSeek coding agents, long-context work, and snapshot-based model comparison.

DeepSeek

Context
1M
AA Index
36
Cost per task
$0.67
Speed
109 tok/s
First chunk
1.73s
Total response
24.63s
47

Current reasoning model

Anthropic - Claude Sonnet - 5.5 (low)

Anthropic Claude Sonnet 5.5 (low) ranks in this snapshot with Intelligence Index 36, 1M context, cost per task $0.42, median 100 tok/s, and first chunk 1.35s.

Best for: Anthropic coding agents, long-context work, and snapshot-based model comparison.

Anthropic

Context
1M
AA Index
36
Cost per task
$0.42
Speed
100 tok/s
First chunk
1.35s
Total response
6.33s
48

Current reasoning model

DeepSeek - DeepSeek - V4 Flash Vision (max)

DeepSeek DeepSeek V4 Flash Vision (max) ranks in this snapshot with Intelligence Index 35, 1M context, cost per task $0.31, median 219 tok/s, and first chunk 0.94s.

Best for: DeepSeek coding agents, long-context work, and snapshot-based model comparison.

DeepSeek

Context
1M
AA Index
35
Cost per task
$0.31
Speed
219 tok/s
First chunk
0.94s
Total response
12.36s
49

Current reasoning model

OpenAI - GPT Luna - 6 (xhigh)

OpenAI GPT Luna 6 (xhigh) ranks in this snapshot with Intelligence Index 35, 1M context, cost per task $0.04, median 138 tok/s, and first chunk 20.86s.

Best for: OpenAI coding agents, long-context work, and snapshot-based model comparison.

OpenAI

Context
1M
AA Index
35
Cost per task
$0.04
Speed
138 tok/s
First chunk
20.86s
Total response
24.48s
50

Current reasoning model

Z AI - GLM - 5.3 (low)

Z AI GLM 5.3 (low) ranks in this snapshot with Intelligence Index 34, 1M context, cost per task $0.85, median 69 tok/s, and first chunk 3.13s.

Best for: Z AI coding agents, long-context work, and snapshot-based model comparison.

Z AI

Context
1M
AA Index
34
Cost per task
$0.85
Speed
69 tok/s
First chunk
3.13s
Total response
39.61s