Back to AI Coding

AI Model Ranking

AI Coding Model Rankings

Snapshot: 2026-07-17 · Last reviewed: 2026-07-19

Methodology

Metrics follow a snapshot of the Artificial Analysis public LLM leaderboard: Intelligence Index, pricing, and speed reflect benchmark performance at the snapshot date, not a hands-on coding review by DevCove. Provider pricing, availability, and rankings change frequently, and a high benchmark score does not guarantee the best result for a specific codebase or workflow.

Source
Artificial Analysis
Last reviewed
2026-07-19

Top 20 models

01

Current reasoning model

Anthropic - Claude Fable - 5

Still the highest Intelligence Index in the snapshot, with a 1M context window.

Best for: Deep repository analysis, complex planning, and high-value agent tasks where latency is acceptable.

Context
1M
AA Index
60
Blended price
$7.70/1M
Speed
66 tok/s
First chunk
148.11s
Total response
155.74s
02

Current reasoning model

OpenAI - GPT Sol - 5.6 (max)

New GPT-5.6 Sol tier at the top of OpenAI's reasoning lineup in this snapshot.

Best for: OpenAI-centered coding agents, long-context planning, and max-reasoning software tasks.

Context
1M
AA Index
59
Blended price
$4.35/1M
Speed
52 tok/s
First chunk
138.64s
Total response
148.19s
03

Current reasoning model

OpenAI - GPT Sol - 5.6 (xhigh)

High-reasoning GPT-5.6 Sol mode with 1M context and much lower latency than max.

Best for: Daily agent coding, code review, and tool-heavy workflows that still need strong reasoning.

Context
1M
AA Index
58
Blended price
$4.35/1M
Speed
52 tok/s
First chunk
41.62s
Total response
51.23s
04

Current reasoning model

Kimi - Kimi - K3

Moonshot's new 2.8T MoE flagship with 1M+ context, strong coding scores, and low first-chunk latency.

Best for: Long-horizon coding agents, frontend generation, regional API stacks, and cost-aware 1M context work.

Context
1.05M
AA Index
57
Blended price
$2.31/1M
Speed
62 tok/s
First chunk
1.99s
Total response
42.30s
05

Current reasoning model

OpenAI - GPT Sol - 5.6 (high)

Balanced GPT-5.6 Sol high tier with 1M context and responsive total response time.

Best for: Interactive coding sessions, refactors, and product engineering loops inside OpenAI tools.

Context
1M
AA Index
56
Blended price
$4.35/1M
Speed
47 tok/s
First chunk
10.17s
Total response
20.74s
06

Current reasoning model

Anthropic - Claude Opus - 4.8 (max)

High-intelligence Claude Opus 4.8 with 1M context and competitive end-to-end latency.

Best for: Large code migrations, multi-file edits, and coding agents that need strong reasoning.

Context
1M
AA Index
56
Blended price
$3.85/1M
Speed
56 tok/s
First chunk
34.49s
Total response
43.42s
07

Current reasoning model

OpenAI - GPT Terra - 5.6 (max)

New GPT-5.6 Terra max tier with lower listed price than Sol and very fast output speed.

Best for: Cost-aware long-context coding where throughput matters more than first-token speed.

Context
1M
AA Index
55
Blended price
$2.17/1M
Speed
138 tok/s
First chunk
138.25s
Total response
141.88s
08

Current reasoning model

OpenAI - GPT - 5.5 (xhigh)

Prior-generation GPT-5.5 xhigh still ranks high, with broad context and strong reasoning.

Best for: Teams still routed on GPT-5.5 endpoints for design-to-code and code review workflows.

Context
922k
AA Index
55
Blended price
$4.35/1M
Speed
67 tok/s
First chunk
87.74s
Total response
95.20s
09

Current reasoning model

SpaceXAI - Grok - 4.5 (high)

First Grok entry in the Top 20, with fast latency and a mid-size 500k context window.

Best for: Fast interactive coding help, medium-context agent loops, and xAI ecosystem experiments.

Context
500k
AA Index
54
Blended price
$1.35/1M
Speed
103 tok/s
First chunk
9.12s
Total response
13.98s
10

Current reasoning model

OpenAI - GPT Sol - 5.6 (medium)

GPT-5.6 Sol medium tier balances intelligence, 1M context, and low total response time.

Best for: Most daily OpenAI coding loops where speed and quality both matter.

Context
1M
AA Index
54
Blended price
$4.35/1M
Speed
52 tok/s
First chunk
5.15s
Total response
14.86s
11

Current reasoning model

Anthropic - Claude Opus - 4.7 (max)

Strong Opus reasoning with 1M context and lower latency than higher Opus/Fable entries.

Best for: Interactive coding sessions that still need frontier reasoning and long context.

Context
1M
AA Index
54
Blended price
$3.85/1M
Speed
50 tok/s
First chunk
19.68s
Total response
29.69s
12

Current reasoning model

Anthropic - Claude Sonnet - 5 (max)

Sonnet-level cost profile with high intelligence and 1M context, trading off long first response.

Best for: Long, careful batch work where quality and context matter more than immediacy.

Context
1M
AA Index
53
Blended price
$1.54/1M
Speed
78 tok/s
First chunk
192.74s
Total response
199.16s
13

Current reasoning model

OpenAI - GPT - 5.5 (high)

A high-reasoning GPT-5.5 mode with the same broad context and much lower latency than xhigh.

Best for: Daily agent coding, code review, refactors, and product engineering loops.

Context
922k
AA Index
53
Blended price
$4.35/1M
Speed
68 tok/s
First chunk
17.77s
Total response
25.11s
14

Current reasoning model

OpenAI - GPT Terra - 5.6 (xhigh)

GPT-5.6 Terra xhigh combines lower listed price, 1M context, and fast median output.

Best for: Cost-sensitive long-context coding with strong throughput and reasonable first-chunk latency.

Context
1M
AA Index
52
Blended price
$2.17/1M
Speed
123 tok/s
First chunk
16.93s
Total response
21.00s
15

Current reasoning model

OpenAI - GPT Luna - 5.6 (max)

New GPT-5.6 Luna max tier with very low listed price, 1M context, and high output speed.

Best for: Budget OpenAI stacks, long-context batch edits, and high-throughput coding assistants.

Context
1M
AA Index
51
Blended price
$0.87/1M
Speed
210 tok/s
First chunk
95.08s
Total response
97.46s
16

Current reasoning model

Z AI - GLM - 5.2 (max)

High ranking with 1M context, low listed blended price, and very fast median output.

Best for: Cost-aware coding assistants, long-context analysis, and latency-sensitive workflows.

Context
1M
AA Index
51
Blended price
$0.90/1M
Speed
156 tok/s
First chunk
1.54s
Total response
17.56s
17

Current reasoning model

Meta - Muse Spark - 1.1 (xhigh)

Updated Muse Spark 1.1 xhigh with full public metrics and a slightly larger context window.

Best for: Low-latency long-context coding experiments and Meta model routing comparisons.

Context
1.05M
AA Index
51
Blended price
$0.78/1M
Speed
117 tok/s
First chunk
1.32s
Total response
22.66s
18

Current reasoning model

OpenAI - GPT - 5.5 (medium)

Balanced GPT-5.5 reasoning tier with broad context and faster response than high/xhigh.

Best for: Most interactive coding loops where quality, speed, and context all matter.

Context
922k
AA Index
50
Blended price
$4.35/1M
Speed
64 tok/s
First chunk
7.40s
Total response
15.27s
19

Current general model

Google - Gemini - 3.5 Flash

Fast long-context Gemini model with strong intelligence and Google ecosystem fit.

Best for: Long-file analysis, multimodal coding workflows, and Google Cloud-oriented stacks.

Context
1M
AA Index
50
Blended price
$1.31/1M
Speed
157 tok/s
First chunk
30.23s
Total response
33.42s
20

Current reasoning model

OpenAI - GPT Sol - 5.6 (low)

Low-latency GPT-5.6 Sol tier with 1M context and the fastest total response in the Sol line.

Best for: Interactive pair programming, quick fixes, and fast planning loops on GPT-5.6 Sol.

Context
1M
AA Index
49
Blended price
$4.35/1M
Speed
50 tok/s
First chunk
3.50s
Total response
13.44s