Best AI for Code in 2026Updated by AA Coding Index

Which AI codes best in 2026? Ranking of 559 models using AA Coding Index as the primary signal, with fallback to LiveCodeBench and SciCode. Built to favor current coding models instead of stale one-off spikes.

Synced: August 31, 2026 β€’ 559 models with coding benchmarks

Use Cases

Code Autocomplete

Inline suggestions as you type. Ideal for IDEs like Cursor and VS Code.

Top models: GPT-5.6 Sol (xhigh), OpenAI: GPT-5.6 Sol (batch), Claude Opus 5

Code Generation

Create functions, classes and full projects from natural language descriptions.

Top models: GPT-5.6 Sol (xhigh), OpenAI: GPT-5.6 Sol (batch), Claude Opus 5

Debug & Code Review

Identify bugs, suggest fixes and review pull requests automatically.

Top models: GPT-5.6 Sol (xhigh), OpenAI: GPT-5.6 Sol (batch), Claude Opus 5

Coding Ranking β€” Top Models

#ModelCompanyCoding ScoreBenchmarkContextInput PriceOpen Source
πŸ₯‡GPT-5.6 Sol (xhigh)OpenAIOpenAI
78.3
AA Coding Indexβ€”$4.00β€”
πŸ₯ˆOpenAI: GPT-5.6 Sol (batch)OpenAIOpenAI
78.3
AA Coding Index1.1M tokens$1.00β€”
πŸ₯‰Claude Opus 5AnthropicAnthropic
78.0
AA Coding Index1.0M tokens$5.00β€”
4GPT-5.6 Sol (max)OpenAIOpenAI
77.4
AA Coding Index1.1M tokens$4.00β€”
5GPT-5.6 Sol (high)OpenAIOpenAI
77.2
AA Coding Indexβ€”$4.00β€”
6SpaceXAI: Grok 4.6xAIxAI
76.8
AA Coding Index500K tokens$2.00β€”
7GPT-5.6 Terra (max)OpenAIOpenAI
76.7
AA Coding Index1.1M tokens$2.00β€”
8Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)AnthropicAnthropic
76.5
AA Coding Index1.0M tokens$10.00β€”
9Anthropic: Claude Fable 5 (batch)AnthropicAnthropic
76.5
AA Coding Index1.0M tokens$5.00β€”
10GPT-5.6 Sol (medium)OpenAIOpenAI
76.3
AA Coding Indexβ€”$4.00β€”
11Kimi K3Moonshot AIMoonshot AI
76.2
AA Coding Index1.0M tokens$3.00β€”
12Gemini 3.7 Flash (high)GoogleGoogle
76.1
AA Coding Index1.0M tokens$0.75β€”
13GPT-5.5OpenAIOpenAI
74.9
AA Coding Index1.1M tokens$5.00β€”
14GLM-5.3Z.aiZ.ai
74.8
AA Coding Indexβ€”$1.40β€”
15Claude Opus 4.8 (Adaptive Reasoning, Max Effort)AnthropicAnthropic
74.3
AA Coding Index1.0M tokens$5.00β€”
16Claude Opus 5 (batch)AnthropicAnthropic
74.3
AA Coding Index1.0M tokens$2.50β€”
17Claude Opus 5 (Fast)AnthropicAnthropic
74.3
AA Coding Index1.0M tokens$10.00β€”
18Anthropic: Claude Opus 4.8 (batch)AnthropicAnthropic
74.3
AA Coding Index1.0M tokens$2.50β€”
19Claude Opus 4.7AnthropicAnthropic
73.6
AA Coding Index1.0M tokens$5.00β€”
20Anthropic: Claude Opus 4.7 (batch)AnthropicAnthropic
73.6
AA Coding Index1.0M tokens$2.50β€”
21Qwen3.8-Flash-NextAlibabaAlibaba
73.1
AA Coding Indexβ€”$0.15β€”
22Grok 4.5SpaceXAISpaceXAI
72.4
AA Coding Indexβ€”$2.00β€”
23SpaceXAI: Grok 4.5xAIxAI
72.4
AA Coding Index500K tokens$2.00β€”
24Muse Spark 1.2 (xhigh)MetaMeta
72.2
AA Coding Indexβ€”$1.25β€”
25Qwen: Qwen3.8 2.4T A95BAlibabaAlibaba
71.9
AA Coding Index1.0M tokens$2.00βœ…
26Qwen: Qwen3.8 MaxAlibabaAlibaba
71.8
AA Coding Index1.0M tokens$2.00βœ…
27Claude Sonnet 5AnthropicAnthropic
71.5
AA Coding Index1.0M tokens$2.00β€”
28GLM-5.3-FlashZ AI
71.5
AA Coding Indexβ€”$0.15β€”
29Gemini 3.7 Flash (medium)GoogleGoogle
71.5
AA Coding Indexβ€”$0.75β€”
30Google: Gemini 3.7 Flash (batch)GoogleGoogle
71.5
AA Coding Index1.0M tokens$0.19β€”
31GPT-5.6 Luna (max)OpenAIOpenAI
71.4
AA Coding Index1.1M tokens$0.20β€”
32Muse Spark 1.1 (xhigh)MetaMeta
71.3
AA Coding Indexβ€”$1.25β€”
33Muse Spark 1.1MetaMeta
71.3
AA Coding Indexβ€”$1.25β€”
34GPT-5.4OpenAIOpenAI
71.1
AA Coding Index1.1M tokens$2.50β€”
35GPT-5.4 ProOpenAIOpenAI
71.1
AA Coding Index1.1M tokens$30.00β€”
36OpenAI: GPT-5.4 (batch)OpenAIOpenAI
71.1
AA Coding Index1.1M tokens$1.25β€”
37Gemini 3.7 Flash (low)GoogleGoogle
71.0
AA Coding Indexβ€”$0.75β€”
38GPT-5.6 Terra (xhigh)OpenAIOpenAI
70.6
AA Coding Indexβ€”$2.00β€”
39Google: Gemini 3.5 FlashGoogleGoogle
70.1
AA Coding Index1.0M tokens$1.50β€”
40Gemini 3.5 FlashGoogleGoogle
70.1
AA Coding Indexβ€”$1.50β€”
41GPT-5.6 Sol (low)OpenAIOpenAI
69.7
AA Coding Indexβ€”$4.00β€”
42Gemini 3.6 Flash (high)GoogleGoogle
69.2
AA Coding Index1.0M tokens$0.75β€”
43Gemini 3.6 FlashGoogleGoogle
69.2
AA Coding Indexβ€”$1.50β€”
44DeepSeek V4 FlashDeepSeekDeepSeek
69.1
AA Coding Index1.0M tokens$0.44βœ…
45Flash 0731GoogleGoogle
69.1
AA Coding Indexβ€”β€”β€”
46Gemini 3.1 Pro PreviewGoogleGoogle
68.8
AA Coding Index1.0M tokens$2.00β€”
47GLM-5.2 (max)Z.aiZ.ai
68.8
AA Coding Indexβ€”$1.40β€”
48DeepSeek V4 ProDeepSeekDeepSeek
68.8
AA Coding Index1.0M tokens$1.32βœ…
49Gemini 3.1GoogleGoogle
68.8
AA Coding Indexβ€”β€”β€”
50GPT-5.6 Luna (xhigh)OpenAIOpenAI
68.6
AA Coding Indexβ€”$0.20β€”
51Qwen: Qwen3.8 27BAlibabaAlibaba
68.1
AA Coding Index1.0M tokens$0.50βœ…
52GPT-5.6 Terra (high)OpenAIOpenAI
67.1
AA Coding Indexβ€”$2.00β€”
53Anthropic: Claude Sonnet 5 (batch)AnthropicAnthropic
66.4
AA Coding Index1.0M tokens$1.00β€”
54Qwen3.7 MaxAlibabaAlibaba
66.0
AA Coding Indexβ€”$2.50β€”
55GPT-5.6 Sol (Non-reasoning)OpenAIOpenAI
65.1
AA Coding Indexβ€”$4.00β€”
56DeepSeek V4 Flash Vision (Reasoning, Max Effort)DeepSeekDeepSeek
65.0
AA Coding Indexβ€”$0.44β€”
57GPT-5.6 Terra (medium)OpenAIOpenAI
64.7
AA Coding Indexβ€”$2.00β€”
58Motif 3Motif Technologies
63.5
AA Coding Indexβ€”β€”β€”
59GPT-5.6 Luna (high)OpenAIOpenAI
63.3
AA Coding Indexβ€”$0.20β€”
60Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort)AnthropicAnthropic
63.0
AA Coding Indexβ€”$3.00β€”
61Anthropic: Claude Sonnet 4.6 (batch)AnthropicAnthropic
63.0
AA Coding Index1.0M tokens$1.50β€”
62Agnes 2.5 Pro BetaSapiens AI
62.3
AA Coding Indexβ€”$0.10β€”
63Motif 3 (Beta)Motif Technologies
62.0
AA Coding Indexβ€”β€”β€”
64MoonshotAI: Kimi K2.6MoonshotAIMoonshotAI
61.8
AA Coding Index262K tokens$0.95βœ…
65OpenAI: GPT-5.5 (batch)OpenAIOpenAI
60.9
AA Coding Index1.1M tokens$2.50β€”
66Kimi K2.7 CodeKimiKimi
60.8
AA Coding Indexβ€”$0.95β€”
67Apodex 1.1Apodex
60.8
AA Coding Indexβ€”$0.30β€”
68Xiaomi: MiMo-V2.5-ProXiaomi
60.2
AA Coding Index1.0M tokens$0.43β€”
69Kwaipilot: KAT-Coder-Pro V2Kwaipilot
59.5
AA Coding Index256K tokens$0.30β€”
70DeepSeek V4 Pro (Non-reasoning)DeepSeekDeepSeek
59.4
AA Coding Indexβ€”$0.43β€”
71DeepSeek V4 Pro (Reasoning, Max Effort)DeepSeekDeepSeek
59.4
AA Coding Indexβ€”$0.43β€”
72Nex-N2-ProNex AGI
59.1
AA Coding Index262K tokens$0.50β€”
73Hy3-preview (Reasoning)Tencent
58.8
AA Coding Index262K tokens$0.14β€”
74Agnes 2.5 Pro AlphaSapiens AI
58.8
AA Coding Indexβ€”$0.45β€”
75Hy3-preview (Non-reasoning)Tencent
58.8
AA Coding Index262K tokens$0.06β€”
76DeepSeek V4 Pro (Reasoning, High Effort)DeepSeekDeepSeek
58.7
AA Coding Indexβ€”$0.43β€”
77Muse SparkMetaMeta
58.6
AA Coding Indexβ€”β€”β€”
78MiniMax-M3MiniMax
58.6
AA Coding Index1.0M tokens$0.30β€”
79GPT-5.6 Terra (low)OpenAIOpenAI
58.1
AA Coding Indexβ€”$2.00β€”
80OpenAI: GPT-5.6 Terra (batch)OpenAIOpenAI
58.1
AA Coding Index1.1M tokens$1.00β€”
81OpenAI: GPT-5.6 Terra ProOpenAIOpenAI
58.1
AA Coding Index1.1M tokens$2.00β€”
82MiMo-V2.5Xiaomi
56.8
AA Coding Indexβ€”$0.14β€”
83DeepSeek-V4-FlashDeepSeekDeepSeek
56.2
AA Coding Indexβ€”$0.14β€”
84DeepSeek V4 Flash (Reasoning, Max Effort)DeepSeekDeepSeek
56.2
AA Coding Indexβ€”$0.13β€”
85GPT-5.4 MiniOpenAIOpenAI
56.1
AA Coding Index400K tokens$0.75β€”
86GPT-5.4 NanoOpenAIOpenAI
56.1
AA Coding Index400K tokens$0.20β€”
87OpenAI: GPT-5.4 Mini (batch)OpenAIOpenAI
56.1
AA Coding Index400K tokens$0.38β€”
88OpenAI: GPT-5.4 Nano (batch)OpenAIOpenAI
56.1
AA Coding Index400K tokens$0.10β€”
89Qwen3.7 PlusAlibabaAlibaba
55.9
AA Coding Indexβ€”$0.40β€”
90GLM-5.1 (Non-reasoning)Z.aiZ.ai
55.8
AA Coding Indexβ€”$1.38β€”
91Z.ai: GLM 5.1Z.aiZ.ai
55.8
AA Coding Index203K tokens$1.28βœ…
92Qwen: Qwen3.6 PlusAlibabaAlibaba
54.5
AA Coding Index1.0M tokens$0.50βœ…
93Qwen: Qwen3.6 27BAlibabaAlibaba
53.7
AA Coding Index262K tokens$0.60βœ…
94Inkling-SmallMira Murati
52.9
AA Coding Index524K tokens$0.30β€”
95Inkling SmallThinking Machines
52.9
AA Coding Indexβ€”$0.30β€”
96Solar Pro 4Upstage
52.7
AA Coding Index524K tokens$0.30β€”
97MiniMax-M2MiniMax
52.6
AA Coding Index205K tokens$0.30β€”
98MiniMax: MiniMax M2.7MiniMax
52.6
AA Coding Index197K tokens$0.30βœ…
99JT-4.1 Flash 236B A21BChina Mobile
52.4
AA Coding Indexβ€”β€”β€”
100GPT-5.6 Terra (Non-reasoning)OpenAIOpenAI
52.3
AA Coding Indexβ€”$2.00β€”
101Claude 4.5 Sonnet (Reasoning)AnthropicAnthropic
52.1
AA Coding Indexβ€”$3.00β€”
102Claude 4.5 Sonnet (Non-reasoning)AnthropicAnthropic
52.1
AA Coding Indexβ€”$3.00β€”
103InklingThinking Machines
52.1
AA Coding Indexβ€”$1.00β€”
104InklingThinkyMachines
52.1
AA Coding Index1.0M tokens$1.00β€”
105DeepSeek V4 Flash (Reasoning, High Effort)DeepSeekDeepSeek
52.0
AA Coding Indexβ€”$0.13β€”
106Grok Build 0.1 0616xAIxAI
51.5
AA Coding Indexβ€”$1.00β€”
107GPT-5.6 Luna (medium)OpenAIOpenAI
50.7
AA Coding Indexβ€”$0.20β€”
108Ling-3.0-flashAnt Bailing
50.6
AA Coding Indexβ€”$0.07β€”
109MiMo-V2-Flash (Reasoning)Xiaomi
49.8
AA Coding Index262K tokens$0.10β€”
110MiMo-V2-Flash (Feb 2026)Xiaomi
49.8
AA Coding Indexβ€”β€”β€”
111GPT-5.1OpenAIOpenAI
49.4
AA Coding Index400K tokens$1.25β€”
112OpenAI: GPT-5.1 (batch)OpenAIOpenAI
49.4
AA Coding Index400K tokens$0.63β€”
113Gemini 3.5 Flash-LiteGoogleGoogle
49.3
AA Coding Index1.0M tokens$0.30β€”
114Nemotron 3 Ultra 550B A55B (Reasoning)NvidiaNvidia
49.3
AA Coding Index1.0M tokens$0.68β€”
115Muse Glimmer (high)MetaMeta
49.0
AA Coding Indexβ€”$0.32β€”
116Qwen: Qwen3.5 397B A17BAlibabaAlibaba
48.2
AA Coding Index262K tokens$0.60βœ…
117Mistral: Mistral Medium 3.5Mistral AIMistral AI
46.9
AA Coding Index262K tokens$1.50βœ…
118Kimi K2 ThinkingKimiKimi
46.8
AA Coding Index262K tokens$0.60β€”
119MoonshotAI: Kimi K2.5MoonshotAIMoonshotAI
46.8
AA Coding Index262K tokens$0.60βœ…
120Gemini 2.5 Pro Preview (Mar' 25)GoogleGoogle
46.7
AA Coding Indexβ€”β€”β€”
121GLM-4.6 (Reasoning)Z.aiZ.ai
45.8
AA Coding Indexβ€”$0.57β€”
122Qwen: Qwen3.5-122B-A10BAlibabaAlibaba
45.7
AA Coding Index262K tokens$0.40βœ…
123GLM-4.7 (Reasoning)Z.aiZ.ai
45.3
AA Coding Indexβ€”$0.60β€”
124LongCat 2.0LongCat
45.3
AA Coding Indexβ€”$0.75β€”
125Solar Open2 250BUpstage
44.7
AA Coding Indexβ€”β€”β€”
126DeepSeek V3.2 Exp (Reasoning)DeepSeekDeepSeek
44.2
AA Coding Indexβ€”$0.28β€”
127DeepSeek V3.2DeepSeekDeepSeek
44.2
AA Coding Index164K tokens$0.28βœ…
128GPT-5.6 Luna (low)OpenAIOpenAI
44.2
AA Coding Indexβ€”$0.20β€”
129Claude 4.5 Haiku (Reasoning)AnthropicAnthropic
43.9
AA Coding Indexβ€”$1.00β€”
130DeepSeek V3.1 TerminusDeepSeekDeepSeek
43.5
AA Coding Index164K tokens$0.27βœ…
131Gemma 4 31BGoogleGoogle
43.4
AA Coding Index262K tokensβ€”β€”
132Ring-2.6-1TInclusionAI
42.8
AA Coding Indexβ€”$0.30β€”
133SpaceXAI: Grok 4.3xAIxAI
42.2
AA Coding Index1.0M tokens$1.25β€”
134Qwen: Qwen3.6 35B A3BAlibabaAlibaba
41.9
AA Coding Index262K tokens$0.38βœ…
135o1OpenAIOpenAI
39.7
AA Coding Index200K tokens$15.00β€”
136OpenAI: o1 (batch)OpenAIOpenAI
39.7
AA Coding Index200K tokens$7.50β€”
137Step 3.7 FlashStepFun
39.6
AA Coding Indexβ€”$0.20β€”
138GPT-5.5 Instant (June 2026)OpenAIOpenAI
39.4
AA Coding Indexβ€”$5.00β€”
139GPT-5.6 Luna (Non-reasoning)OpenAIOpenAI
39.3
AA Coding Indexβ€”$0.20β€”
140Gemma 4 26B A4B GoogleGoogle
39.3
AA Coding Index262K tokens$0.13β€”
141A.X-K2SK Telecom
38.8
AA Coding Indexβ€”β€”β€”
142GPT-5OpenAIOpenAI
37.8
AA Coding Index400K tokens$1.25β€”
143OpenAI: GPT-5 (batch)OpenAIOpenAI
37.8
AA Coding Index400K tokens$0.63β€”
144NVIDIA Nemotron 3 Super 120B A12B (Reasoning)NvidiaNvidia
37.7
AA Coding Index1.0M tokens$0.20β€”
145Claude 4 Sonnet (Reasoning)AnthropicAnthropic
37.6
AA Coding Indexβ€”$3.00β€”
146Qwen: Qwen3.5-35B-A3BAlibabaAlibaba
37.0
AA Coding Index262K tokens$0.25βœ…
147North Mini CodeCohereCohere
36.5
AA Coding Indexβ€”β€”β€”
148Claude 3.7 Sonnet (thinking)AnthropicAnthropic
36.4
AA Coding Index200K tokensβ€”β€”
149Claude 3.7 SonnetAnthropicAnthropic
36.4
AA Coding Index200K tokens$3.00β€”
150Qwen: Qwen3 Coder NextAlibabaAlibaba
36.2
AA Coding Index262K tokens$0.35βœ…
151Gemini 3.1 Flash Lite PreviewGoogleGoogle
34.7
AA Coding Index1.0M tokens$0.25β€”
152Gemini 3.1 Flash LiteGoogleGoogle
34.7
AA Coding Index1.0M tokens$0.25β€”
153Google: Gemini 3.1 Flash Lite (batch)GoogleGoogle
34.7
AA Coding Index1.0M tokens$0.13β€”
154Nova 2.0 Pro Preview (medium)AmazonAmazon
34.0
AA Coding Indexβ€”$1.25β€”
155o1-previewOpenAIOpenAI
34.0
AA Coding Indexβ€”$16.50β€”
156Gemini 2.5GoogleGoogle
33.3
AA Coding Indexβ€”β€”β€”
157Gemini 2.5 ProGoogleGoogle
33.3
AA Coding Index1.0M tokens$1.25β€”
158G9v3-39A5BAI9Stars
33.1
AA Coding Indexβ€”β€”β€”
159Devstral 2MistralMistral
31.3
AA Coding Indexβ€”β€”β€”
160Inception: Mercury 2Inception
31.1
AA Coding Index128K tokens$0.25β€”
161Gemma 4 12B (Reasoning)GoogleGoogle
31.0
AA Coding Indexβ€”$0.10β€”
162gpt-oss-120bOpenAIOpenAI
30.4
AA Coding Index131K tokens$0.15β€”
163Claude 3.5 Sonnet (Oct '24)AnthropicAnthropic
30.2
AA Coding Indexβ€”$3.00β€”
164Granite 4.2 30BIBM
29.9
AA Coding Indexβ€”$0.16β€”
165Devstral Small 2MistralMistral
29.3
AA Coding Indexβ€”β€”β€”
166Qwen3.5 9B (Reasoning)AlibabaAlibaba
28.7
AA Coding Indexβ€”$0.14β€”
167Cohere: Command ACohereCohere
27.8
AA Coding Index256K tokens$2.50β€”
168Command A+CohereCohere
27.8
AA Coding Indexβ€”β€”β€”
169Nemotron 3.5 LightningNVIDIANVIDIA
26.8
AA Coding Index262K tokens$0.07β€”
170Mistral: Mistral Small 4Mistral AIMistral AI
26.6
AA Coding Index262K tokens$0.15βœ…
171Ling 3.0 TinyInclusionAI
26.5
AA Coding Indexβ€”β€”β€”
172Mistral Small 3.1MistralMistral
26.3
AA Coding Indexβ€”$0.10β€”
173Claude 3.5 Sonnet (June '24)AnthropicAnthropic
26.0
AA Coding Indexβ€”$3.00β€”
174Arcee AI: Trinity Large ThinkingArcee AI
25.8
AA Coding Index262K tokens$0.23βœ…
175Gemini 2.0 Pro Experimental (Feb '25)GoogleGoogle
25.5
AA Coding Indexβ€”β€”β€”
176Nemotron Cascade 2 30B A3BNvidiaNvidia
25.3
AA Coding Indexβ€”β€”β€”
177Ling 2.6 FlashInclusion AI
25.3
AA Coding Indexβ€”$0.10β€”
178DeepSeek R1 (Jan '25)DeepSeekDeepSeek
24.6
AA Coding Indexβ€”$1.35β€”
179DeepSeek: R1DeepSeekDeepSeek
24.6
AA Coding Index64K tokens$0.70βœ…
180GPT-4o (March 2025, chatgpt-4o-latest)OpenAIOpenAI
24.2
AA Coding Indexβ€”β€”β€”
181OpenAI: GPT-4o (2024-05-13)OpenAIOpenAI
24.2
AA Coding Index128K tokens$5.00β€”
182OpenAI: GPT-4o (batch)OpenAIOpenAI
24.2
AA Coding Index128K tokens$1.25β€”
183Gemini 2.0 Flash Thinking Experimental (Jan '25)GoogleGoogle
24.1
AA Coding Indexβ€”β€”β€”
184Gemini 1.5 Pro (Sep '24)GoogleGoogle
23.6
AA Coding Indexβ€”β€”β€”
185EXAONE 4.5 33BLG AI
23.6
AA Coding Indexβ€”β€”β€”
186HyperNova 60B 2605Multiverse Computing
23.2
AA Coding Indexβ€”$0.04β€”
187Nova 2.0 Lite (high)AmazonAmazon
23.0
AA Coding Indexβ€”$0.30β€”
188DeepSeek V3DeepSeekDeepSeek
23.0
AA Coding Index164K tokens$0.36βœ…
189Qwen3.5 4B (Reasoning)AlibabaAlibaba
22.6
AA Coding Indexβ€”$0.03β€”
190Granite 4.2 8BIBM
22.4
AA Coding Indexβ€”$0.06β€”
191Qwen: Qwen3 235B A22B Instruct 2507AlibabaAlibaba
22.1
AA Coding Index262K tokens$0.23βœ…
192Qwen: Qwen3 235B A22B Thinking 2507AlibabaAlibaba
22.1
AA Coding Index131K tokens$0.15βœ…
193OpenAI: GPT-4 Turbo (older v1106)OpenAIOpenAI
21.5
AA Coding Index128K tokens$10.00β€”
194GPT-4 TurboOpenAIOpenAI
21.5
AA Coding Index128K tokens$10.00β€”
195OpenAI: GPT-4 Turbo (batch)OpenAIOpenAI
21.5
AA Coding Index128K tokens$5.00β€”
196Magistral Medium 1.2Mistral AIMistral AI
21.3
AA Coding Indexβ€”$2.00β€”
197DeepSeek V3 0324DeepSeekDeepSeek
21.2
AA Coding Indexβ€”$0.27β€”
198K2 Think V2MBZUAI Institute of Foundation Models
21.0
AA Coding Indexβ€”β€”β€”
199gpt-oss-20bOpenAIOpenAI
20.7
AA Coding Index131K tokens$0.06β€”
200Mistral: Mistral Medium 3.1Mistral AIMistral AI
20.5
AA Coding Index131K tokens$0.40βœ…
201Qwen3.5 4B (Non-reasoning)AlibabaAlibaba
20.3
AA Coding Indexβ€”$0.03β€”
202GPT-4.1 MiniOpenAIOpenAI
20.2
AA Coding Index1.0M tokens$0.40β€”
203OpenAI: GPT-4.1 Mini (batch)OpenAIOpenAI
20.2
AA Coding Index1.0M tokens$0.20β€”
204Mistral Large 3MistralMistral
20.1
AA Coding Indexβ€”$0.50β€”
205Gemini 1.5 Pro (May '24)GoogleGoogle
19.8
AA Coding Indexβ€”β€”β€”
206DiffusionGemma 26B A4BGoogleGoogle
19.7
AA Coding Indexβ€”β€”β€”
207Claude 3 OpusAnthropicAnthropic
19.5
AA Coding Indexβ€”$15.00β€”
208Gemini 1.0 UltraGoogleGoogle
17.6
AA Coding Indexβ€”β€”β€”
209Granite 4.2 3BIBM
17.5
AA Coding Indexβ€”$0.03β€”
210Qwen3 Next 80B A3B (Reasoning)AlibabaAlibaba
17.4
AA Coding Indexβ€”$0.15β€”
211o3 Mini HighOpenAIOpenAI
16.3
AA Coding Index200K tokens$1.10β€”
212Llama 4 MaverickMetaMeta
16.3
AA Coding Index1.0M tokens$0.26βœ…
213Solar Pro 3Upstage
16.2
AA Coding Index128K tokens$0.15β€”
214Claude 3.5 HaikuAnthropicAnthropic
15.9
AA Coding Index200K tokensβ€”β€”
215GPT-5 MiniOpenAIOpenAI
15.6
AA Coding Index400K tokens$0.25β€”
216OpenAI: GPT-5 Mini (batch)OpenAIOpenAI
15.6
AA Coding Index400K tokens$0.13β€”
217Qwen3 32B (Reasoning)AlibabaAlibaba
15.3
AA Coding Indexβ€”$0.16β€”
218Magistral Small 1.2MistralMistral
14.7
AA Coding Indexβ€”$0.50β€”
219NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)NvidiaNvidia
14.4
AA Coding Indexβ€”$0.05β€”
220NVIDIA: Nemotron 3 Nano 30B A3BNvidiaNvidia
14.4
AA Coding Index262K tokens$0.05β€”
221Ministral 3 14BMistralMistral
14.4
AA Coding Indexβ€”$0.20β€”
222Celeris-1Unknown
14.4
AA Coding Indexβ€”$0.20β€”
223Claude 2.1AnthropicAnthropic
14.0
AA Coding Indexβ€”β€”β€”
224Qwen3 14B (Reasoning)AlibabaAlibaba
13.8
AA Coding Indexβ€”$0.35β€”
225Nemotron 3 Nano Omni 30B A3B ReasoningNvidiaNvidia
13.8
AA Coding Indexβ€”$0.07β€”
226OpenAI: GPT-4OpenAIOpenAI
13.1
AA Coding Index8K tokens$30.00β€”
227Claude 2.0AnthropicAnthropic
12.9
AA Coding Indexβ€”β€”β€”
228Mistral Small 3.2MistralMistral
12.5
AA Coding Indexβ€”$0.10β€”
229Qwen3 30B A3B 2507 (Reasoning)AlibabaAlibaba
12.1
AA Coding Indexβ€”$0.20β€”
230Qwen3 30B A3B 2507 InstructAlibabaAlibaba
12.1
AA Coding Indexβ€”$0.20β€”
231Llama 3.3 70B InstructMetaMeta
11.9
AA Coding Index131K tokens$0.66βœ…
232OpenAI: GPT-4o-miniOpenAIOpenAI
11.4
AA Coding Index128K tokens$0.15β€”
233OpenAI: GPT-4o-mini (batch)OpenAIOpenAI
11.4
AA Coding Index128K tokens$0.07β€”
234GPT-4.1 NanoOpenAIOpenAI
11.1
AA Coding Index1.0M tokens$0.10β€”
235OpenAI: GPT-4.1 Nano (batch)OpenAIOpenAI
11.1
AA Coding Index1.0M tokens$0.05β€”
236OpenAI: GPT-3.5 Turbo (batch)OpenAIOpenAI
10.7
AA Coding Index16K tokens$0.25β€”
237GPT-3.5 TurboOpenAIOpenAI
10.7
AA Coding Indexβ€”$0.50β€”
238Granite 4.1 30BIBM
10.4
AA Coding Indexβ€”β€”β€”
239Gemma 3 27BGoogleGoogle
10.1
AA Coding Index131K tokensβ€”β€”
240G9v3-3BAI9Stars
9.9
AA Coding Indexβ€”β€”β€”
241Ministral 3 8BMistralMistral
9.7
AA Coding Indexβ€”$0.15β€”
242Nanbeige4.1-3BNanbeige
9.6
AA Coding Indexβ€”β€”β€”
243Granite 4.1 8BIBM
9.5
AA Coding Indexβ€”$0.05β€”
244Gemma 4 E4B (Reasoning)GoogleGoogle
9.4
AA Coding Indexβ€”$0.02β€”
245Qwen3 8B (Reasoning)AlibabaAlibaba
9.0
AA Coding Indexβ€”$0.18β€”
246Llama 4 ScoutMetaMeta
8.2
AA Coding Index1.3M tokens$0.18βœ…
247NVIDIA Nemotron 3 Nano 4BNvidiaNvidia
8.0
AA Coding Indexβ€”β€”β€”
248Claude InstantAnthropicAnthropic
7.8
AA Coding Indexβ€”β€”β€”
249LFM2.5-2.6Bβ€”
7.7
AA Coding Indexβ€”β€”β€”
250Gemma 4 E2B (Reasoning)GoogleGoogle
7.2
AA Coding Indexβ€”β€”β€”
251Gemma 3 12BGoogleGoogle
5.8
AA Coding Index131K tokensβ€”β€”
252Llama 3.1 8B InstructMetaMeta
5.4
AA Coding Index16K tokens$0.02βœ…
253Ministral 3 3BMistralMistral
4.8
AA Coding Indexβ€”$0.10β€”
254Granite 4.1 3BIBM
4.7
AA Coding Indexβ€”β€”β€”
255PALM-2GoogleGoogle
4.6
AA Coding Indexβ€”β€”β€”
256Phi-4 Mini InstructMicrosoftMicrosoft
3.8
AA Coding Indexβ€”β€”β€”
257Gemma 3n E4B InstructGoogleGoogle
3.2
AA Coding Indexβ€”$0.06β€”
258Qwen3.5 2B (Reasoning)AlibabaAlibaba
2.9
AA Coding Indexβ€”β€”β€”
259Gemma 3 4BGoogleGoogle
2.7
AA Coding Index131K tokensβ€”β€”
260Qwen3.5 0.8B (Non-reasoning)AlibabaAlibaba
1.2
AA Coding Indexβ€”β€”β€”
261MiniCPM-V 4.6 1.3BOpenBMB
0.7
AA Coding Indexβ€”β€”β€”
262Qwen3.5 0.8B (Reasoning)AlibabaAlibaba
0.0
AA Coding Indexβ€”β€”β€”
263Gemini 3 Pro Preview (high)GoogleGoogle
92.0
LiveCodeBenchβ€”$2.00β€”
264Gemini 3 Flash Preview (Reasoning)GoogleGoogle
91.0
LiveCodeBenchβ€”$0.50β€”
265Google: Gemini 3 Flash Preview (batch)GoogleGoogle
90.8
LiveCodeBench1.0M tokens$0.25β€”
266DeepSeek V3.2 SpecialeDeepSeekDeepSeek
90.0
LiveCodeBench164K tokensβ€”βœ…
267GPT-5.2OpenAIOpenAI
89.0
LiveCodeBench400K tokens$1.75β€”
268Claude Opus 4.5 (Reasoning)AnthropicAnthropic
87.0
LiveCodeBenchβ€”$5.00β€”
269Gemini 3 Pro Preview (low)GoogleGoogle
86.0
LiveCodeBenchβ€”$2.00β€”
270o4 MiniOpenAIOpenAI
86.0
LiveCodeBench200K tokens$1.10β€”
271o4 Mini HighOpenAIOpenAI
85.9
LiveCodeBench200K tokens$1.10β€”
272GPT-5.1-CodexOpenAIOpenAI
85.0
LiveCodeBench400K tokens$1.25β€”
273GPT-5.1-Codex-MaxOpenAIOpenAI
84.9
LiveCodeBench400K tokens$1.25β€”
274GPT-5.1-Codex-MiniOpenAIOpenAI
84.0
LiveCodeBench400K tokens$0.25β€”
275GPT-5 CodexOpenAIOpenAI
84.0
LiveCodeBench400K tokens$1.25β€”
276Grok 4 FastxAIxAI
83.0
LiveCodeBench2.0M tokens$0.20β€”
277Grok 4xAIxAI
82.0
LiveCodeBench256K tokens$3.00β€”
278Grok 4.1 FastxAIxAI
82.0
LiveCodeBench2.0M tokensβ€”β€”
279MiniMax: MiniMax M2.1MiniMax
81.0
LiveCodeBench197K tokens$0.30βœ…
280o3OpenAIOpenAI
81.0
LiveCodeBench200K tokens$2.00β€”
281ERNIE 5.0 Thinking PreviewBaidu
81.0
LiveCodeBenchβ€”β€”β€”
282Apriel-v1.6-15B-ThinkerServiceNow
81.0
LiveCodeBenchβ€”β€”β€”
283o3 ProOpenAIOpenAI
80.8
LiveCodeBench200K tokens$20.00β€”
284Gemini 3 Flash Preview (Non-reasoning)GoogleGoogle
80.0
LiveCodeBenchβ€”$0.50β€”
285GPT-5 NanoOpenAIOpenAI
79.0
LiveCodeBench400K tokens$0.05β€”
286DeepSeek V3.2 ExpDeepSeekDeepSeek
78.9
LiveCodeBench164K tokens$0.27βœ…
287INTELLECT-3Prime Intellect
78.0
LiveCodeBench131K tokensβ€”βœ…
288DeepSeek V3.1DeepSeekDeepSeek
78.0
LiveCodeBench164K tokens$0.57βœ…
289Gemini 2.5 Pro Preview (May' 25)GoogleGoogle
77.0
LiveCodeBenchβ€”$1.25β€”
290Qwen: Qwen3 MaxAlibabaAlibaba
77.0
LiveCodeBench262K tokens$1.20βœ…
291Seed-OSS-36B-InstructByteDance Seed
77.0
LiveCodeBenchβ€”$0.21β€”
292OpenAI: GPT-5 Nano (batch)OpenAIOpenAI
76.3
LiveCodeBench400K tokens$0.03β€”
293Claude Sonnet 4.5AnthropicAnthropic
76.1
LiveBench Coding1.0M tokens$3.00β€”
294KAT-Coder-Pro V1KwaiKAT
75.0
LiveCodeBenchβ€”β€”β€”
295EXAONE 4.0 32B (Reasoning)LG AI Research
75.0
LiveCodeBenchβ€”β€”β€”
296Gemini 3.5GoogleGoogle
74.6
LiveBench Codingβ€”β€”β€”
297Claude Opus 4.5AnthropicAnthropic
74.0
LiveCodeBench200K tokens$5.00β€”
298GLM-4.5 (Reasoning)Z.aiZ.ai
74.0
LiveCodeBench131K tokensβ€”β€”
299Llama Nemotron Super 49B v1.5 (Reasoning)NvidiaNvidia
74.0
LiveCodeBenchβ€”$0.40β€”
300Qwen3 VL 32B (Reasoning)AlibabaAlibaba
74.0
LiveCodeBenchβ€”$0.16β€”
301Anthropic: Claude Opus 4.5 (batch)AnthropicAnthropic
73.8
LiveCodeBench200K tokens$2.50β€”
302Apriel-v1.5-15B-ThinkerServiceNow
73.0
LiveCodeBenchβ€”β€”β€”
303o3 MiniOpenAIOpenAI
72.0
LiveCodeBench200K tokens$1.10β€”
304Falcon-H1R-7BTII UAE
72.0
LiveCodeBenchβ€”β€”β€”
305NVIDIA Nemotron Nano 9B V2 (Reasoning)NvidiaNvidia
72.0
LiveCodeBenchβ€”$0.04β€”
306Gemini 2.5 Flash Preview (Sep '25) (Reasoning)GoogleGoogle
71.0
LiveCodeBenchβ€”β€”β€”
307MiniMax M1 80kMiniMax
71.0
LiveCodeBenchβ€”$0.55β€”
308NVIDIA: Nemotron Nano 9B V2NvidiaNvidia
70.1
LiveCodeBench131K tokens$0.05β€”
309Grok 3 MinixAIxAI
70.0
LiveCodeBench131K tokens$0.30β€”
310Gemini 2.5 Flash Preview (Reasoning)GoogleGoogle
70.0
LiveCodeBenchβ€”$0.30β€”
311Olmo 3.1 32B ThinkAllen Institute for AI
70.0
LiveCodeBenchβ€”β€”β€”
312Qwen3 VL 30B A3B (Reasoning)AlibabaAlibaba
70.0
LiveCodeBenchβ€”$0.20β€”
313NVIDIA Nemotron Nano 9B V2 (Non-reasoning)NvidiaNvidia
70.0
LiveCodeBench131K tokens$0.05β€”
314Cogito v2.1 (Reasoning)Deep Cogito
69.0
LiveCodeBenchβ€”$1.25β€”
315Gemini 2.5 Flash-Lite Preview (Sep '25) (Reasoning)GoogleGoogle
69.0
LiveCodeBenchβ€”$0.10β€”
316NVIDIA Nemotron Nano 12B v2 VL (Reasoning)NvidiaNvidia
69.0
LiveCodeBenchβ€”$0.20β€”
317Hermes 4 - Llama-3.1 405B (Reasoning)Nous Research
69.0
LiveCodeBenchβ€”$1.00β€”
318Ling-1TInclusionAI
68.0
LiveCodeBenchβ€”β€”β€”
319Qwen3 Omni 30B A3B (Reasoning)AlibabaAlibaba
68.0
LiveCodeBenchβ€”$0.25β€”
320Qwen: Qwen3 Next 80B A3B InstructAlibabaAlibaba
68.0
LiveCodeBench262K tokens$0.15βœ…
321Olmo 3 32B ThinkAllenAI
67.0
LiveCodeBench66K tokensβ€”βœ…
322GPT-5.2-CodexOpenAIOpenAI
66.9
LiveCodeBench400K tokens$1.75β€”
323OpenAI: GPT-5.2 (batch)OpenAIOpenAI
66.9
LiveCodeBench400K tokens$0.88β€”
324MiniMax M1 40kMiniMax
66.0
LiveCodeBenchβ€”β€”β€”
325Nova 2.0 Omni (medium)AmazonAmazon
66.0
LiveCodeBenchβ€”$0.30β€”
326Grok Code Fast 1xAIxAI
66.0
LiveCodeBench256K tokensβ€”β€”
327Mi:dm K 2.5 ProKorea Telecom
66.0
LiveCodeBenchβ€”β€”β€”
328Claude 4.1 Opus (Non-reasoning)AnthropicAnthropic
65.4
LiveCodeBenchβ€”$15.00β€”
329xAI: Grok Build 0.1xAIxAI
65.4
LiveBench Coding256K tokens$1.00β€”
330Claude 4.1 Opus (Reasoning)AnthropicAnthropic
65.0
LiveCodeBenchβ€”$15.00β€”
331Qwen3 VL 235B A22B (Reasoning)AlibabaAlibaba
65.0
LiveCodeBenchβ€”$0.40β€”
332Qwen3 Max (Preview)AlibabaAlibaba
65.0
LiveCodeBenchβ€”$1.20β€”
333Hermes 4 - Llama-3.1 70B (Reasoning)Nous Research
65.0
LiveCodeBenchβ€”$0.13β€”
334Motif-2-12.7B-ReasoningMotif Technologies
65.0
LiveCodeBenchβ€”β€”β€”
335Claude 4 Opus (Reasoning)AnthropicAnthropic
64.0
LiveCodeBenchβ€”$15.00β€”
336Ring-1TInclusionAI
64.0
LiveCodeBenchβ€”β€”β€”
337Llama 3.1 Nemotron Ultra 253B v1 (Reasoning)NvidiaNvidia
64.0
LiveCodeBenchβ€”$0.60β€”
338Gemini 2.5 Flash-Lite Preview (Sep '25) (Non-reasoning)GoogleGoogle
64.0
LiveCodeBenchβ€”$0.10β€”
339Qwen3 4B 2507 (Reasoning)AlibabaAlibaba
64.0
LiveCodeBenchβ€”β€”β€”
340QwQ 32BAlibabaAlibaba
63.0
LiveCodeBenchβ€”$0.66β€”
341HyperCLOVA X SEED Think (32B)Naver
63.0
LiveCodeBenchβ€”β€”β€”
342Ring-flash-2.0InclusionAI
63.0
LiveCodeBenchβ€”$0.14β€”
343Qwen3 235B A22B (Reasoning)AlibabaAlibaba
62.0
LiveCodeBenchβ€”$0.70β€”
344Solar Pro 2 (Non-reasoning)Upstage
62.0
LiveCodeBenchβ€”β€”β€”
345Olmo 3 7B ThinkAllen Institute for AI
62.0
LiveCodeBenchβ€”β€”β€”
346MoonshotAI: Kimi K2 0905MoonshotAIMoonshotAI
61.0
LiveCodeBench262K tokens$0.60βœ…
347GLM-4.5V (Reasoning)Z.aiZ.ai
60.0
LiveCodeBenchβ€”$0.60β€”
348Qwen: Qwen3 VL 235B A22B InstructAlibabaAlibaba
59.0
LiveCodeBench262K tokens$0.40βœ…
349Qwen3 Coder 480B A35B InstructAlibabaAlibaba
59.0
LiveCodeBenchβ€”$1.50β€”
350Nova 2.0 Omni (low)AmazonAmazon
59.0
LiveCodeBenchβ€”$0.30β€”
351Ling-flash-2.0InclusionAI
59.0
LiveCodeBenchβ€”$0.14β€”
352Gemini 2.5 Flash LiteGoogleGoogle
59.0
LiveCodeBench1.0M tokens$0.10β€”
353o1-miniOpenAIOpenAI
58.0
LiveCodeBenchβ€”β€”β€”
354Mi:dm K 2.5 Pro PreviewKorea Telecom
58.0
LiveCodeBenchβ€”β€”β€”
355GPT-5 (minimal)OpenAIOpenAI
56.0
LiveCodeBenchβ€”$1.25β€”
356Kimi K2Moonshot AIMoonshot AI
56.0
LiveCodeBench131K tokens$0.57β€”
357DeepSeek V3.2 Exp (Non-reasoning)DeepSeekDeepSeek
55.0
LiveCodeBenchβ€”$0.28β€”
358GPT-5 mini (minimal)OpenAIOpenAI
55.0
LiveCodeBenchβ€”$0.25β€”
359Hermes 4 - Llama-3.1 405B (Non-reasoning)Nous Research
55.0
LiveCodeBenchβ€”$1.00β€”
360Claude Opus 4AnthropicAnthropic
54.0
LiveCodeBench200K tokens$15.00β€”
361Qwen3 Max Thinking (Preview)AlibabaAlibaba
54.0
LiveCodeBenchβ€”$1.20β€”
362GPT-5 (ChatGPT)OpenAIOpenAI
54.0
LiveCodeBenchβ€”β€”β€”
363K2-V2 (medium)MBZUAI Institute of Foundation Models
54.0
LiveCodeBenchβ€”β€”β€”
364Magistral Medium 1MistralMistral
53.0
LiveCodeBenchβ€”β€”β€”
365GPT-5.3-CodexOpenAIOpenAI
53.0
SciCode400K tokens$1.75β€”
366Exaone 4.0 1.2B (Non-reasoning)LG AI Research
52.0
LiveCodeBenchβ€”β€”β€”
367Claude Opus 4.6 (Adaptive Reasoning, Max Effort)AnthropicAnthropic
52.0
SciCodeβ€”$5.00β€”
368Claude Haiku 4.5AnthropicAnthropic
51.0
LiveCodeBench200K tokens$1.00β€”
369Qwen: Qwen3 VL 32B InstructAlibabaAlibaba
51.0
LiveCodeBench131K tokens$0.16βœ…
370Qwen3 30B A3B (Reasoning)AlibabaAlibaba
51.0
LiveCodeBenchβ€”$0.20β€”
371Magistral Small 1MistralMistral
51.0
LiveCodeBenchβ€”β€”β€”
372DeepSeek R1 0528 Qwen3 8BDeepSeekDeepSeek
51.0
LiveCodeBenchβ€”β€”β€”
373Gemini 2.5 FlashGoogleGoogle
50.0
LiveCodeBench1.0M tokens$0.30β€”
374GPT-5.5 Instant (May 2026)OpenAIOpenAI
50.0
SciCodeβ€”$5.00β€”
375Google: Gemini 2.5 Flash (batch)GoogleGoogle
49.5
LiveCodeBench1.0M tokens$0.15β€”
376GPT-5.1 ChatOpenAIOpenAI
49.4
LiveCodeBench128K tokens$1.25β€”
377Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)NvidiaNvidia
49.0
LiveCodeBenchβ€”β€”β€”
378Gemini 3.5 Flash (minimal)GoogleGoogle
49.0
SciCodeβ€”$1.50β€”
379Qwen: Qwen3 VL 30B A3B InstructAlibabaAlibaba
48.0
LiveCodeBench131K tokens$0.20βœ…
380Baidu: ERNIE 4.5 300B A47B Baidu
47.0
LiveCodeBench123K tokens$0.28βœ…
381GPT-5 nano (minimal)OpenAIOpenAI
47.0
LiveCodeBenchβ€”$0.05β€”
382EXAONE 4.0 32B (Non-reasoning)LG AI Research
47.0
LiveCodeBenchβ€”β€”β€”
383Qwen3 4B (Reasoning)AlibabaAlibaba
47.0
LiveCodeBenchβ€”β€”β€”
384Qwen3.6 Max PreviewAlibabaAlibaba
47.0
SciCodeβ€”$1.30β€”
385Claude Sonnet 4.6AnthropicAnthropic
47.0
SciCode1.0M tokens$3.00β€”
386GPT-4.1OpenAIOpenAI
46.0
LiveCodeBench1.0M tokens$2.00β€”
387Solar Pro 2 (Preview) (Reasoning)Upstage
46.0
LiveCodeBenchβ€”β€”β€”
388Claude Opus 4.6AnthropicAnthropic
46.0
SciCode1.0M tokens$5.00β€”
389Claude Sonnet 4AnthropicAnthropic
45.0
LiveCodeBench1.0M tokens$3.00β€”
390Grok 4.20 0309 (Reasoning)xAIxAI
45.0
SciCodeβ€”$2.00β€”
391Reka Flash 3Reka Flash 3
44.0
LiveCodeBench66K tokens$0.20βœ…
392GLM-5-TurboZ.aiZ.ai
44.0
SciCode203K tokensβ€”β€”
393Claude Sonnet 4.6 (Non-reasoning, Low Effort)AnthropicAnthropic
44.0
SciCodeβ€”$3.00β€”
394Grok 3xAIxAI
43.0
LiveCodeBench131K tokens$4.00β€”
395Ling-mini-2.0InclusionAI
43.0
LiveCodeBenchβ€”β€”β€”
396Xiaomi: MiMo-V2-ProXiaomi
43.0
SciCode1.0M tokensβ€”β€”
397MiniMax: MiniMax M2.5MiniMax
43.0
SciCode197K tokens$0.30βœ…
398Qwen: Qwen3 Max ThinkingAlibabaAlibaba
43.0
SciCode262K tokens$0.78βœ…
399GPT-4o (2024-11-20)OpenAIOpenAI
42.5
LiveCodeBench128K tokens$2.50β€”
400Qwen3 Omni 30B A3B InstructAlibabaAlibaba
42.0
LiveCodeBenchβ€”$0.25β€”
401Gemini 2.5 Flash Preview (Non-reasoning)GoogleGoogle
41.0
LiveCodeBenchβ€”β€”β€”
402Qwen3.5 Omni PlusAlibabaAlibaba
41.0
SciCodeβ€”$0.40β€”
403Mistral: Mistral Medium 3Mistral AIMistral AI
40.0
LiveCodeBench131K tokens$0.40βœ…
404Qwen: Qwen3 Coder 30B A3B InstructAlibabaAlibaba
40.0
LiveCodeBench160K tokens$0.45βœ…
405Google: Gemini 2.5 Flash Lite (batch)GoogleGoogle
40.0
LiveCodeBench1.0M tokens$0.05β€”
406MiMo-V2-Omni-0327Xiaomi
40.0
SciCodeβ€”β€”β€”
407Qwen: Qwen3.5-27BAlibabaAlibaba
40.0
SciCode262K tokens$0.30βœ…
408Step 3.5 FlashStepFun
40.0
SciCodeβ€”$0.10β€”
409Solar Pro 2 (Preview) (Non-reasoning)Upstage
39.0
LiveCodeBenchβ€”β€”β€”
410DeepSeek R1 Distill Qwen 14BDeepSeekDeepSeek
38.0
LiveCodeBenchβ€”β€”β€”
411Kimi Linear 48B A3B InstructKimiKimi
38.0
LiveCodeBenchβ€”β€”β€”
412Qwen3 4B 2507 InstructAlibabaAlibaba
38.0
LiveCodeBenchβ€”β€”β€”
413GLM-5 (Non-reasoning)Z.aiZ.ai
38.0
SciCode205K tokens$1.00β€”
414Ling-2.6-1TInclusion AI
37.0
SciCodeβ€”$0.30β€”
415Xiaomi: MiMo-V2-OmniXiaomi
37.0
SciCode262K tokensβ€”β€”
416Qwen2.5 MaxAlibabaAlibaba
36.0
LiveCodeBenchβ€”β€”β€”
417NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)NvidiaNvidia
36.0
LiveCodeBench262K tokens$0.05β€”
418Qwen3 VL 8B (Reasoning)AlibabaAlibaba
35.0
LiveCodeBenchβ€”$0.18β€”
419NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning)NvidiaNvidia
35.0
LiveCodeBenchβ€”$0.20β€”
420Mistral: Devstral MediumMistral AIMistral AI
34.0
LiveCodeBench131K tokensβ€”βœ…
421QwQ 32B-PreviewAlibabaAlibaba
34.0
LiveCodeBenchβ€”β€”β€”
422Gemini 2.0 FlashGoogleGoogle
33.0
LiveCodeBench1.0M tokensβ€”β€”
423Qwen: Qwen3 VL 8B InstructAlibabaAlibaba
33.0
LiveCodeBench131K tokens$0.18βœ…
424GPT-4o (ChatGPT)OpenAIOpenAI
33.0
SciCodeβ€”β€”β€”
425Qwen: Qwen3 30B A3B Thinking 2507AlibabaAlibaba
32.2
LiveCodeBench131K tokens$0.08βœ…
426GPT-4o (2024-08-06)OpenAIOpenAI
32.0
LiveCodeBench128K tokens$2.50β€”
427Qwen: Qwen3 30B A3B Instruct 2507AlibabaAlibaba
32.0
LiveCodeBench262K tokens$0.20βœ…
428Qwen3 VL 4B (Reasoning)AlibabaAlibaba
32.0
LiveCodeBenchβ€”β€”β€”
429OpenAI: GPT-4oOpenAIOpenAI
31.0
LiveCodeBench128K tokens$2.50β€”
430Llama 3.1 Instruct 405BMetaMeta
31.0
LiveCodeBenchβ€”$2.50β€”
431Nova 2.0 Omni (Non-reasoning)AmazonAmazon
31.0
LiveCodeBenchβ€”$0.30β€”
432Qwen3 1.7B (Reasoning)AlibabaAlibaba
31.0
LiveCodeBenchβ€”β€”β€”
433Step3 VL 10BStepFun
31.0
SciCodeβ€”β€”β€”
434Qwen2.5 Coder 32B InstructAlibabaAlibaba
30.0
LiveCodeBench33K tokensβ€”βœ…
435SonarPerplexityPerplexity
30.0
LiveCodeBench127K tokensβ€”β€”
436Sarvam M (Reasoning)Sarvam
30.0
LiveCodeBenchβ€”β€”β€”
437Llama 3.1 Tulu3 405BAllen Institute for AI
29.0
LiveCodeBenchβ€”β€”β€”
438Mistral Large 2 (Nov '24)MistralMistral
29.0
LiveCodeBenchβ€”$4.00β€”
439Qwen3 32B (Non-reasoning)AlibabaAlibaba
29.0
LiveCodeBenchβ€”$0.16β€”
440Llama Nemotron Super 49B v1.5 (Non-reasoning)NvidiaNvidia
29.0
LiveCodeBenchβ€”$0.40β€”
441Qwen3 VL 4B InstructAlibabaAlibaba
29.0
LiveCodeBenchβ€”β€”β€”
442JT-35B-FlashChina Mobile
29.0
SciCodeβ€”β€”β€”
443Llama 3.3 Nemotron Super 49B v1 (Reasoning)NvidiaNvidia
28.0
LiveCodeBenchβ€”β€”β€”
444Qwen3 14B (Non-reasoning)AlibabaAlibaba
28.0
LiveCodeBenchβ€”$0.35β€”
445Qwen2.5 72B InstructAlibabaAlibaba
28.0
LiveCodeBench33K tokens$0.47βœ…
446Llama 3.3 Nemotron Super 49B v1 (Non-reasoning)NvidiaNvidia
28.0
LiveCodeBenchβ€”β€”β€”
447Sonar Reasoning ProPerplexityPerplexity
28.0
LiveCodeBench128K tokensβ€”β€”
448LongCat Flash LiteLongCat
28.0
SciCodeβ€”β€”β€”
449DeepSeek: R1 Distill Qwen 32BDeepSeekDeepSeek
27.0
LiveCodeBench128K tokensβ€”βœ…
450R1 Distill Llama 70BDeepSeekDeepSeek
27.0
LiveCodeBench8K tokens$0.70βœ…
451Hermes 4 - Llama-3.1 70B (Non-reasoning)Nous Research
27.0
LiveCodeBenchβ€”$0.13β€”
452Grok 2 (Dec '24)xAIxAI
27.0
LiveCodeBenchβ€”β€”β€”
453Gemini 1.5 Flash (Sep '24)GoogleGoogle
27.0
LiveCodeBenchβ€”β€”β€”
454Mistral Large 2 (Jul '24)MistralMistral
27.0
LiveCodeBench131K tokens$2.00β€”
455Olmo 3 7B InstructAllen Institute for AI
27.0
LiveCodeBenchβ€”$0.10β€”
456JT-MINIChina Mobile
27.0
SciCodeβ€”β€”β€”
457Solar Open 100B (Reasoning)Upstage
27.0
SciCodeβ€”β€”β€”
458Mistral: Pixtral Large 2411Mistral AIMistral AI
26.0
LiveCodeBench131K tokensβ€”β€”
459Devstral Small (May '25)MistralMistral
26.0
LiveCodeBenchβ€”β€”β€”
460Qwen3.5 Omni FlashAlibabaAlibaba
26.0
SciCodeβ€”$0.10β€”
461Sarvam 105B (high)Sarvam
26.0
SciCodeβ€”$0.04β€”
462Devstral Small (Jul '25)MistralMistral
25.0
LiveCodeBench131K tokensβ€”β€”
463Mistral Small 3MistralMistral
25.0
LiveCodeBenchβ€”$0.10β€”
464Qwen2.5 Instruct 32BAlibabaAlibaba
25.0
LiveCodeBenchβ€”β€”β€”
465Granite 4.0 H SmallIBM
25.0
LiveCodeBenchβ€”$0.06β€”
466Grok BetaxAIxAI
24.0
LiveCodeBenchβ€”β€”β€”
467Mistral: SabaMistral AIMistral AI
24.0
SciCode33K tokens$0.20βœ…
468GPT-4o-mini (2024-07-18)OpenAIOpenAI
23.4
LiveCodeBench128K tokens$0.15β€”
469Llama 3.1 70B InstructMetaMeta
23.0
LiveCodeBench131K tokens$0.56βœ…
470Microsoft: Phi 4MicrosoftMicrosoft
23.0
LiveCodeBench16K tokens$0.13βœ…
471Qwen3 4B (Non-reasoning)AlibabaAlibaba
23.0
LiveCodeBenchβ€”β€”β€”
472DeepSeek R1 Distill Llama 8BDeepSeekDeepSeek
23.0
LiveCodeBenchβ€”β€”β€”
473Gemini 1.5 Flash-8BGoogleGoogle
22.0
LiveCodeBenchβ€”β€”β€”
474Gemini 2.0 Flash (experimental)GoogleGoogle
21.0
LiveCodeBenchβ€”β€”β€”
475Llama 3.2 Instruct 90B (Vision)MetaMeta
21.0
LiveCodeBenchβ€”β€”β€”
476Jamba Reasoning 3BAI21 Labs
21.0
LiveCodeBenchβ€”β€”β€”
477DeepHermes 3 - Mistral 24B Preview (Non-reasoning)Nous Research
20.0
LiveCodeBenchβ€”β€”β€”
478Llama 3 70B InstructMetaMeta
20.0
LiveCodeBench8K tokens$0.65βœ…
479Gemini 1.5 Flash (May '24)GoogleGoogle
20.0
LiveCodeBenchβ€”β€”β€”
480Qwen3 8B (Non-reasoning)AlibabaAlibaba
20.0
LiveCodeBenchβ€”$0.18β€”
481Gemma 4 E2B (Non-reasoning)GoogleGoogle
20.0
SciCodeβ€”β€”β€”
482Gemini 2.0 Flash-Lite (Feb '25)GoogleGoogle
19.0
LiveCodeBenchβ€”β€”β€”
483Hermes 3 - Llama-3.1 70BNous Research
19.0
LiveCodeBenchβ€”$0.70β€”
484Sarvam 30BSarvam
19.0
SciCodeβ€”$0.03β€”
485Gemini 2.0 Flash LiteGoogleGoogle
18.5
LiveCodeBench1.0M tokens$0.07β€”
486Gemini 2.0 Flash-Lite (Preview)GoogleGoogle
18.0
LiveCodeBenchβ€”β€”β€”
487Claude 3 SonnetAnthropicAnthropic
18.0
LiveCodeBenchβ€”$3.00β€”
488AI21: Jamba Large 1.7AI21 Labs
18.0
LiveCodeBench256K tokens$2.00βœ…
489Granite 4.0 MicroIBM
18.0
LiveCodeBench131K tokensβ€”βœ…
490Tri-21B-think PreviewTrillion Labs
18.0
SciCodeβ€”β€”β€”
491Mistral LargeMistral AIMistral AI
17.8
LiveCodeBench128K tokens$2.00βœ…
492Llama 3.1 Nemotron 70B InstructNvidiaNvidia
17.0
LiveCodeBench131K tokens$1.20βœ…
493Jamba 1.6 LargeAI21 Labs
17.0
LiveCodeBenchβ€”$2.00β€”
494Tri-21B-ThinkTrillion Labs
17.0
SciCodeβ€”β€”β€”
495Olmo 3.1 32B InstructAllenAI
17.0
SciCode66K tokensβ€”βœ…
496GLM-4.6V (Reasoning)Z.aiZ.ai
16.0
LiveCodeBenchβ€”$0.30β€”
497Qwen2 Instruct 72BAlibabaAlibaba
16.0
LiveCodeBenchβ€”β€”β€”
498Qwen: Qwen-TurboAlibabaAlibaba
16.0
LiveCodeBench131K tokens$0.05βœ…
499DeepSeek Coder V2 Lite InstructDeepSeekDeepSeek
16.0
LiveCodeBenchβ€”β€”β€”
500Mixtral 8x22B InstructMistralMistral
15.0
LiveCodeBenchβ€”β€”β€”
501Anthropic: Claude 3 HaikuAnthropicAnthropic
15.0
LiveCodeBench200K tokens$0.25β€”
502LFM2 8B A1BLiquid AI
15.0
LiveCodeBenchβ€”β€”β€”
503Mistral: Mixtral 8x22B InstructMistral AIMistral AI
14.8
LiveCodeBench66K tokens$2.00βœ…
504Mistral Small (Sep '24)MistralMistral
14.0
LiveCodeBenchβ€”$0.20β€”
505Jamba 1.5 LargeAI21 Labs
14.0
LiveCodeBenchβ€”$2.00β€”
506Gemma 3n E4B Instruct Preview (May '25)GoogleGoogle
14.0
LiveCodeBenchβ€”β€”β€”
507Qwen2.5 Coder Instruct 7B AlibabaAlibaba
13.0
LiveCodeBenchβ€”β€”β€”
508Phi-4 Multimodal InstructMicrosoftMicrosoft
13.0
LiveCodeBenchβ€”β€”β€”
509Granite 3.3 8B (Non-reasoning)IBM
13.0
LiveCodeBenchβ€”$0.03β€”
510Qwen3 1.7B (Non-reasoning)AlibabaAlibaba
13.0
LiveCodeBenchβ€”β€”β€”
511Molmo2-8BAllen Institute for AI
13.0
SciCodeβ€”β€”β€”
512Cohere: Command R+ (08-2024)CohereCohere
12.2
LiveCodeBench128K tokens$2.50β€”
513Command-R+ (Apr '24)CohereCohere
12.0
LiveCodeBenchβ€”$3.00β€”
514Gemini 1.0 ProGoogleGoogle
12.0
LiveCodeBenchβ€”β€”β€”
515Phi-3 Mini Instruct 3.8BMicrosoftMicrosoft
12.0
LiveCodeBenchβ€”β€”β€”
516Granite 4.0 H 1BIBM
12.0
LiveCodeBenchβ€”β€”β€”
517Qwen3 0.6B (Reasoning)AlibabaAlibaba
12.0
LiveCodeBenchβ€”β€”β€”
518OpenChat 3.5 (1210)OpenChat
12.0
LiveCodeBenchβ€”β€”β€”
519Mistral Small (Feb '24)MistralMistral
11.0
LiveCodeBenchβ€”$0.15β€”
520Llama 3.2 11B Vision InstructMetaMeta
11.0
LiveCodeBench131K tokens$0.34βœ…
521LFM2-24B-A2BLiquidAI
11.0
SciCode33K tokensβ€”βœ…
522Llama 3 8B InstructMetaMeta
10.0
LiveCodeBench8K tokens$0.04βœ…
523Mistral MediumMistralMistral
10.0
LiveCodeBenchβ€”$1.50β€”
524Llama 2 Chat 13BMetaMeta
10.0
LiveCodeBenchβ€”β€”β€”
525LFM 40BLiquid AI
10.0
LiveCodeBenchβ€”β€”β€”
526Gemma 3n E2B InstructGoogleGoogle
10.0
LiveCodeBenchβ€”β€”β€”
527Llama 2 Chat 70BMetaMeta
10.0
LiveCodeBenchβ€”β€”β€”
528DBRX InstructDatabricks
9.0
LiveCodeBenchβ€”β€”β€”
529DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning)Nous Research
9.0
LiveCodeBenchβ€”β€”β€”
530Llama 3.2 3B InstructMetaMeta
8.0
LiveCodeBench80K tokensβ€”βœ…
531LFM2 2.6BLiquid AI
8.0
LiveCodeBenchβ€”β€”β€”
532LFM2.5-8B-A1BLiquid AI
8.0
SciCodeβ€”β€”β€”
533Jamba 1.6 MiniAI21 Labs
7.0
LiveCodeBenchβ€”$0.20β€”
534OLMo 2 32BAllen Institute for AI
7.0
LiveCodeBenchβ€”β€”β€”
535DeepSeek R1 Distill Qwen 1.5BDeepSeekDeepSeek
7.0
LiveCodeBenchβ€”β€”β€”
536Qwen3 0.6B (Non-reasoning)AlibabaAlibaba
7.0
LiveCodeBenchβ€”β€”β€”
537Mistral: Mixtral 8x7B InstructMistral AIMistral AI
7.0
LiveCodeBench33K tokens$0.45βœ…
538Jamba 1.7 MiniAI21 Labs
6.0
LiveCodeBenchβ€”β€”β€”
539Jamba 1.5 MiniAI21 Labs
6.0
LiveCodeBenchβ€”$0.20β€”
540Apertus 70B InstructSwiss AI Initiative
6.0
SciCodeβ€”$0.82β€”
541Granite 4.0 1BIBM
5.0
LiveCodeBenchβ€”β€”β€”
542Command-R (Mar '24)CohereCohere
5.0
LiveCodeBenchβ€”$0.50β€”
543Mistral 7B InstructMistralMistral
5.0
LiveCodeBenchβ€”$0.25β€”
544OLMo 2 7BAllen Institute for AI
4.0
LiveCodeBenchβ€”β€”β€”
545Molmo 7B-DAllen Institute for AI
4.0
LiveCodeBenchβ€”β€”β€”
546MiniCPM5-1B (Non-reasoning)OpenBMB
4.0
SciCodeβ€”β€”β€”
547Gemma 4 E4B (Non-reasoning)GoogleGoogle
4.0
SciCodeβ€”$0.02β€”
548Tiny Aya GlobalCohereCohere
4.0
SciCodeβ€”β€”β€”
549LFM2.5-1.2B-ThinkingLiquid AI
4.0
SciCodeβ€”β€”β€”
550Apertus 8B InstructSwiss AI Initiative
4.0
SciCodeβ€”$0.10β€”
551LFM2.5-VL-1.6BLiquid AI
3.0
SciCodeβ€”β€”β€”
552LFM2 1.2BLiquid AI
2.0
LiveCodeBenchβ€”β€”β€”
553Granite 4.0 H 350MIBM
2.0
LiveCodeBenchβ€”β€”β€”
554Llama 3.2 1B InstructMetaMeta
2.0
LiveCodeBench60K tokensβ€”βœ…
555Granite 4.0 350MIBM
2.0
LiveCodeBenchβ€”β€”β€”
556Gemma 3 1B InstructGoogleGoogle
2.0
LiveCodeBenchβ€”β€”β€”
557LFM2.5-1.2B-InstructLiquid AI
2.0
SciCodeβ€”β€”β€”
558Gemma 3 270MGoogleGoogle
0.0
LiveCodeBenchβ€”β€”β€”
559Llama 2 Chat 7BMetaMeta
0.0
LiveCodeBenchβ€”$0.05β€”

+ 212 models without coding benchmarks available.View all models

Complete Guide: AI for Programming in 2026

The State of AI for Code in 2026

Artificial intelligence has fundamentally changed software development. In 2026, large language models (LLMs) can generate working code in dozens of languages, fix bugs in production codebases, and build complete applications from plain-English descriptions. SWE-bench β€” the most rigorous coding benchmark β€” evaluates models on real software engineering tasks pulled from GitHub issues.

SWE-bench: The Gold Standard

SWE-bench (Software Engineering Benchmark) is widely considered the gold standard for evaluating LLM coding ability. Unlike academic benchmarks like HumanEval (which tests isolated functions), SWE-bench presents real issues from popular repositories such as Django, Flask, scikit-learn, and requests. The model must understand the project context, locate the relevant files, and generate a patch that resolves the bug β€” mirroring the actual workflow of a professional developer.

The β€œVerified” variant (SWE-bench Verified) is curated by human engineers to ensure every task has a clear, verifiable solution. Scores on this benchmark correlate strongly with real-world coding performance, making it the single most informative metric when choosing an AI coding assistant.

HumanEval and LiveCodeBench

HumanEval, created by OpenAI, tests a model's ability to generate Python functions from docstrings. It is simpler than SWE-bench but useful for gauging basic code fluency. LiveCodeBench raises the bar by using problems that are refreshed regularly, reducing the risk of data contamination β€” a concern when a model may have seen the answers during training.

How to Choose the Best AI Model for Code

The right model depends on your specific use case. For real-time code autocomplete (Cursor, Copilot), speed and latency matter more than peak benchmark scores β€” lighter mini and flash-class models often deliver the best speed-to-quality ratio. For full project generation or complex debugging, frontier models like Claude Fable 5, GPT-5.5 and Gemini 3.5-class systems are better suited, despite higher costs.

Teams with strict data control requirements (compliance, security) should consider open-source models like DeepSeek Coder, Code Llama, and StarCoder, which can be deployed on-premises with competitive performance. The trade-off between proprietary and open-source involves cost, latency, privacy, and quality considerations.

AI-Powered Coding Tools

The leading AI-assisted development tools in 2026 include Cursor, GitHub Copilot, Windsurf and agentic IDE workflows that let you swap between multiple frontier models. Each tool uses different models under the hood, and the quality of generated code depends directly on the LLM powering it.

Trends for 2026 and Beyond

The most significant trends in AI for code include autonomous software engineering agents (that solve complex tasks without supervision), automated test generation, intelligent refactoring, and native CI/CD pipeline integration. The frontier is shifting from β€œcode assistant” to β€œautonomous engineer”, with models increasingly capable of navigating large codebases and making architectural decisions.

Frequently Asked Questions

What is the best AI for coding?

In 2026, the top models on coding benchmarks are GPT-5.6 Sol (xhigh), OpenAI: GPT-5.6 Sol (batch), Claude Opus 5. The best choice depends on your use case: code autocomplete, full project generation, debugging, or code review.

ChatGPT or Claude for code?

Today the strongest current options usually cluster around Claude Fable 5 / Opus 4.8, GPT-5.5 and Gemini 3.5-class models. Claude tends to shine in long-context refactors; GPT is strong in fast generation and tool use. Test with your own repository and workflow.

What is SWE-bench?

SWE-bench (Software Engineering Benchmark) evaluates how well models can resolve real issues from open-source GitHub repositories. It is considered the most realistic coding benchmark because it tests bug resolution in real projects, not academic exercises.

Which free LLMs are good for coding?

Open-source models like DeepSeek Coder, Qwen Coder, and Code Llama offer excellent coding performance with no API cost. You can run them locally with Ollama or access them for free on platforms like Together AI and Groq.

What coding benchmarks matter most?

SWE-bench Verified is the gold standard for real-world coding ability. HumanEval tests basic function generation, while LiveCodeBench uses regularly updated problems to reduce data contamination. For a complete picture, look at all three.

Explore Other Categories