Melhor IA para Código em 2026Atualizado por AA Coding Index

Qual IA programa melhor em 2026? Ranking de 555 modelos usando o AA Coding Index como sinal principal, com fallback para LiveCodeBench e SciCode. A lógica prioriza modelos atuais e evita distorções por picos isolados.

Sincronizado: 13 de setembro de 2026 555 modelos com benchmarks de código

Casos de Uso

Autocompletar Código

Sugestões inline enquanto você digita. Ideal para IDEs como Cursor e VS Code.

Top modelos: Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)

Geração de Código

Criar funções, classes e projetos completos a partir de descrições em linguagem natural.

Top modelos: Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)

Debug e Code Review

Identificar bugs, sugerir correções e revisar pull requests automaticamente.

Top modelos: Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)

Ranking de Coding — Top Modelos

#ModeloEmpresaScore de CodingBenchmarkContextPreço InputOpen Source
🥇Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)AnthropicAnthropic
81.6
AA Coding Index1.0M tokens$10.00
🥈Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback)AnthropicAnthropic
80.7
AA Coding Index$10.00
🥉Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback)AnthropicAnthropic
79.1
AA Coding Index$10.00
4GPT-5.6 Sol (xhigh)OpenAIOpenAI
78.3
AA Coding Index$4.00
5Claude Opus 5AnthropicAnthropic
78.0
AA Coding Index1.0M tokens$5.00
6GPT-5.6 Sol (max)OpenAIOpenAI
77.4
AA Coding Index1.1M tokens$4.00
7GPT-5.6 Sol (high)OpenAIOpenAI
77.2
AA Coding Index$4.00
8GPT-6 Astra (high)OpenAIOpenAI
77.1
AA Coding Index$10.00
9Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback)AnthropicAnthropic
77.1
AA Coding Index$10.00
10Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)AnthropicAnthropic
77.0
AA Coding Index$5.00
11GPT-6 Astra (max)OpenAIOpenAI
76.9
AA Coding Index1.1M tokens$10.00
12SpaceXAI: Grok 4.6xAIxAI
76.8
AA Coding Index500K tokens$2.00
13GPT-5.6 Terra (max)OpenAIOpenAI
76.7
AA Coding Index1.1M tokens$2.00
14GPT-6 Astra (medium)OpenAIOpenAI
76.7
AA Coding Index$10.00
15Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)AnthropicAnthropic
76.5
AA Coding Index1.0M tokens$10.00
16Muse Spark 1.3 (xhigh)MetaMeta
76.5
AA Coding Index$1.25
17Claude Opus 5 (Adaptive Reasoning, High Effort)AnthropicAnthropic
76.5
AA Coding Index1.0M tokens$5.00
18GPT-5.6 Sol (medium)OpenAIOpenAI
76.3
AA Coding Index$4.00
19Gemini 3.8 Flash (high)GoogleGoogle
76.3
AA Coding Index1.0M tokens$0.75
20GPT-6 Astra (Non-reasoning)OpenAIOpenAI
76.2
AA Coding Index$10.00
21Kimi K3Moonshot AIMoonshot AI
76.2
AA Coding Index1.0M tokens$3.00
22Gemini 3.7 Flash (high)GoogleGoogle
76.1
AA Coding Index1.0M tokens$0.75
23GPT-6 Astra (xhigh)OpenAIOpenAI
75.9
AA Coding Index$10.00
24Grok 4.6 (xhigh)SpaceXAISpaceXAI
75.9
AA Coding Index$2.00
25Muse Spark 1.3 (max)MetaMeta
75.8
AA Coding Index$1.25
26GPT-6 Astra (low)OpenAIOpenAI
75.7
AA Coding Index$10.00
27Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback)AnthropicAnthropic
75.2
AA Coding Index$10.00
28GPT-5.5OpenAIOpenAI
74.9
AA Coding Index1.1M tokens$5.00
29GLM-5.3Z.aiZ.ai
74.8
AA Coding Index$1.40
30Grok 4.6 (medium)SpaceXAISpaceXAI
74.4
AA Coding Index$2.00
31Claude Opus 4.8 (Adaptive Reasoning, Max Effort)AnthropicAnthropic
74.3
AA Coding Index1.0M tokens$5.00
32Claude Opus 5 (Adaptive Reasoning, Medium Effort)AnthropicAnthropic
74.3
AA Coding Index$5.00
33Gemini 3.8 Flash (medium)GoogleGoogle
74.1
AA Coding Index$0.75
34Claude Opus 4.7AnthropicAnthropic
73.6
AA Coding Index1.0M tokens$5.00
35Gemini 3.8 Flash (low)GoogleGoogle
73.5
AA Coding Index$0.75
36Qwen3.8-Flash-NextAlibabaAlibaba
73.1
AA Coding Index$0.15
37SpaceXAI: Grok 4.5xAIxAI
72.4
AA Coding Index500K tokens$2.00
38Muse Spark 1.2 (xhigh)MetaMeta
72.2
AA Coding Index$1.25
39Kimi K3 (low)KimiKimi
72.0
AA Coding Index1.0M tokens$3.00
40Qwen: Qwen3.8 2.4T A95BAlibabaAlibaba
71.9
AA Coding Index1.0M tokens$2.00
41Qwen: Qwen3.8 MaxAlibabaAlibaba
71.8
AA Coding Index1.0M tokens$2.00
42GPT-5.5 (high)OpenAIOpenAI
71.6
AA Coding Index$5.00
43Claude Sonnet 5AnthropicAnthropic
71.5
AA Coding Index1.0M tokens$2.00
44GLM-5.3-FlashZ AI
71.5
AA Coding Index$0.15
45Gemini 3.7 Flash (medium)GoogleGoogle
71.5
AA Coding Index$0.75
46GPT-5.5 (medium)OpenAIOpenAI
71.5
AA Coding Index$5.00
47GPT-5.6 Luna (max)OpenAIOpenAI
71.4
AA Coding Index1.1M tokens$0.20
48Muse Spark 1.1 (xhigh)MetaMeta
71.3
AA Coding Index$1.25
49GPT-5.4OpenAIOpenAI
71.1
AA Coding Index1.1M tokens$2.50
50Gemini 3.7 Flash (low)GoogleGoogle
71.0
AA Coding Index$0.75
51GPT-5.6 Terra (xhigh)OpenAIOpenAI
70.6
AA Coding Index$2.00
52Google: Gemini 3.5 FlashGoogleGoogle
70.1
AA Coding Index1.0M tokens$1.50
53GPT-5.6 Sol (low)OpenAIOpenAI
69.7
AA Coding Index$4.00
54Gemini 3.6 Flash (high)GoogleGoogle
69.2
AA Coding Index1.0M tokens$0.75
55DeepSeek-V4-FlashDeepSeekDeepSeek
69.1
AA Coding Index1.0M tokens$0.44
56Gemini 3.1 Pro PreviewGoogleGoogle
68.8
AA Coding Index1.0M tokens$2.00
57DeepSeek V4 ProDeepSeekDeepSeek
68.8
AA Coding Index1.0M tokens$1.32
58GLM-5.2 (Non-reasoning)Z AI
68.8
AA Coding Index$1.40
59GPT-5.6 Luna (xhigh)OpenAIOpenAI
68.6
AA Coding Index$0.20
60Qwen: Qwen3.8 27BAlibabaAlibaba
68.1
AA Coding Index1.0M tokens$0.21
61Qwen3.8 27B (medium)AlibabaAlibaba
68.1
AA Coding Index$0.50
62GPT-5.6 Terra (high)OpenAIOpenAI
67.1
AA Coding Index$2.00
63Claude Opus 5 (Adaptive Reasoning, Low Effort)AnthropicAnthropic
66.9
AA Coding Index$5.00
64Claude Sonnet 5 (Non-reasoning, High Effort)AnthropicAnthropic
66.4
AA Coding Index1.0M tokens$2.00
65Grok 4.6 (low)SpaceXAISpaceXAI
66.3
AA Coding Index$2.00
66Qwen3.7 MaxAlibabaAlibaba
66.0
AA Coding Index$2.50
67GPT-5.6 Sol (Non-reasoning)OpenAIOpenAI
65.1
AA Coding Index$4.00
68DeepSeek V4 Flash Vision (Reasoning, Max Effort)DeepSeekDeepSeek
65.0
AA Coding Index$0.44
69GPT-5.6 Terra (medium)OpenAIOpenAI
64.7
AA Coding Index$2.00
70Motif 3Motif Technologies
63.5
AA Coding Index
71GPT-5.6 Luna (high)OpenAIOpenAI
63.3
AA Coding Index$0.20
72Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort)AnthropicAnthropic
63.0
AA Coding Index$3.00
73Agnes 2.5 Pro BetaSapiens AI
62.3
AA Coding Index$0.10
74Motif 3 (Beta)Motif Technologies
62.0
AA Coding Index
75Kimi K2.6 (Non-reasoning)KimiKimi
61.8
AA Coding Index$0.95
76K2 Horizon 375B A23BMBZUAI Institute of Foundation Models
61.5
AA Coding Index
77Quasar 438BMultiverse Computing
61.2
AA Coding Index$0.60
78GPT-5.5 (low)OpenAIOpenAI
60.9
AA Coding Index$5.00
79Kimi K2.7 CodeKimiKimi
60.8
AA Coding Index$0.95
80Apodex 1.1Apodex
60.8
AA Coding Index$0.30
81Xiaomi: MiMo-V2.5-ProXiaomi
60.2
AA Coding Index1.0M tokens$0.43
82Kwaipilot: KAT-Coder-Pro V2Kwaipilot
59.5
AA Coding Index256K tokens$0.30
83DeepSeek V4 Pro (Non-reasoning)DeepSeekDeepSeek
59.4
AA Coding Index$0.43
84DeepSeek V4 Pro (Reasoning, High Effort)DeepSeekDeepSeek
59.4
AA Coding Index$0.43
85Nex-N2-ProNex AGI
59.1
AA Coding Index262K tokens$0.50
86Hy3-preview (Reasoning)Tencent
58.8
AA Coding Index262K tokens$0.14
87Agnes 2.5 Pro AlphaSapiens AI
58.8
AA Coding Index$0.45
88Muse SparkMetaMeta
58.6
AA Coding Index
89MiniMax-M3MiniMax
58.6
AA Coding Index1.0M tokens$0.30
90Qwen3.8 27B (low)AlibabaAlibaba
58.2
AA Coding Index$0.50
91GPT-5.6 Terra (low)OpenAIOpenAI
58.1
AA Coding Index$2.00
92Ling-3.0-flash-VLInclusionAI
57.0
AA Coding Index
93MiMo-V2.5Xiaomi
56.8
AA Coding Index$0.14
94GPT-5.5 (Non-reasoning)OpenAIOpenAI
56.5
AA Coding Index$5.00
95DeepSeek V4 Flash (Reasoning, High Effort)DeepSeekDeepSeek
56.2
AA Coding Index$0.13
96GPT-5.4 MiniOpenAIOpenAI
56.1
AA Coding Index400K tokens$0.75
97GPT-5.4 NanoOpenAIOpenAI
56.1
AA Coding Index400K tokens$0.20
98GPT-5.4 nano (Non-Reasoning)OpenAIOpenAI
56.1
AA Coding Index$0.20
99Qwen3.7 PlusAlibabaAlibaba
55.9
AA Coding Index$0.40
100Z.ai: GLM 5.1Z.aiZ.ai
55.8
AA Coding Index203K tokens$1.20
101Qwen: Qwen3.6 PlusAlibabaAlibaba
54.5
AA Coding Index1.0M tokens$0.50
102Qwen: Qwen3.6 27BAlibabaAlibaba
53.7
AA Coding Index262K tokens$0.30
103Qwen3.6 27B (Non-reasoning)AlibabaAlibaba
53.7
AA Coding Index$0.60
104Inkling SmallThinking Machines
52.9
AA Coding Index$0.30
105Solar Pro 4Upstage
52.7
AA Coding Index524K tokens$0.30
106MiniMax: MiniMax M2.7MiniMax
52.6
AA Coding Index197K tokens$0.30
107JT-4.1 Flash 236B A21BChina Mobile
52.4
AA Coding Index
108GPT-5.6 Terra (Non-reasoning)OpenAIOpenAI
52.3
AA Coding Index$2.00
109Claude 4.5 Sonnet (Reasoning)AnthropicAnthropic
52.1
AA Coding Index$3.00
110InklingThinking Machines
52.1
AA Coding Index$1.00
111Grok Build 0.1 0616xAIxAI
51.5
AA Coding Index$1.00
112GPT-5.6 Luna (medium)OpenAIOpenAI
50.7
AA Coding Index$0.20
113Ling 3.0 FlashInclusionAI
50.6
AA Coding Index$0.07
114MiMo-V2-Flash (Reasoning)Xiaomi
49.8
AA Coding Index262K tokens$0.10
115GPT-5.1OpenAIOpenAI
49.4
AA Coding Index400K tokens$1.25
116Gemini 3.5 Flash-LiteGoogleGoogle
49.3
AA Coding Index1.0M tokens$0.30
117Nemotron 3 Ultra 550B A55B (Reasoning)NvidiaNvidia
49.3
AA Coding Index1.0M tokens$0.60
118Muse Glimmer (high)MetaMeta
49.0
AA Coding Index$0.35
119Qwen: Qwen3.5 397B A17BAlibabaAlibaba
48.2
AA Coding Index262K tokens$0.60
120Mistral Medium 3.5MistralMistral
46.9
AA Coding Index262K tokens$1.50
121MoonshotAI: Kimi K2.5MoonshotAIMoonshotAI
46.8
AA Coding Index262K tokens$0.60
122Kimi K2.5 (Non-reasoning)KimiKimi
46.8
AA Coding Index$0.60
123Gemini 2.5 Pro Preview (Mar' 25)GoogleGoogle
46.7
AA Coding Index
124GLM-4.6 (Reasoning)Z.aiZ.ai
45.8
AA Coding Index$0.57
125Qwen: Qwen3.5-122B-A10BAlibabaAlibaba
45.7
AA Coding Index262K tokens$0.40
126LongCat 2.0LongCat
45.3
AA Coding Index$0.30
127GLM-4.7 (Non-reasoning)Z AI
45.3
AA Coding Index$0.60
128Solar Open2 250BUpstage
45.0
AA Coding Index
129Qwen3.8 27B (Non-reasoning)AlibabaAlibaba
44.6
AA Coding Index$0.50
130DeepSeek V3.2DeepSeekDeepSeek
44.2
AA Coding Index164K tokens$0.28
131GPT-5.6 Luna (low)OpenAIOpenAI
44.2
AA Coding Index$0.20
132DeepSeek V3.2 (Reasoning)DeepSeekDeepSeek
44.2
AA Coding Index$0.28
133Claude 4.5 Haiku (Reasoning)AnthropicAnthropic
43.9
AA Coding Index$1.00
134DeepSeek V3.1 TerminusDeepSeekDeepSeek
43.5
AA Coding Index164K tokens$0.27
135DeepSeek V3.1 Terminus (Reasoning)DeepSeekDeepSeek
43.5
AA Coding Index$1.64
136Gemma 4 31BGoogleGoogle
43.4
AA Coding Index262K tokens
137Qwen3.5 122B A10B (Non-reasoning)AlibabaAlibaba
43.3
AA Coding Index$0.40
138Ring-2.6-1TInclusionAI
42.8
AA Coding Index$0.30
139Grok 4.3 (low)SpaceXAISpaceXAI
42.2
AA Coding Index$1.25
140SpaceXAI: Grok 4.3xAIxAI
42.2
AA Coding Index1.0M tokens$1.25
141Qwen: Qwen3.6 35B A3BAlibabaAlibaba
41.9
AA Coding Index262K tokens$0.38
142K-EXAONE 2.0 0803LG AI Research
40.6
AA Coding Index
143o1OpenAIOpenAI
39.7
AA Coding Index200K tokens$15.00
144Step 3.7 FlashStepFun
39.6
AA Coding Index$0.20
145GPT-5.5 Instant (June 2026)OpenAIOpenAI
39.4
AA Coding Index$5.00
146GPT-5.6 Luna (Non-reasoning)OpenAIOpenAI
39.3
AA Coding Index$0.20
147Gemma 4 26B A4B GoogleGoogle
39.3
AA Coding Index262K tokens$0.12
148A.X-K2SK Telecom
38.8
AA Coding Index
149GPT-5OpenAIOpenAI
37.8
AA Coding Index400K tokens$1.25
150NVIDIA Nemotron 3 Super 120B A12B (Reasoning)NvidiaNvidia
37.7
AA Coding Index1.0M tokens$0.19
151Claude 4 Sonnet (Reasoning)AnthropicAnthropic
37.6
AA Coding Index$3.00
152Qwen: Qwen3.5-35B-A3BAlibabaAlibaba
37.0
AA Coding Index262K tokens$0.25
153Qwen3.5 35B A3B (Non-reasoning)AlibabaAlibaba
37.0
AA Coding Index$0.25
154North Mini CodeCohereCohere
36.5
AA Coding Index
155Claude 3.7 Sonnet (thinking)AnthropicAnthropic
36.4
AA Coding Index200K tokens
156Qwen: Qwen3 Coder NextAlibabaAlibaba
36.2
AA Coding Index262K tokens$0.35
157Grok 4.3 (Non-reasoning)SpaceXAISpaceXAI
35.2
AA Coding Index$1.25
158Gemini 3.1 Flash Lite PreviewGoogleGoogle
34.7
AA Coding Index1.0M tokens$0.25
159Nova 2.0 Pro Preview (medium)AmazonAmazon
34.0
AA Coding Index$1.25
160o1-previewOpenAIOpenAI
34.0
AA Coding Index$16.50
161Gemini 2.5 ProGoogleGoogle
33.3
AA Coding Index1.0M tokens$1.25
162Gemma 4 31B (Non-reasoning)GoogleGoogle
33.2
AA Coding Index$0.14
163G9v3-39A5BAI9Stars
33.1
AA Coding Index
164K-EXAONE (Non-reasoning)LG AI Research
32.1
AA Coding Index
165Devstral 2MistralMistral
31.3
AA Coding Index
166Inception: Mercury 2Inception
31.1
AA Coding Index128K tokens$0.25
167Gemma 4 12B (Reasoning)GoogleGoogle
31.0
AA Coding Index$0.10
168gpt-oss-120bOpenAIOpenAI
30.4
AA Coding Index131K tokens$0.15
169Claude 3.5 Sonnet (Oct '24)AnthropicAnthropic
30.2
AA Coding Index$3.00
170Granite 4.2 30BIBM
29.9
AA Coding Index$0.16
171Devstral Small 2MistralMistral
29.3
AA Coding Index
172Qwen3.5 9B (Reasoning)AlibabaAlibaba
28.7
AA Coding Index$0.14
173Qwen3.6 35B A3B (Non-reasoning)AlibabaAlibaba
28.1
AA Coding Index$0.38
174Command A+CohereCohere
27.8
AA Coding Index
175Nemotron 3.5 LightningNVIDIANVIDIA
26.8
AA Coding Index262K tokens$0.06
176Mistral Small 4 (Non-reasoning)MistralMistral
26.6
AA Coding Index$0.15
177Ling 3.0 TinyInclusionAI
26.5
AA Coding Index
178Mistral Small 3.1MistralMistral
26.3
AA Coding Index$0.10
179Claude 3.5 Sonnet (June '24)AnthropicAnthropic
26.0
AA Coding Index$3.00
180Nova 2.0 Pro Preview (low)AmazonAmazon
25.9
AA Coding Index$1.25
181Arcee AI: Trinity Large ThinkingArcee AI
25.8
AA Coding Index262K tokens$0.25
182Gemini 2.0 Pro Experimental (Feb '25)GoogleGoogle
25.5
AA Coding Index
183Nemotron Cascade 2 30B A3BNvidiaNvidia
25.3
AA Coding Index
184Ling 2.6 FlashInclusion AI
25.3
AA Coding Index$0.10
185DeepSeek R1 (Jan '25)DeepSeekDeepSeek
24.6
AA Coding Index$2.00
186OpenAI: GPT-4o (2024-05-13)OpenAIOpenAI
24.2
AA Coding Index128K tokens$5.00
187Gemini 2.0 Flash Thinking Experimental (Jan '25)GoogleGoogle
24.1
AA Coding Index
188Gemini 1.5 Pro (Sep '24)GoogleGoogle
23.6
AA Coding Index
189EXAONE 4.5 33B (Non-reasoning)LG AI Research
23.6
AA Coding Index
190Qwen3.5 9B (Non-reasoning)AlibabaAlibaba
23.5
AA Coding Index$0.17
191HyperNova 60B 2605Multiverse Computing
23.2
AA Coding Index$0.04
192Nova 2.0 Lite (high)AmazonAmazon
23.0
AA Coding Index$0.30
193DeepSeek V3DeepSeekDeepSeek
23.0
AA Coding Index164K tokens$0.32
194Qwen3.5 4B (Reasoning)AlibabaAlibaba
22.6
AA Coding Index$0.03
195Granite 4.2 8BIBM
22.4
AA Coding Index$0.06
196Qwen: Qwen3 235B A22B Instruct 2507AlibabaAlibaba
22.1
AA Coding Index262K tokens$0.70
197Qwen3 235B A22B 2507 (Reasoning)AlibabaAlibaba
22.1
AA Coding Index$0.23
198GPT-4 TurboOpenAIOpenAI
21.5
AA Coding Index128K tokens$10.00
199Magistral Medium 1.2Mistral AIMistral AI
21.3
AA Coding Index$2.00
200DeepSeek V3 0324DeepSeekDeepSeek
21.2
AA Coding Index$0.84
201gpt-oss-120b (low)OpenAIOpenAI
21.2
AA Coding Index131K tokens$0.15
202K2 Think V2MBZUAI Institute of Foundation Models
21.0
AA Coding Index
203gpt-oss-20bOpenAIOpenAI
20.7
AA Coding Index131K tokens$0.06
204Mistral: Mistral Medium 3.1Mistral AIMistral AI
20.5
AA Coding Index131K tokens$0.40
205Qwen3.5 4B (Non-reasoning)AlibabaAlibaba
20.3
AA Coding Index$0.03
206GPT-4.1 MiniOpenAIOpenAI
20.2
AA Coding Index1.0M tokens$0.40
207Mistral Large 3MistralMistral
20.1
AA Coding Index$0.50
208Gemini 1.5 Pro (May '24)GoogleGoogle
19.8
AA Coding Index
209DiffusionGemma 26B A4BGoogleGoogle
19.7
AA Coding Index
210Claude 3 OpusAnthropicAnthropic
19.5
AA Coding Index$15.00
211Gemini 1.0 UltraGoogleGoogle
17.6
AA Coding Index
212Granite 4.2 3BIBM
17.5
AA Coding Index$0.03
213Qwen3 Next 80B A3B (Reasoning)AlibabaAlibaba
17.4
AA Coding Index$0.15
214o3 Mini HighOpenAIOpenAI
16.3
AA Coding Index200K tokens$1.10
215Llama 4 MaverickMetaMeta
16.3
AA Coding Index1.0M tokens$0.26
216Solar Pro 3Upstage
16.2
AA Coding Index128K tokens$0.15
217Claude 3.5 HaikuAnthropicAnthropic
15.9
AA Coding Index200K tokens
218GPT-5 MiniOpenAIOpenAI
15.6
AA Coding Index400K tokens$0.25
219Qwen3 32B (Reasoning)AlibabaAlibaba
15.3
AA Coding Index$0.16
220Magistral Small 1.2MistralMistral
14.7
AA Coding Index$0.50
221MiniCPM5-2BOpenBMB
14.5
AA Coding Index
222NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)NvidiaNvidia
14.4
AA Coding Index$0.05
223Ministral 3 14BMistralMistral
14.4
AA Coding Index$0.20
224Celeris-1Unknown
14.4
AA Coding Index$0.20
225Claude 2.1AnthropicAnthropic
14.0
AA Coding Index
226Qwen3 14B (Reasoning)AlibabaAlibaba
13.8
AA Coding Index$0.35
227Nemotron 3 Nano Omni 30B A3B ReasoningNvidiaNvidia
13.8
AA Coding Index$0.09
228OpenAI: GPT-4OpenAIOpenAI
13.1
AA Coding Index8K tokens$30.00
229Claude 2.0AnthropicAnthropic
12.9
AA Coding Index
230Mistral Small 3.2MistralMistral
12.5
AA Coding Index$0.10
231Qwen3 30B A3B 2507 (Reasoning)AlibabaAlibaba
12.1
AA Coding Index$0.20
232Llama 3.3 70B InstructMetaMeta
11.9
AA Coding Index131K tokens$0.66
233OpenAI: GPT-4o-miniOpenAIOpenAI
11.4
AA Coding Index128K tokens$0.15
234GPT-4.1 NanoOpenAIOpenAI
11.1
AA Coding Index1.0M tokens$0.10
235GPT-3.5 TurboOpenAIOpenAI
10.7
AA Coding Index$0.50
236Granite 4.1 30BIBM
10.4
AA Coding Index
237Gemma 3 27BGoogleGoogle
10.1
AA Coding Index131K tokens
238G9v3-3BAI9Stars
9.9
AA Coding Index
239Ministral 3 8BMistralMistral
9.7
AA Coding Index$0.15
240Nanbeige4.1-3BNanbeige
9.6
AA Coding Index
241Granite 4.1 8BIBM
9.5
AA Coding Index$0.05
242Gemma 4 E4B (Reasoning)GoogleGoogle
9.4
AA Coding Index$0.02
243Qwen3 8B (Reasoning)AlibabaAlibaba
9.0
AA Coding Index$0.18
244Llama 4 ScoutMetaMeta
8.2
AA Coding Index1.3M tokens$0.19
245NVIDIA Nemotron 3 Nano 4BNvidiaNvidia
8.0
AA Coding Index
246Claude InstantAnthropicAnthropic
7.8
AA Coding Index
247LFM2.5-2.6B
7.7
AA Coding Index
248Gemma 4 E2B (Reasoning)GoogleGoogle
7.2
AA Coding Index
249Gemma 3 12BGoogleGoogle
5.8
AA Coding Index131K tokens
250Llama 3.1 8B InstructMetaMeta
5.4
AA Coding Index16K tokens$0.02
251Ministral 3 3BMistralMistral
4.8
AA Coding Index$0.10
252Granite 4.1 3BIBM
4.7
AA Coding Index
253PALM-2GoogleGoogle
4.6
AA Coding Index
254Phi-4 Mini InstructMicrosoftMicrosoft
3.8
AA Coding Index
255Gemma 3n E4B InstructGoogleGoogle
3.2
AA Coding Index$0.06
256Qwen3.5 2B (Reasoning)AlibabaAlibaba
2.9
AA Coding Index
257Gemma 3 4BGoogleGoogle
2.7
AA Coding Index131K tokens
258Qwen3.5 2B (Non-reasoning)AlibabaAlibaba
2.4
AA Coding Index
259Qwen3.5 0.8B (Non-reasoning)AlibabaAlibaba
1.2
AA Coding Index
260MiniCPM-V 4.6 1.3BOpenBMB
0.7
AA Coding Index
261Gemini 3 Pro Preview (high)GoogleGoogle
92.0
LiveCodeBench$2.00
262Gemini 3 Flash Preview (Reasoning)GoogleGoogle
91.0
LiveCodeBench$0.50
263DeepSeek V3.2 SpecialeDeepSeekDeepSeek
90.0
LiveCodeBench164K tokens
264GPT-5.2OpenAIOpenAI
89.0
LiveCodeBench400K tokens$1.75
265Claude Opus 4.5 (Reasoning)AnthropicAnthropic
87.0
LiveCodeBench$5.00
266MiMo-V2-Flash (Feb 2026)Xiaomi
86.8
LiveCodeBench
267Gemini 3 Pro Preview (low)GoogleGoogle
86.0
LiveCodeBench$2.00
268o4 MiniOpenAIOpenAI
86.0
LiveCodeBench200K tokens$1.10
269DeepSeek V3.2 Exp (Reasoning)DeepSeekDeepSeek
86.0
LiveCodeBench$0.28
270o4 Mini HighOpenAIOpenAI
85.9
LiveCodeBench200K tokens$1.10
271Kimi K2 ThinkingKimiKimi
85.0
LiveCodeBench262K tokens$0.60
272GPT-5.1-CodexOpenAIOpenAI
85.0
LiveCodeBench400K tokens$1.25
273GPT-5.1-Codex-MiniOpenAIOpenAI
84.0
LiveCodeBench400K tokens$0.25
274GPT-5 CodexOpenAIOpenAI
84.0
LiveCodeBench400K tokens$1.25
275Grok 4 FastxAIxAI
83.0
LiveCodeBench2.0M tokens$0.20
276MiniMax-M2MiniMax
83.0
LiveCodeBench205K tokens$0.30
277Grok 4xAIxAI
82.0
LiveCodeBench256K tokens$3.00
278Grok 4.1 FastxAIxAI
82.0
LiveCodeBench2.0M tokens
279MiniMax: MiniMax M2.1MiniMax
81.0
LiveCodeBench197K tokens$0.30
280o3OpenAIOpenAI
81.0
LiveCodeBench200K tokens$2.00
281ERNIE 5.0 Thinking PreviewBaidu
81.0
LiveCodeBench
282Apriel-v1.6-15B-ThinkerServiceNow
81.0
LiveCodeBench
283o3 ProOpenAIOpenAI
80.8
LiveCodeBench200K tokens$20.00
284Gemini 3 Flash Preview (Non-reasoning)GoogleGoogle
80.0
LiveCodeBench$0.50
285GPT-5 NanoOpenAIOpenAI
79.0
LiveCodeBench400K tokens$0.05
286INTELLECT-3Prime Intellect
78.0
LiveCodeBench131K tokens
287DeepSeek V3.1DeepSeekDeepSeek
78.0
LiveCodeBench164K tokens$0.57
288GPT-5.4 ProOpenAIOpenAI
77.5
LiveBench Coding1.1M tokens$30.00
289Gemini 2.5 Pro Preview (May' 25)GoogleGoogle
77.0
LiveCodeBench$1.25
290Qwen: Qwen3 MaxAlibabaAlibaba
77.0
LiveCodeBench262K tokens$1.20
291Seed-OSS-36B-InstructByteDance Seed
77.0
LiveCodeBench$0.21
292Claude Sonnet 4.5AnthropicAnthropic
76.1
LiveBench Coding1.0M tokens$3.00
293KAT-Coder-Pro V1KwaiKAT
75.0
LiveCodeBench
294EXAONE 4.0 32B (Reasoning)LG AI Research
75.0
LiveCodeBench
295Claude Opus 4.5AnthropicAnthropic
74.0
LiveCodeBench200K tokens$5.00
296GLM-4.5 (Reasoning)Z.aiZ.ai
74.0
LiveCodeBench131K tokens
297Llama Nemotron Super 49B v1.5 (Reasoning)NvidiaNvidia
74.0
LiveCodeBench$0.40
298Qwen3 VL 32B (Reasoning)AlibabaAlibaba
74.0
LiveCodeBench$0.16
299Apriel-v1.5-15B-ThinkerServiceNow
73.0
LiveCodeBench
300o3 MiniOpenAIOpenAI
72.0
LiveCodeBench200K tokens$1.10
301Falcon-H1R-7BTII UAE
72.0
LiveCodeBench
302NVIDIA Nemotron Nano 9B V2 (Reasoning)NvidiaNvidia
72.0
LiveCodeBench$0.04
303Gemini 2.5 Flash Preview (Sep '25) (Reasoning)GoogleGoogle
71.0
LiveCodeBench
304MiniMax M1 80kMiniMax
71.0
LiveCodeBench$0.55
305NVIDIA: Nemotron Nano 9B V2NvidiaNvidia
70.1
LiveCodeBench131K tokens$0.05
306Gemini 2.5 Flash Preview (Reasoning)GoogleGoogle
70.0
LiveCodeBench$0.30
307Olmo 3.1 32B ThinkAllen Institute for AI
70.0
LiveCodeBench
308Qwen3 VL 30B A3B (Reasoning)AlibabaAlibaba
70.0
LiveCodeBench$0.20
309Cogito v2.1 (Reasoning)Deep Cogito
69.0
LiveCodeBench$1.25
310Gemini 2.5 Flash-Lite Preview (Sep '25) (Reasoning)GoogleGoogle
69.0
LiveCodeBench$0.10
311NVIDIA Nemotron Nano 12B v2 VL (Reasoning)NvidiaNvidia
69.0
LiveCodeBench$0.20
312Hermes 4 - Llama-3.1 405B (Reasoning)Nous Research
69.0
LiveCodeBench$1.00
313Ling-1TInclusionAI
68.0
LiveCodeBench
314Qwen3 Omni 30B A3B (Reasoning)AlibabaAlibaba
68.0
LiveCodeBench$0.25
315Qwen: Qwen3 Next 80B A3B InstructAlibabaAlibaba
68.0
LiveCodeBench262K tokens$0.15
316Olmo 3 32B ThinkAllenAI
67.0
LiveCodeBench66K tokens
317GPT-5.2-CodexOpenAIOpenAI
66.9
LiveCodeBench400K tokens$1.75
318MiniMax M1 40kMiniMax
66.0
LiveCodeBench
319Nova 2.0 Omni (medium)AmazonAmazon
66.0
LiveCodeBench$0.30
320Grok Code Fast 1xAIxAI
66.0
LiveCodeBench256K tokens
321Mi:dm K 2.5 ProKorea Telecom
66.0
LiveCodeBench
322Claude 4.1 Opus (Non-reasoning)AnthropicAnthropic
65.4
LiveCodeBench$15.00
323SpaceXAI: Grok Build 0.1xAIxAI
65.4
LiveBench Coding256K tokens$1.00
324Claude 4.1 Opus (Reasoning)AnthropicAnthropic
65.0
LiveCodeBench$15.00
325Qwen3 Max (Preview)AlibabaAlibaba
65.0
LiveCodeBench$1.20
326Hermes 4 - Llama-3.1 70B (Reasoning)Nous Research
65.0
LiveCodeBench$0.13
327Motif-2-12.7B-ReasoningMotif Technologies
65.0
LiveCodeBench
328Qwen3 VL 235B A22B (Reasoning)AlibabaAlibaba
65.0
LiveCodeBench$0.40
329Claude 4 Opus (Reasoning)AnthropicAnthropic
64.0
LiveCodeBench$15.00
330Ring-1TInclusionAI
64.0
LiveCodeBench
331Llama 3.1 Nemotron Ultra 253B v1 (Reasoning)NvidiaNvidia
64.0
LiveCodeBench$0.60
332Gemini 2.5 Flash-Lite Preview (Sep '25) (Non-reasoning)GoogleGoogle
64.0
LiveCodeBench$0.10
333Qwen3 4B 2507 (Reasoning)AlibabaAlibaba
64.0
LiveCodeBench
334QwQ 32BAlibabaAlibaba
63.0
LiveCodeBench$0.66
335HyperCLOVA X SEED Think (32B)Naver
63.0
LiveCodeBench
336Ring-flash-2.0InclusionAI
63.0
LiveCodeBench$0.14
337Solar Pro 2 (Non-reasoning)Upstage
62.0
LiveCodeBench
338Olmo 3 7B ThinkAllen Institute for AI
62.0
LiveCodeBench
339Qwen3 235B A22B (Reasoning)AlibabaAlibaba
62.0
LiveCodeBench$0.70
340DeepSeek: R1DeepSeekDeepSeek
61.7
LiveCodeBench64K tokens$1.35
341MoonshotAI: Kimi K2 0905MoonshotAIMoonshotAI
61.0
LiveCodeBench262K tokens$0.60
342GLM-4.5V (Reasoning)Z.aiZ.ai
60.0
LiveCodeBench$0.60
343Claude 4.5 Sonnet (Non-reasoning)AnthropicAnthropic
59.0
LiveCodeBench$3.00
344Qwen: Qwen3 VL 235B A22B InstructAlibabaAlibaba
59.0
LiveCodeBench262K tokens$0.40
345Qwen3 Coder 480B A35B InstructAlibabaAlibaba
59.0
LiveCodeBench$1.50
346Nova 2.0 Omni (low)AmazonAmazon
59.0
LiveCodeBench$0.30
347Ling-flash-2.0InclusionAI
59.0
LiveCodeBench$0.14
348Gemini 2.5 Flash LiteGoogleGoogle
59.0
LiveCodeBench1.0M tokens$0.10
349o1-miniOpenAIOpenAI
58.0
LiveCodeBench
350Mi:dm K 2.5 Pro PreviewKorea Telecom
58.0
LiveCodeBench
351GPT-5 (minimal)OpenAIOpenAI
56.0
LiveCodeBench$1.25
352Kimi K2Moonshot AIMoonshot AI
56.0
LiveCodeBench131K tokens$0.57
353DeepSeek V3.2 Exp (Non-reasoning)DeepSeekDeepSeek
55.0
LiveCodeBench$0.28
354GPT-5 mini (minimal)OpenAIOpenAI
55.0
LiveCodeBench$0.25
355Hermes 4 - Llama-3.1 405B (Non-reasoning)Nous Research
55.0
LiveCodeBench$1.00
356Claude Opus 4AnthropicAnthropic
54.0
LiveCodeBench200K tokens$15.00
357Qwen3 Max Thinking (Preview)AlibabaAlibaba
54.0
LiveCodeBench$1.20
358GPT-5 (ChatGPT)OpenAIOpenAI
54.0
LiveCodeBench
359K2-V2 (medium)MBZUAI Institute of Foundation Models
54.0
LiveCodeBench
360Magistral Medium 1MistralMistral
53.0
LiveCodeBench
361GPT-5.3-CodexOpenAIOpenAI
53.0
SciCode400K tokens$1.75
362Qwen3 30B A3B 2507 InstructAlibabaAlibaba
52.0
LiveCodeBench$0.20
363Claude Opus 4.6 (Adaptive Reasoning, Max Effort)AnthropicAnthropic
52.0
SciCode$5.00
364Claude Haiku 4.5AnthropicAnthropic
51.0
LiveCodeBench200K tokens$1.00
365Qwen: Qwen3 VL 32B InstructAlibabaAlibaba
51.0
LiveCodeBench131K tokens$0.16
366Qwen3 30B A3B (Reasoning)AlibabaAlibaba
51.0
LiveCodeBench$0.20
367Magistral Small 1MistralMistral
51.0
LiveCodeBench
368DeepSeek R1 0528 Qwen3 8BDeepSeekDeepSeek
51.0
LiveCodeBench
369Gemini 2.5 FlashGoogleGoogle
50.0
LiveCodeBench1.0M tokens$0.30
370GPT-5.5 Instant (May 2026)OpenAIOpenAI
50.0
SciCode$5.00
371Llama 3.1 Nemotron Nano 4B v1.1 (Reasoning)NvidiaNvidia
49.0
LiveCodeBench
372Gemini 3.5 Flash (minimal)GoogleGoogle
49.0
SciCode$1.50
373Qwen: Qwen3 VL 30B A3B InstructAlibabaAlibaba
48.0
LiveCodeBench131K tokens$0.20
374Baidu: ERNIE 4.5 300B A47B Baidu
47.0
LiveCodeBench123K tokens$0.28
375GPT-5 nano (minimal)OpenAIOpenAI
47.0
LiveCodeBench$0.05
376EXAONE 4.0 32B (Non-reasoning)LG AI Research
47.0
LiveCodeBench
377Qwen3 4B (Reasoning)AlibabaAlibaba
47.0
LiveCodeBench
378Qwen3.6 Max PreviewAlibabaAlibaba
47.0
SciCode$1.30
379Claude Sonnet 4.6AnthropicAnthropic
47.0
SciCode1.0M tokens$3.00
380GPT-4.1OpenAIOpenAI
46.0
LiveCodeBench1.0M tokens$2.00
381Solar Pro 2 (Preview) (Reasoning)Upstage
46.0
LiveCodeBench
382Claude Opus 4.6AnthropicAnthropic
46.0
SciCode1.0M tokens$5.00
383Claude Sonnet 4AnthropicAnthropic
45.0
LiveCodeBench1.0M tokens$3.00
384Grok 4.20 0309 (Reasoning)xAIxAI
45.0
SciCode$2.00
385Reka Flash 3Reka Flash 3
44.0
LiveCodeBench66K tokens$0.20
386GLM-5-TurboZ.aiZ.ai
44.0
SciCode203K tokens
387Claude Sonnet 4.6 (Non-reasoning, Low Effort)AnthropicAnthropic
44.0
SciCode$3.00
388GPT-4o (March 2025, chatgpt-4o-latest)OpenAIOpenAI
43.0
LiveCodeBench
389Ling-mini-2.0InclusionAI
43.0
LiveCodeBench
390Xiaomi: MiMo-V2-ProXiaomi
43.0
SciCode1.0M tokens
391MiniMax: MiniMax M2.5MiniMax
43.0
SciCode197K tokens$0.30
392Qwen: Qwen3 Max ThinkingAlibabaAlibaba
43.0
SciCode262K tokens$0.78
393GPT-4o (2024-11-20)OpenAIOpenAI
42.5
LiveCodeBench128K tokens$2.50
394Qwen3 Omni 30B A3B InstructAlibabaAlibaba
42.0
LiveCodeBench$0.25
395Gemini 2.5 Flash Preview (Non-reasoning)GoogleGoogle
41.0
LiveCodeBench
396Qwen3.5 Omni PlusAlibabaAlibaba
41.0
SciCode$0.40
397MiMo-V2-Omni-0327Xiaomi
40.0
SciCode
398Qwen: Qwen3.5-27BAlibabaAlibaba
40.0
SciCode262K tokens$0.30
399Step 3.5 FlashStepFun
40.0
SciCode$0.10
400Claude 3.7 SonnetAnthropicAnthropic
39.0
LiveCodeBench200K tokens$3.00
401Solar Pro 2 (Preview) (Non-reasoning)Upstage
39.0
LiveCodeBench
402DeepSeek R1 Distill Qwen 14BDeepSeekDeepSeek
38.0
LiveCodeBench
403Kimi Linear 48B A3B InstructKimiKimi
38.0
LiveCodeBench
404Qwen3 4B 2507 InstructAlibabaAlibaba
38.0
LiveCodeBench
405GLM-5 (Non-reasoning)Z.aiZ.ai
38.0
SciCode205K tokens$1.00
406Ling-2.6-1TInclusion AI
37.0
SciCode$0.30
407Xiaomi: MiMo-V2-OmniXiaomi
37.0
SciCode262K tokens
408Qwen2.5 MaxAlibabaAlibaba
36.0
LiveCodeBench
409NVIDIA: Nemotron 3 Nano 30B A3BNvidiaNvidia
36.0
LiveCodeBench262K tokens$0.05
410GLM-5.1 (Non-reasoning)Z.aiZ.ai
36.0
SciCode$1.38
411Qwen3 VL 8B (Reasoning)AlibabaAlibaba
35.0
LiveCodeBench$0.18
412NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning)NvidiaNvidia
35.0
LiveCodeBench$0.20
413Qwen: Qwen3 235B A22B Thinking 2507AlibabaAlibaba
34.3
LiveCodeBench131K tokens$0.15
414Mistral: Devstral MediumMistral AIMistral AI
34.0
LiveCodeBench131K tokens
415QwQ 32B-PreviewAlibabaAlibaba
34.0
LiveCodeBench
416Gemini 2.0 FlashGoogleGoogle
33.0
LiveCodeBench1.0M tokens
417Qwen: Qwen3 VL 8B InstructAlibabaAlibaba
33.0
LiveCodeBench131K tokens$0.18
418Mistral: Mistral Medium 3Mistral AIMistral AI
33.0
SciCode131K tokens$0.40
419GPT-4o (ChatGPT)OpenAIOpenAI
33.0
SciCode
420Qwen: Qwen3 30B A3B Thinking 2507AlibabaAlibaba
32.2
LiveCodeBench131K tokens$0.08
421GPT-4o (2024-08-06)OpenAIOpenAI
32.0
LiveCodeBench128K tokens$2.50
422Qwen: Qwen3 30B A3B Instruct 2507AlibabaAlibaba
32.0
LiveCodeBench262K tokens$0.20
423Qwen3 VL 4B (Reasoning)AlibabaAlibaba
32.0
LiveCodeBench
424OpenAI: GPT-4oOpenAIOpenAI
31.0
LiveCodeBench128K tokens$2.50
425Llama 3.1 Instruct 405BMetaMeta
31.0
LiveCodeBench$2.50
426Nova 2.0 Omni (Non-reasoning)AmazonAmazon
31.0
LiveCodeBench$0.30
427Qwen3 1.7B (Reasoning)AlibabaAlibaba
31.0
LiveCodeBench
428Step3 VL 10BStepFun
31.0
SciCode
429Qwen2.5 Coder 32B InstructAlibabaAlibaba
30.0
LiveCodeBench33K tokens
430SonarPerplexityPerplexity
30.0
LiveCodeBench127K tokens
431Sarvam M (Reasoning)Sarvam
30.0
LiveCodeBench
432Llama 3.1 Tulu3 405BAllen Institute for AI
29.0
LiveCodeBench
433Mistral Large 2 (Nov '24)MistralMistral
29.0
LiveCodeBench$4.00
434Qwen3 32B (Non-reasoning)AlibabaAlibaba
29.0
LiveCodeBench$0.16
435Cohere: Command ACohereCohere
29.0
LiveCodeBench256K tokens$2.50
436Llama Nemotron Super 49B v1.5 (Non-reasoning)NvidiaNvidia
29.0
LiveCodeBench$0.40
437Qwen3 VL 4B InstructAlibabaAlibaba
29.0
LiveCodeBench
438JT-35B-FlashChina Mobile
29.0
SciCode
439Llama 3.3 Nemotron Super 49B v1 (Reasoning)NvidiaNvidia
28.0
LiveCodeBench
440Qwen3 14B (Non-reasoning)AlibabaAlibaba
28.0
LiveCodeBench$0.35
441Qwen2.5 72B InstructAlibabaAlibaba
28.0
LiveCodeBench33K tokens$0.47
442Llama 3.3 Nemotron Super 49B v1 (Non-reasoning)NvidiaNvidia
28.0
LiveCodeBench
443Sonar Reasoning ProPerplexityPerplexity
28.0
LiveCodeBench128K tokens
444LongCat Flash LiteLongCat
28.0
SciCode
445Qwen: Qwen3 Coder 30B A3B InstructAlibabaAlibaba
28.0
SciCode160K tokens$0.45
446DeepSeek: R1 Distill Qwen 32BDeepSeekDeepSeek
27.0
LiveCodeBench128K tokens
447R1 Distill Llama 70BDeepSeekDeepSeek
27.0
LiveCodeBench8K tokens$0.70
448Hermes 4 - Llama-3.1 70B (Non-reasoning)Nous Research
27.0
LiveCodeBench$0.13
449Grok 2 (Dec '24)xAIxAI
27.0
LiveCodeBench
450Gemini 1.5 Flash (Sep '24)GoogleGoogle
27.0
LiveCodeBench
451Mistral Large 2 (Jul '24)MistralMistral
27.0
LiveCodeBench131K tokens$2.00
452Olmo 3 7B InstructAllen Institute for AI
27.0
LiveCodeBench$0.10
453JT-MINIChina Mobile
27.0
SciCode
454Solar Open 100B (Reasoning)Upstage
27.0
SciCode
455Mistral: Pixtral Large 2411Mistral AIMistral AI
26.0
LiveCodeBench131K tokens
456Devstral Small (May '25)MistralMistral
26.0
LiveCodeBench
457Qwen3.5 Omni FlashAlibabaAlibaba
26.0
SciCode$0.10
458Sarvam 105B (high)Sarvam
26.0
SciCode$0.04
459Devstral Small (Jul '25)MistralMistral
25.0
LiveCodeBench131K tokens
460Mistral Small 3MistralMistral
25.0
LiveCodeBench$0.10
461Qwen2.5 Instruct 32BAlibabaAlibaba
25.0
LiveCodeBench
462Granite 4.0 H SmallIBM
25.0
LiveCodeBench$0.06
463Grok BetaxAIxAI
24.0
LiveCodeBench
464Mistral: SabaMistral AIMistral AI
24.0
SciCode33K tokens$0.20
465GPT-4o-mini (2024-07-18)OpenAIOpenAI
23.4
LiveCodeBench128K tokens$0.15
466Llama 3.1 70B InstructMetaMeta
23.0
LiveCodeBench131K tokens$0.56
467Microsoft: Phi 4MicrosoftMicrosoft
23.0
LiveCodeBench16K tokens$0.13
468Qwen3 4B (Non-reasoning)AlibabaAlibaba
23.0
LiveCodeBench
469DeepSeek R1 Distill Llama 8BDeepSeekDeepSeek
23.0
LiveCodeBench
470Gemini 1.5 Flash-8BGoogleGoogle
22.0
LiveCodeBench
471Gemini 2.0 Flash (experimental)GoogleGoogle
21.0
LiveCodeBench
472Llama 3.2 Instruct 90B (Vision)MetaMeta
21.0
LiveCodeBench
473Jamba Reasoning 3BAI21 Labs
21.0
LiveCodeBench
474DeepHermes 3 - Mistral 24B Preview (Non-reasoning)Nous Research
20.0
LiveCodeBench
475Llama 3 70B InstructMetaMeta
20.0
LiveCodeBench8K tokens$0.65
476Gemini 1.5 Flash (May '24)GoogleGoogle
20.0
LiveCodeBench
477Qwen3 8B (Non-reasoning)AlibabaAlibaba
20.0
LiveCodeBench$0.18
478Gemma 4 E2B (Non-reasoning)GoogleGoogle
20.0
SciCode
479Gemini 2.0 Flash-Lite (Feb '25)GoogleGoogle
19.0
LiveCodeBench
480Hermes 3 - Llama-3.1 70BNous Research
19.0
LiveCodeBench$0.70
481Sarvam 30BSarvam
19.0
SciCode$0.03
482Gemini 2.0 Flash LiteGoogleGoogle
18.5
LiveCodeBench1.0M tokens$0.07
483Gemini 2.0 Flash-Lite (Preview)GoogleGoogle
18.0
LiveCodeBench
484Claude 3 SonnetAnthropicAnthropic
18.0
LiveCodeBench$3.00
485AI21: Jamba Large 1.7AI21 Labs
18.0
LiveCodeBench256K tokens$2.00
486Granite 4.0 MicroIBM
18.0
LiveCodeBench131K tokens
487Tri-21B-think PreviewTrillion Labs
18.0
SciCode
488Mistral LargeMistral AIMistral AI
17.8
LiveCodeBench128K tokens$4.00
489Llama 3.1 Nemotron 70B InstructNvidiaNvidia
17.0
LiveCodeBench131K tokens$1.20
490Jamba 1.6 LargeAI21 Labs
17.0
LiveCodeBench$2.00
491Tri-21B-ThinkTrillion Labs
17.0
SciCode
492Olmo 3.1 32B InstructAllenAI
17.0
SciCode66K tokens
493GLM-4.6V (Reasoning)Z.aiZ.ai
16.0
LiveCodeBench$0.30
494Qwen2 Instruct 72BAlibabaAlibaba
16.0
LiveCodeBench
495DeepSeek Coder V2 Lite InstructDeepSeekDeepSeek
16.0
LiveCodeBench
496Mixtral 8x22B InstructMistralMistral
15.0
LiveCodeBench
497Anthropic: Claude 3 HaikuAnthropicAnthropic
15.0
LiveCodeBench200K tokens$0.25
498LFM2 8B A1BLiquid AI
15.0
LiveCodeBench
499Mistral Small (Sep '24)MistralMistral
14.0
LiveCodeBench$0.20
500Jamba 1.5 LargeAI21 Labs
14.0
LiveCodeBench$2.00
501Gemma 3n E4B Instruct Preview (May '25)GoogleGoogle
14.0
LiveCodeBench
502Qwen2.5 Coder Instruct 7B AlibabaAlibaba
13.0
LiveCodeBench
503Phi-4 Multimodal InstructMicrosoftMicrosoft
13.0
LiveCodeBench
504Granite 3.3 8B (Non-reasoning)IBM
13.0
LiveCodeBench$0.03
505Qwen3 1.7B (Non-reasoning)AlibabaAlibaba
13.0
LiveCodeBench
506Molmo2-8BAllen Institute for AI
13.0
SciCode
507Cohere: Command R+ (08-2024)CohereCohere
12.2
LiveCodeBench128K tokens$2.50
508Command-R+ (Apr '24)CohereCohere
12.0
LiveCodeBench$3.00
509Gemini 1.0 ProGoogleGoogle
12.0
LiveCodeBench
510Phi-3 Mini Instruct 3.8BMicrosoftMicrosoft
12.0
LiveCodeBench
511Granite 4.0 H 1BIBM
12.0
LiveCodeBench
512Qwen3 0.6B (Reasoning)AlibabaAlibaba
12.0
LiveCodeBench
513OpenChat 3.5 (1210)OpenChat
12.0
LiveCodeBench
514Mistral Small (Feb '24)MistralMistral
11.0
LiveCodeBench$0.15
515Llama 3.2 11B Vision InstructMetaMeta
11.0
LiveCodeBench131K tokens$0.34
516LFM2-24B-A2BLiquidAI
11.0
SciCode33K tokens
517Llama 3 8B InstructMetaMeta
10.0
LiveCodeBench8K tokens$0.04
518Mistral MediumMistralMistral
10.0
LiveCodeBench$1.50
519Llama 2 Chat 13BMetaMeta
10.0
LiveCodeBench
520LFM 40BLiquid AI
10.0
LiveCodeBench
521Gemma 3n E2B InstructGoogleGoogle
10.0
LiveCodeBench
522Llama 2 Chat 70BMetaMeta
10.0
LiveCodeBench
523DBRX InstructDatabricks
9.0
LiveCodeBench
524DeepHermes 3 - Llama-3.1 8B Preview (Non-reasoning)Nous Research
9.0
LiveCodeBench
525Llama 3.2 3B InstructMetaMeta
8.0
LiveCodeBench80K tokens
526LFM2 2.6BLiquid AI
8.0
LiveCodeBench
527LFM2.5-8B-A1BLiquid AI
8.0
SciCode
528Jamba 1.6 MiniAI21 Labs
7.0
LiveCodeBench$0.20
529OLMo 2 32BAllen Institute for AI
7.0
LiveCodeBench
530DeepSeek R1 Distill Qwen 1.5BDeepSeekDeepSeek
7.0
LiveCodeBench
531Qwen3 0.6B (Non-reasoning)AlibabaAlibaba
7.0
LiveCodeBench
532Mistral: Mixtral 8x7B InstructMistral AIMistral AI
7.0
LiveCodeBench33K tokens$0.45
533Jamba 1.7 MiniAI21 Labs
6.0
LiveCodeBench
534Jamba 1.5 MiniAI21 Labs
6.0
LiveCodeBench$0.20
535Apertus 70B InstructSwiss AI Initiative
6.0
SciCode$0.82
536Granite 4.0 1BIBM
5.0
LiveCodeBench
537Command-R (Mar '24)CohereCohere
5.0
LiveCodeBench$0.50
538Mistral 7B InstructMistralMistral
5.0
LiveCodeBench$0.25
539OLMo 2 7BAllen Institute for AI
4.0
LiveCodeBench
540Molmo 7B-DAllen Institute for AI
4.0
LiveCodeBench
541MiniCPM5-1B (Non-reasoning)OpenBMB
4.0
SciCode
542Gemma 4 E4B (Non-reasoning)GoogleGoogle
4.0
SciCode$0.02
543Tiny Aya GlobalCohereCohere
4.0
SciCode
544LFM2.5-1.2B-ThinkingLiquid AI
4.0
SciCode
545Apertus 8B InstructSwiss AI Initiative
4.0
SciCode$0.10
546LFM2.5-VL-1.6BLiquid AI
3.0
SciCode
547LFM2 1.2BLiquid AI
2.0
LiveCodeBench
548Granite 4.0 H 350MIBM
2.0
LiveCodeBench
549Llama 3.2 1B InstructMetaMeta
2.0
LiveCodeBench60K tokens
550Granite 4.0 350MIBM
2.0
LiveCodeBench
551Gemma 3 1B InstructGoogleGoogle
2.0
LiveCodeBench
552LFM2.5-1.2B-InstructLiquid AI
2.0
SciCode
553Gemma 3 270MGoogleGoogle
0.0
LiveCodeBench
554Llama 2 Chat 7BMetaMeta
0.0
LiveCodeBench$0.05
555Qwen3.5 0.8B (Reasoning)AlibabaAlibaba
0.0
SciCode

+ 247 modelos sem benchmark de coding disponível.Ver todos os modelos

Guia Completo: IA para Programação em 2026

O Estado da IA para Código em 2026

A inteligência artificial transformou radicalmente o desenvolvimento de software nos últimos anos. Em 2026, modelos de linguagem (LLMs) são capazes de gerar código funcional em dezenas de linguagens, resolver bugs em projetos reais e até criar aplicações completas a partir de descrições em linguagem natural. O SWE-bench — o benchmark mais rigoroso para coding — avalia modelos em tarefas reais de engenharia de software extraídas de issues do GitHub.

SWE-bench: O Benchmark de Referência

O SWE-bench (Software Engineering Benchmark) é considerado o padrão ouro para avaliar capacidade de coding de LLMs. Diferente de benchmarks acadêmicos como HumanEval (que testa funções isoladas), o SWE-bench apresenta issues reais de repositórios populares como Django, Flask, scikit-learn e requests. O modelo precisa entender o contexto do projeto, localizar os arquivos relevantes e gerar um patch que resolva o bug — simulando o trabalho real de um desenvolvedor.

A versão “Verified” do SWE-bench (SWE-bench Verified) é curada por engenheiros humanos para garantir que cada tarefa tem uma solução clara e verificável. Os scores neste benchmark são particularmente informativos porque correlacionam fortemente com a experiência real de uso para coding.

HumanEval e LiveCodeBench

HumanEval, criado pela OpenAI, testa a capacidade de gerar funções Python a partir de docstrings. É um benchmark mais simples que o SWE-bench, mas útil para avaliar fluência básica em código. LiveCodeBench adiciona uma camada de complexidade ao testar com problemas que são atualizados regularmente, reduzindo o risco de contaminação (quando o modelo já viu as respostas durante o treinamento).

Como Escolher o Melhor Modelo para Código

A escolha do modelo ideal depende do caso de uso específico. Para autocompletar código em tempo real (Cursor, Copilot), velocidade e latência são mais importantes que score máximo — modelos das classes mini e flash costumam oferecer a melhor relação velocidade/qualidade. Para geração de projetos completos ou debug complexo, modelos frontier como Claude Fable 5, GPT-5.5 e Gemini 3.5 tendem a ser mais adequados, apesar do custo maior.

Para equipes que precisam de controle sobre os dados (compliance, segurança), modelos open source como DeepSeek Coder, Code Llama e StarCoder permitem deploy on-premises com performance competitiva. A decisão entre proprietário e open source envolve tradeoffs de custo, latência, privacidade e qualidade.

Ferramentas de Coding com IA

As principais ferramentas de desenvolvimento assistido por IA em 2026 incluem Cursor, GitHub Copilot, Windsurf e fluxos agentic integrados ao editor. Cada ferramenta usa diferentes modelos por baixo, e a qualidade do código gerado depende diretamente da capacidade do LLM utilizado.

Para desenvolvedores brasileiros, um fator importante é a capacidade do modelo de entender comentários, nomes de variáveis e documentação em português — algo que varia significativamente entre modelos e que não é capturado pelos benchmarks tradicionais em inglês.

Tendências para 2026 e Além

As tendências mais relevantes em IA para código incluem: agentes autônomos de engenharia (que resolvem tarefas complexas sem supervisão), geração de testes automatizados, refatoração inteligente, e integração nativa com pipelines de CI/CD. A fronteira está se movendo de “assistente de código” para “engenheiro autônomo”, com modelos cada vez mais capazes de navegar codebases grandes e tomar decisões arquiteturais.

Perguntas Frequentes

Qual é a melhor IA para programar?

Em 2026, os modelos que lideram em benchmarks de código são Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback), Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback). No entanto, a melhor escolha depende do caso de uso: autocompletar código, geração de projetos completos, debug ou code review.

ChatGPT ou Claude para código?

Hoje as opções mais fortes costumam orbitar Claude Fable 5 / Opus 4.8, GPT-5.5 e modelos da classe Gemini 3.5. Claude tende a brilhar em refactors longos; GPT é forte em geração rápida e tool use. O teste final precisa ser feito no seu próprio repositório.

O que é o SWE-bench?

SWE-bench (Software Engineering Benchmark) avalia a capacidade de modelos de resolver issues reais de repositórios open source no GitHub. É considerado o benchmark mais realista para coding, pois testa resolução de bugs em projetos reais, não exercícios acadêmicos.

Quais LLMs gratuitas são boas para código?

Modelos open source como DeepSeek Coder, Qwen Coder e Code Llama oferecem excelente performance em coding sem custo de API. Podem ser rodados localmente via Ollama ou acessados gratuitamente em plataformas como Together AI e Groq.

Cursor ou GitHub Copilot?

Cursor e Copilot são IDEs/extensões que usam LLMs por baixo. Cursor permite escolher o modelo (Claude, GPT, etc.), enquanto Copilot usa modelos da OpenAI. A qualidade do código gerado depende mais do modelo escolhido do que da ferramenta em si.

Explorar Outras Categorias