Grok 4.6
Same $2 / $6 API price as Grok 4.5. xAI reports an Artificial Analysis Intelligence Index of 61 — tied with GPT-5.6 Sol Max, one point behind Fable 5 Max. Strong on knowledge-work evals; 26% on Terminal-Bench v3.
Read the Grok 4.6 briefingBenchmarks de IA
A practical view of published model scores that helps you compare capability, category strengths, and run cost without losing the detail.
Latest benchmark & model updates
The leaderboard on this page is LiveBench's 25 June release, re-captured on 22 September 2026. It now includes Grok 4.7 xHigh and DeepSeek V4.1 Flash Max Effort. The model notes below are from the vendors' own pages, labelled as such.
Same $2 / $6 API price as Grok 4.5. xAI reports an Artificial Analysis Intelligence Index of 61 — tied with GPT-5.6 Sol Max, one point behind Fable 5 Max. Strong on knowledge-work evals; 26% on Terminal-Bench v3.
Read the Grok 4.6 briefingDeepSeek V4.1 Flash replaced V4 Flash on 10 Sep 2026 on the new deepseek-flash API id, at $0.15 / $0.60 per 1M tokens off-peak and $0.30 / $1.20 at peak. From 14 Sep, V4 Pro requests are routed to it too, so switch model ids now. LiveBench now scores V4.1 Flash Max Effort at 81.1 overall.
See it in the model trackerLeaderboard rows come only from published LiveBench data. We do not invent scores.
Best overall
Claude Fable 5.1 Max Effort leads this snapshot at 83.4 overall, $1.212 per successful task.
Best value per task
Union Alpha runs at $0.000 per successful task and still scores 76.1 overall.
Best open weights
DeepSeek V4.1 Flash Max Effort is the strongest open-weight model here at 81.1 overall — the option you can host yourself.
Best per capability
If your work is mostly one kind of task, these lead their category in this snapshot.
Every number above comes from the 2026-06-25 release dated 2026-06-25, published by a third-party benchmark and reproduced unaltered.
O primeiro gráfico responde “quem é o mais forte”. O segundo responde à pergunta que realmente custa dinheiro: qual deles vale a pena pagar.
Mais longo é melhor. As barras mostram a pontuação Geral de 100; o número ao lado de cada barra é o valor exato. Selecione uma linha para levar esse modelo para o gráfico de custos abaixo.
Mostrando os 12 primeiros de 54. Mude para a visão de comparação completa para todos os modelos.
Para cima é mais forte, para a esquerda é mais barato — então as melhores compras ficam no canto superior esquerdo. A linha em degraus liga os modelos que nada mais barato supera; qualquer coisa abaixo dela é superada por algo que custa menos.
Selected
Claude Fable 5.1 Max Effort
Anthropic
#5 · Na fronteira de valor.
Best buys (cheapest → strongest)
Use Tab para entrar no gráfico e depois as setas para percorrer os modelos do mais barato ao mais caro.
| Modelo | Organização | Geral | por tarefa bem-sucedida |
|---|---|---|---|
| Claude Fable 5.1 Max Effort#5 · Melhor valor pelo preço | Anthropic | 83.4 | US$ 1,212 |
| Claude Fable 5 Max Effort | Anthropic | 83.0 | US$ 1,439 |
| GPT-6 Astra Max Effort#4 · Melhor valor pelo preço | OpenAI | 82.2 | US$ 0,736 |
| Muse Spark 1.3 xHigh Effort#3 · Melhor valor pelo preço | Meta | 81.6 | US$ 0,219 |
| DeepSeek V4.1 Flash Max Effort#2 · Melhor valor pelo preço | DeepSeek | 81.1 | US$ 0,029 |
| GPT-5.6 Sol Max Effort | OpenAI | 81.0 | US$ 0,515 |
| GPT-5.5 Thinking xHigh Effort | OpenAI | 80.2 | US$ 0,435 |
| Claude 5 Opus Thinking Max Effort | Anthropic | 80.1 | US$ 0,699 |
| Kimi K3 | Moonshot AI | 79.2 | US$ 0,348 |
| Gemini 3.7 Flash High | 78.8 | US$ 0,157 | |
| Qwen 3.8 Max | Alibaba | 78.5 | US$ 0,275 |
| Grok 4.6 | xAI | 78.0 | US$ 0,207 |
| Muse Spark 1.2 xHigh Effort | Meta | 78.0 | US$ 0,375 |
| GPT-5.4 Thinking xHigh Effort | OpenAI | 78.0 | US$ 0,387 |
| GPT-5.6 Terra Max Effort | OpenAI | 77.9 | US$ 0,352 |
| DeepSeek V4 Pro 0813 | DeepSeek | 77.4 | US$ 0,044 |
| Grok 4.7 xHigh | xAI | 77.4 | US$ 0,718 |
| Gemini 3.1 Pro Preview High | 77.0 | US$ 0,286 | |
| DeepSeek V4 Flash Vision Exp | DeepSeek | 76.8 | US$ 0,051 |
| Claude 4.7 Opus Thinking xHigh Effort | Anthropic | 76.5 | US$ 0,528 |
| Qwen 3.8 Flash Next | Alibaba | 76.2 | US$ 0,042 |
| Claude 4.8 Opus Thinking Max Effort | Anthropic | 76.2 | US$ 0,983 |
| GLM-5.3 | Z.AI | 76.1 | US$ 0,450 |
| Claude Sonnet 5 xHigh Effort | Anthropic | 76.0 | US$ 0,505 |
| Grok 4.5 | xAI | 75.8 | US$ 0,131 |
| Gemini 3.8 Flash High | 75.8 | US$ 0,307 | |
| Qwen3.8 27B | Alibaba | 75.3 | US$ 0,094 |
| Muse Spark 1.1 xHigh Effort | Meta | 75.3 | US$ 0,198 |
| GPT-5.2 High | OpenAI | 74.6 | US$ 0,234 |
| Gemini 3.5 Flash High | 74.6 | US$ 0,249 | |
| Claude 4.6 Opus Thinking High Effort | Anthropic | 74.5 | US$ 0,404 |
| DeepSeek V4 Flash 0731 | DeepSeek | 74.2 | US$ 0,060 |
| GPT-5.2 Codex | OpenAI | 74.0 | US$ 0,187 |
| GPT-5.6 Luna Max Effort | OpenAI | 73.6 | US$ 0,169 |
| Gemini 3.6 Flash High | 73.6 | US$ 0,235 | |
| GLM-5.2 | Z.AI | 73.2 | US$ 0,225 |
| Qwen 3.7 Max | Alibaba | 73.1 | US$ 0,182 |
| Claude 4.6 Sonnet Thinking Medium Effort | Anthropic | 73.0 | US$ 0,306 |
| Claude 4.5 Opus Thinking High Effort | Anthropic | 72.6 | US$ 0,610 |
| Inkling xHigh Effort | Thinking Machines | 71.9 | US$ 0,310 |
| GLM-5.3 Flash | Z.AI | 71.6 | US$ 0,031 |
| Kimi K2.6 Thinking | Moonshot AI | 70.5 | US$ 0,169 |
| GPT-5.4 Nano xHigh | OpenAI | 69.6 | US$ 0,091 |
| Qwen 3.6 Plus | Alibaba | 68.9 | US$ 0,227 |
| Kimi K2.7 Code | Moonshot AI | 68.4 | US$ 0,100 |
| Grok Build 0.1#1 · Melhor valor pelo preço | xAI | 67.8 | US$ 0,024 |
| Nemotron 3 Ultra 550B A55B | NVIDIA | 67.4 | US$ 0,371 |
| Minimax M3 | Minimax | 67.3 | US$ 0,060 |
| GPT-5.4 Mini xHigh | OpenAI | 66.4 | US$ 0,334 |
| Qwen 3.6 27B | Alibaba | 64.0 | US$ 0,202 |
| Gemini 3.5 Flash-Lite High | 63.9 | US$ 0,069 | |
| Grok 4.3 | xAI | 62.3 | US$ 0,061 |
Selecione até três modelos do ranking para comparar suas pontuações em sete categorias.
Cada faixa vai de 0 a 100. Quanto mais à direita, mais forte; o número em negrito é o líder nessa linha.
| Categoria | Claude Fable 5.1 Max Effort | Union Alpha | DeepSeek V4.1 Flash Max Effort |
|---|---|---|---|
| Raciocínio | 91.7 | 80.8 | 86.7 |
| Codificação | 86.4 | 82.1 | 80.0 |
| Codificação agêntica | 66.1 | 54.7 | 77.3 |
| Matemática | 97.0 | 95.3 | 93.3 |
| Análise de dados | 80.3 | 74.6 | 79.3 |
| Linguagem | 89.5 | 85.9 | 81.2 |
| Seguimento de instruções | 73.0 | 59.5 | 70.0 |
Classifique cada categoria, restrinja a lista e abra um modelo para sua divisão de subtarefas publicada.
Claude Fable 5.1 Max Effort
Anthropic
83.4
Geral
US$ 1,212
por tarefa bem-sucedida
Claude Fable 5 Max Effort
Anthropic
83.0
Geral
US$ 1,439
por tarefa bem-sucedida
GPT-6 Astra Max Effort
OpenAI
82.2
Geral
US$ 0,736
por tarefa bem-sucedida
Muse Spark 1.3 xHigh Effort
Meta
81.6
Geral
US$ 0,219
por tarefa bem-sucedida
DeepSeek V4.1 Flash Max Effort
abertoDeepSeek
81.1
Geral
US$ 0,029
por tarefa bem-sucedida
GPT-5.6 Sol Max Effort
OpenAI
81.0
Geral
US$ 0,515
por tarefa bem-sucedida
GPT-5.5 Thinking xHigh Effort
OpenAI
80.2
Geral
US$ 0,435
por tarefa bem-sucedida
Claude 5 Opus Thinking Max Effort
Anthropic
80.1
Geral
US$ 0,699
por tarefa bem-sucedida
Kimi K3
abertoMoonshot AI
79.2
Geral
US$ 0,348
por tarefa bem-sucedida
Gemini 3.7 Flash High
78.8
Geral
US$ 0,157
por tarefa bem-sucedida
Qwen 3.8 Max
abertoAlibaba
78.5
Geral
US$ 0,275
por tarefa bem-sucedida
Muse Spark 1.2 xHigh Effort
Meta
78.0
Geral
US$ 0,375
por tarefa bem-sucedida
GPT-5.4 Thinking xHigh Effort
OpenAI
78.0
Geral
US$ 0,387
por tarefa bem-sucedida
Grok 4.6
xAI
78.0
Geral
US$ 0,207
por tarefa bem-sucedida
GPT-5.6 Terra Max Effort
OpenAI
77.9
Geral
US$ 0,352
por tarefa bem-sucedida
DeepSeek V4 Pro 0813
abertoDeepSeek
77.4
Geral
US$ 0,044
por tarefa bem-sucedida
Grok 4.7 xHigh
xAI
77.4
Geral
US$ 0,718
por tarefa bem-sucedida
Gemini 3.1 Pro Preview High
77.0
Geral
US$ 0,286
por tarefa bem-sucedida
DeepSeek V4 Flash Vision Exp
abertoDeepSeek
76.8
Geral
US$ 0,051
por tarefa bem-sucedida
Claude 4.7 Opus Thinking xHigh Effort
Anthropic
76.5
Geral
US$ 0,528
por tarefa bem-sucedida
Qwen 3.8 Flash Next
abertoAlibaba
76.2
Geral
US$ 0,042
por tarefa bem-sucedida
Claude 4.8 Opus Thinking Max Effort
Anthropic
76.2
Geral
US$ 0,983
por tarefa bem-sucedida
GLM-5.3
abertoZ.AI
76.1
Geral
US$ 0,450
por tarefa bem-sucedida
Union Alpha
Stealth
76.1
Geral
US$ 0,000
por tarefa bem-sucedida
Claude Sonnet 5 xHigh Effort
Anthropic
76.0
Geral
US$ 0,505
por tarefa bem-sucedida
Gemini 3.8 Flash High
75.8
Geral
US$ 0,307
por tarefa bem-sucedida
Grok 4.5
xAI
75.8
Geral
US$ 0,131
por tarefa bem-sucedida
Qwen3.8 27B
abertoAlibaba
75.3
Geral
US$ 0,094
por tarefa bem-sucedida
Muse Spark 1.1 xHigh Effort
Meta
75.3
Geral
US$ 0,198
por tarefa bem-sucedida
GPT-5.2 High
OpenAI
74.6
Geral
US$ 0,234
por tarefa bem-sucedida
Gemini 3.5 Flash High
74.6
Geral
US$ 0,249
por tarefa bem-sucedida
Claude 4.6 Opus Thinking High Effort
Anthropic
74.5
Geral
US$ 0,404
por tarefa bem-sucedida
DeepSeek V4 Flash 0731
abertoDeepSeek
74.2
Geral
US$ 0,060
por tarefa bem-sucedida
GPT-5.2 Codex
OpenAI
74.0
Geral
US$ 0,187
por tarefa bem-sucedida
Gemini 3.6 Flash High
73.6
Geral
US$ 0,235
por tarefa bem-sucedida
GPT-5.6 Luna Max Effort
OpenAI
73.6
Geral
US$ 0,169
por tarefa bem-sucedida
GLM-5.2
abertoZ.AI
73.2
Geral
US$ 0,225
por tarefa bem-sucedida
Qwen 3.7 Max
Alibaba
73.1
Geral
US$ 0,182
por tarefa bem-sucedida
Claude 4.6 Sonnet Thinking Medium Effort
Anthropic
73.0
Geral
US$ 0,306
por tarefa bem-sucedida
Claude 4.5 Opus Thinking High Effort
Anthropic
72.6
Geral
US$ 0,610
por tarefa bem-sucedida
Inkling xHigh Effort
abertoThinking Machines
71.9
Geral
US$ 0,310
por tarefa bem-sucedida
GLM-5.3 Flash
abertoZ.AI
71.6
Geral
US$ 0,031
por tarefa bem-sucedida
Kimi K2.6 Thinking
abertoMoonshot AI
70.5
Geral
US$ 0,169
por tarefa bem-sucedida
GPT-5.4 Nano xHigh
OpenAI
69.6
Geral
US$ 0,091
por tarefa bem-sucedida
ox-alpha-max
Stealth
69.2
Geral
US$ 0,000
por tarefa bem-sucedida
Qwen 3.6 Plus
Alibaba
68.9
Geral
US$ 0,227
por tarefa bem-sucedida
Kimi K2.7 Code
abertoMoonshot AI
68.4
Geral
US$ 0,100
por tarefa bem-sucedida
Grok Build 0.1
xAI
67.8
Geral
US$ 0,024
por tarefa bem-sucedida
Nemotron 3 Ultra 550B A55B
abertoNVIDIA
67.4
Geral
US$ 0,371
por tarefa bem-sucedida
Minimax M3
Minimax
67.3
Geral
US$ 0,060
por tarefa bem-sucedida
GPT-5.4 Mini xHigh
OpenAI
66.4
Geral
US$ 0,334
por tarefa bem-sucedida
Qwen 3.6 27B
abertoAlibaba
64.0
Geral
US$ 0,202
por tarefa bem-sucedida
Gemini 3.5 Flash-Lite High
63.9
Geral
US$ 0,069
por tarefa bem-sucedida
Grok 4.3
xAI
62.3
Geral
US$ 0,061
por tarefa bem-sucedida
Cost is the source-provided cost-per-successful-task metric for the selected scope.
These scores say which model is strongest at a set of held-out tasks. They don’t say which one fits your workflow, your budget, or the tools you already pay for — and the cheapest model that clears your bar usually beats the highest scorer.
What changed recently
Dated releases from every major lab, with the vendor's own claims labelled as claims.
OpenWhat it costs to run
Our tool catalogue carries the price we last verified and the date we checked it.
OpenWhat to actually build
A deployment plan picks the stack for one workflow at your budget — models included.
Open