Grok 4.6
Same $2 / $6 API price as Grok 4.5. xAI reports an Artificial Analysis Intelligence Index of 61 — tied with GPT-5.6 Sol Max, one point behind Fable 5 Max. Strong on knowledge-work evals; 26% on Terminal-Bench v3.
Read the Grok 4.6 briefingBenchmark AI
A practical view of published model scores that helps you compare capability, category strengths, and run cost without losing the detail.
Latest benchmark & model updates
The leaderboard on this page is LiveBench's 25 June release, re-captured on 22 September 2026. It now includes Grok 4.7 xHigh and DeepSeek V4.1 Flash Max Effort. The model notes below are from the vendors' own pages, labelled as such.
Same $2 / $6 API price as Grok 4.5. xAI reports an Artificial Analysis Intelligence Index of 61 — tied with GPT-5.6 Sol Max, one point behind Fable 5 Max. Strong on knowledge-work evals; 26% on Terminal-Bench v3.
Read the Grok 4.6 briefingDeepSeek V4.1 Flash replaced V4 Flash on 10 Sep 2026 on the new deepseek-flash API id, at $0.15 / $0.60 per 1M tokens off-peak and $0.30 / $1.20 at peak. From 14 Sep, V4 Pro requests are routed to it too, so switch model ids now. LiveBench now scores V4.1 Flash Max Effort at 81.1 overall.
See it in the model trackerLeaderboard rows come only from published LiveBench data. We do not invent scores.
Best overall
Claude Fable 5.1 Max Effort leads this snapshot at 83.4 overall, $1.212 per successful task.
Best value per task
Union Alpha runs at $0.000 per successful task and still scores 76.1 overall.
Best open weights
DeepSeek V4.1 Flash Max Effort is the strongest open-weight model here at 81.1 overall — the option you can host yourself.
Best per capability
If your work is mostly one kind of task, these lead their category in this snapshot.
Every number above comes from the 2026-06-25 release dated 2026-06-25, published by a third-party benchmark and reproduced unaltered.
Il primo grafico risponde a “chi è il più forte”. Il secondo risponde alla domanda che ti costa davvero denaro: quale tra questi vale la pena pagare.
Più lungo è meglio. Le barre mostrano il punteggio Generale su 100; il numero accanto a ciascuna barra è il valore esatto. Seleziona una riga per trasportare quel modello nel grafico dei costi sottostante.
Mostrando i primi 12 di 54. Passa alla vista comparativa completa per ogni modello.
In alto è più forte, a sinistra è più economico — quindi i migliori acquisti stanno in alto a sinistra. La linea a gradini collega i modelli che nessun modello più economico batte; tutto ciò che sta sotto è superato da qualcosa che costa meno.
Selected
Claude Fable 5.1 Max Effort
Anthropic
#5 · Sulla frontiera del valore.
Best buys (cheapest → strongest)
Usa Tab per entrare nel grafico, poi le frecce per scorrere i modelli dal più economico al più caro.
| Modello | Organizzazione | Generale | per attività riuscita |
|---|---|---|---|
| Claude Fable 5.1 Max Effort#5 · Miglior rapporto qualità-prezzo al suo prezzo | Anthropic | 83.4 | 1,212 USD |
| Claude Fable 5 Max Effort | Anthropic | 83.0 | 1,439 USD |
| GPT-6 Astra Max Effort#4 · Miglior rapporto qualità-prezzo al suo prezzo | OpenAI | 82.2 | 0,736 USD |
| Muse Spark 1.3 xHigh Effort#3 · Miglior rapporto qualità-prezzo al suo prezzo | Meta | 81.6 | 0,219 USD |
| DeepSeek V4.1 Flash Max Effort#2 · Miglior rapporto qualità-prezzo al suo prezzo | DeepSeek | 81.1 | 0,029 USD |
| GPT-5.6 Sol Max Effort | OpenAI | 81.0 | 0,515 USD |
| GPT-5.5 Thinking xHigh Effort | OpenAI | 80.2 | 0,435 USD |
| Claude 5 Opus Thinking Max Effort | Anthropic | 80.1 | 0,699 USD |
| Kimi K3 | Moonshot AI | 79.2 | 0,348 USD |
| Gemini 3.7 Flash High | 78.8 | 0,157 USD | |
| Qwen 3.8 Max | Alibaba | 78.5 | 0,275 USD |
| Grok 4.6 | xAI | 78.0 | 0,207 USD |
| Muse Spark 1.2 xHigh Effort | Meta | 78.0 | 0,375 USD |
| GPT-5.4 Thinking xHigh Effort | OpenAI | 78.0 | 0,387 USD |
| GPT-5.6 Terra Max Effort | OpenAI | 77.9 | 0,352 USD |
| DeepSeek V4 Pro 0813 | DeepSeek | 77.4 | 0,044 USD |
| Grok 4.7 xHigh | xAI | 77.4 | 0,718 USD |
| Gemini 3.1 Pro Preview High | 77.0 | 0,286 USD | |
| DeepSeek V4 Flash Vision Exp | DeepSeek | 76.8 | 0,051 USD |
| Claude 4.7 Opus Thinking xHigh Effort | Anthropic | 76.5 | 0,528 USD |
| Qwen 3.8 Flash Next | Alibaba | 76.2 | 0,042 USD |
| Claude 4.8 Opus Thinking Max Effort | Anthropic | 76.2 | 0,983 USD |
| GLM-5.3 | Z.AI | 76.1 | 0,450 USD |
| Claude Sonnet 5 xHigh Effort | Anthropic | 76.0 | 0,505 USD |
| Grok 4.5 | xAI | 75.8 | 0,131 USD |
| Gemini 3.8 Flash High | 75.8 | 0,307 USD | |
| Qwen3.8 27B | Alibaba | 75.3 | 0,094 USD |
| Muse Spark 1.1 xHigh Effort | Meta | 75.3 | 0,198 USD |
| GPT-5.2 High | OpenAI | 74.6 | 0,234 USD |
| Gemini 3.5 Flash High | 74.6 | 0,249 USD | |
| Claude 4.6 Opus Thinking High Effort | Anthropic | 74.5 | 0,404 USD |
| DeepSeek V4 Flash 0731 | DeepSeek | 74.2 | 0,060 USD |
| GPT-5.2 Codex | OpenAI | 74.0 | 0,187 USD |
| GPT-5.6 Luna Max Effort | OpenAI | 73.6 | 0,169 USD |
| Gemini 3.6 Flash High | 73.6 | 0,235 USD | |
| GLM-5.2 | Z.AI | 73.2 | 0,225 USD |
| Qwen 3.7 Max | Alibaba | 73.1 | 0,182 USD |
| Claude 4.6 Sonnet Thinking Medium Effort | Anthropic | 73.0 | 0,306 USD |
| Claude 4.5 Opus Thinking High Effort | Anthropic | 72.6 | 0,610 USD |
| Inkling xHigh Effort | Thinking Machines | 71.9 | 0,310 USD |
| GLM-5.3 Flash | Z.AI | 71.6 | 0,031 USD |
| Kimi K2.6 Thinking | Moonshot AI | 70.5 | 0,169 USD |
| GPT-5.4 Nano xHigh | OpenAI | 69.6 | 0,091 USD |
| Qwen 3.6 Plus | Alibaba | 68.9 | 0,227 USD |
| Kimi K2.7 Code | Moonshot AI | 68.4 | 0,100 USD |
| Grok Build 0.1#1 · Miglior rapporto qualità-prezzo al suo prezzo | xAI | 67.8 | 0,024 USD |
| Nemotron 3 Ultra 550B A55B | NVIDIA | 67.4 | 0,371 USD |
| Minimax M3 | Minimax | 67.3 | 0,060 USD |
| GPT-5.4 Mini xHigh | OpenAI | 66.4 | 0,334 USD |
| Qwen 3.6 27B | Alibaba | 64.0 | 0,202 USD |
| Gemini 3.5 Flash-Lite High | 63.9 | 0,069 USD | |
| Grok 4.3 | xAI | 62.3 | 0,061 USD |
Seleziona fino a tre modelli dalla classifica per confrontare i loro punteggi nelle sette categorie.
Ogni traccia va da 0 a 100. Più a destra è più forte; il numero in grassetto è il leader di quella riga.
| Categoria | Claude Fable 5.1 Max Effort | Union Alpha | DeepSeek V4.1 Flash Max Effort |
|---|---|---|---|
| Ragionamento | 91.7 | 80.8 | 86.7 |
| Programmazione | 86.4 | 82.1 | 80.0 |
| Programmazione agentica | 66.1 | 54.7 | 77.3 |
| Matematica | 97.0 | 95.3 | 93.3 |
| Analisi dei dati | 80.3 | 74.6 | 79.3 |
| Linguaggio | 89.5 | 85.9 | 81.2 |
| Seguimento delle istruzioni | 73.0 | 59.5 | 70.0 |
Ordina ogni categoria, restringi la lista e apri un modello per la sua suddivisione dei sotto-compiti pubblicata.
Claude Fable 5.1 Max Effort
Anthropic
83.4
Generale
1,212 USD
per attività riuscita
Claude Fable 5 Max Effort
Anthropic
83.0
Generale
1,439 USD
per attività riuscita
GPT-6 Astra Max Effort
OpenAI
82.2
Generale
0,736 USD
per attività riuscita
Muse Spark 1.3 xHigh Effort
Meta
81.6
Generale
0,219 USD
per attività riuscita
DeepSeek V4.1 Flash Max Effort
apertoDeepSeek
81.1
Generale
0,029 USD
per attività riuscita
GPT-5.6 Sol Max Effort
OpenAI
81.0
Generale
0,515 USD
per attività riuscita
GPT-5.5 Thinking xHigh Effort
OpenAI
80.2
Generale
0,435 USD
per attività riuscita
Claude 5 Opus Thinking Max Effort
Anthropic
80.1
Generale
0,699 USD
per attività riuscita
Kimi K3
apertoMoonshot AI
79.2
Generale
0,348 USD
per attività riuscita
Gemini 3.7 Flash High
78.8
Generale
0,157 USD
per attività riuscita
Qwen 3.8 Max
apertoAlibaba
78.5
Generale
0,275 USD
per attività riuscita
Muse Spark 1.2 xHigh Effort
Meta
78.0
Generale
0,375 USD
per attività riuscita
GPT-5.4 Thinking xHigh Effort
OpenAI
78.0
Generale
0,387 USD
per attività riuscita
Grok 4.6
xAI
78.0
Generale
0,207 USD
per attività riuscita
GPT-5.6 Terra Max Effort
OpenAI
77.9
Generale
0,352 USD
per attività riuscita
DeepSeek V4 Pro 0813
apertoDeepSeek
77.4
Generale
0,044 USD
per attività riuscita
Grok 4.7 xHigh
xAI
77.4
Generale
0,718 USD
per attività riuscita
Gemini 3.1 Pro Preview High
77.0
Generale
0,286 USD
per attività riuscita
DeepSeek V4 Flash Vision Exp
apertoDeepSeek
76.8
Generale
0,051 USD
per attività riuscita
Claude 4.7 Opus Thinking xHigh Effort
Anthropic
76.5
Generale
0,528 USD
per attività riuscita
Qwen 3.8 Flash Next
apertoAlibaba
76.2
Generale
0,042 USD
per attività riuscita
Claude 4.8 Opus Thinking Max Effort
Anthropic
76.2
Generale
0,983 USD
per attività riuscita
GLM-5.3
apertoZ.AI
76.1
Generale
0,450 USD
per attività riuscita
Union Alpha
Stealth
76.1
Generale
0,000 USD
per attività riuscita
Claude Sonnet 5 xHigh Effort
Anthropic
76.0
Generale
0,505 USD
per attività riuscita
Gemini 3.8 Flash High
75.8
Generale
0,307 USD
per attività riuscita
Grok 4.5
xAI
75.8
Generale
0,131 USD
per attività riuscita
Qwen3.8 27B
apertoAlibaba
75.3
Generale
0,094 USD
per attività riuscita
Muse Spark 1.1 xHigh Effort
Meta
75.3
Generale
0,198 USD
per attività riuscita
GPT-5.2 High
OpenAI
74.6
Generale
0,234 USD
per attività riuscita
Gemini 3.5 Flash High
74.6
Generale
0,249 USD
per attività riuscita
Claude 4.6 Opus Thinking High Effort
Anthropic
74.5
Generale
0,404 USD
per attività riuscita
DeepSeek V4 Flash 0731
apertoDeepSeek
74.2
Generale
0,060 USD
per attività riuscita
GPT-5.2 Codex
OpenAI
74.0
Generale
0,187 USD
per attività riuscita
Gemini 3.6 Flash High
73.6
Generale
0,235 USD
per attività riuscita
GPT-5.6 Luna Max Effort
OpenAI
73.6
Generale
0,169 USD
per attività riuscita
GLM-5.2
apertoZ.AI
73.2
Generale
0,225 USD
per attività riuscita
Qwen 3.7 Max
Alibaba
73.1
Generale
0,182 USD
per attività riuscita
Claude 4.6 Sonnet Thinking Medium Effort
Anthropic
73.0
Generale
0,306 USD
per attività riuscita
Claude 4.5 Opus Thinking High Effort
Anthropic
72.6
Generale
0,610 USD
per attività riuscita
Inkling xHigh Effort
apertoThinking Machines
71.9
Generale
0,310 USD
per attività riuscita
GLM-5.3 Flash
apertoZ.AI
71.6
Generale
0,031 USD
per attività riuscita
Kimi K2.6 Thinking
apertoMoonshot AI
70.5
Generale
0,169 USD
per attività riuscita
GPT-5.4 Nano xHigh
OpenAI
69.6
Generale
0,091 USD
per attività riuscita
ox-alpha-max
Stealth
69.2
Generale
0,000 USD
per attività riuscita
Qwen 3.6 Plus
Alibaba
68.9
Generale
0,227 USD
per attività riuscita
Kimi K2.7 Code
apertoMoonshot AI
68.4
Generale
0,100 USD
per attività riuscita
Grok Build 0.1
xAI
67.8
Generale
0,024 USD
per attività riuscita
Nemotron 3 Ultra 550B A55B
apertoNVIDIA
67.4
Generale
0,371 USD
per attività riuscita
Minimax M3
Minimax
67.3
Generale
0,060 USD
per attività riuscita
GPT-5.4 Mini xHigh
OpenAI
66.4
Generale
0,334 USD
per attività riuscita
Qwen 3.6 27B
apertoAlibaba
64.0
Generale
0,202 USD
per attività riuscita
Gemini 3.5 Flash-Lite High
63.9
Generale
0,069 USD
per attività riuscita
Grok 4.3
xAI
62.3
Generale
0,061 USD
per attività riuscita
Cost is the source-provided cost-per-successful-task metric for the selected scope.
These scores say which model is strongest at a set of held-out tasks. They don’t say which one fits your workflow, your budget, or the tools you already pay for — and the cheapest model that clears your bar usually beats the highest scorer.
What changed recently
Dated releases from every major lab, with the vendor's own claims labelled as claims.
OpenWhat it costs to run
Our tool catalogue carries the price we last verified and the date we checked it.
OpenWhat to actually build
A deployment plan picks the stack for one workflow at your budget — models included.
Open