Grok 4.6
Same $2 / $6 API price as Grok 4.5. xAI reports an Artificial Analysis Intelligence Index of 61 — tied with GPT-5.6 Sol Max, one point behind Fable 5 Max. Strong on knowledge-work evals; 26% on Terminal-Bench v3.
Read the Grok 4.6 briefingTolok ukur AI
A practical view of published model scores that helps you compare capability, category strengths, and run cost without losing the detail.
Latest benchmark & model updates
The leaderboard on this page is LiveBench's 25 June release, re-captured on 22 September 2026. It now includes Grok 4.7 xHigh and DeepSeek V4.1 Flash Max Effort. The model notes below are from the vendors' own pages, labelled as such.
Same $2 / $6 API price as Grok 4.5. xAI reports an Artificial Analysis Intelligence Index of 61 — tied with GPT-5.6 Sol Max, one point behind Fable 5 Max. Strong on knowledge-work evals; 26% on Terminal-Bench v3.
Read the Grok 4.6 briefingDeepSeek V4.1 Flash replaced V4 Flash on 10 Sep 2026 on the new deepseek-flash API id, at $0.15 / $0.60 per 1M tokens off-peak and $0.30 / $1.20 at peak. From 14 Sep, V4 Pro requests are routed to it too, so switch model ids now. LiveBench now scores V4.1 Flash Max Effort at 81.1 overall.
See it in the model trackerLeaderboard rows come only from published LiveBench data. We do not invent scores.
Best overall
Claude Fable 5.1 Max Effort leads this snapshot at 83.4 overall, $1.212 per successful task.
Best value per task
Union Alpha runs at $0.000 per successful task and still scores 76.1 overall.
Best open weights
DeepSeek V4.1 Flash Max Effort is the strongest open-weight model here at 81.1 overall — the option you can host yourself.
Best per capability
If your work is mostly one kind of task, these lead their category in this snapshot.
Every number above comes from the 2026-06-25 release dated 2026-06-25, published by a third-party benchmark and reproduced unaltered.
Grafik pertama menjawab “siapa yang terkuat”. Grafik kedua menjawab pertanyaan yang sebenarnya menghabiskan uang Anda: mana yang layak dibayar.
Semakin panjang semakin baik. Batang menunjukkan skor Keseluruhan dari 100; angka di samping setiap batang adalah nilai pastinya. Pilih satu baris untuk membawa model itu ke grafik biaya di bawah.
Menampilkan 12 teratas dari 54. Beralih ke tampilan perbandingan lengkap untuk semua model.
Ke atas lebih kuat, ke kiri lebih murah — jadi yang terbaik ada di kiri atas. Garis bertingkat menghubungkan model-model yang tidak bisa dikalahkan oleh yang lebih murah; apa pun di bawahnya kalah skor oleh sesuatu yang lebih murah.
Selected
Claude Fable 5.1 Max Effort
Anthropic
#5 · Di garis depan nilai.
Best buys (cheapest → strongest)
Tekan Tab ke dalam plot, lalu gunakan tombol panah untuk menelusuri model dari yang termurah hingga termahal.
| Model | Organisasi | Keseluruhan | per tugas berhasil |
|---|---|---|---|
| Claude Fable 5.1 Max Effort#5 · Nilai terbaik di harga tersebut | Anthropic | 83.4 | US$1,212 |
| Claude Fable 5 Max Effort | Anthropic | 83.0 | US$1,439 |
| GPT-6 Astra Max Effort#4 · Nilai terbaik di harga tersebut | OpenAI | 82.2 | US$0,736 |
| Muse Spark 1.3 xHigh Effort#3 · Nilai terbaik di harga tersebut | Meta | 81.6 | US$0,219 |
| DeepSeek V4.1 Flash Max Effort#2 · Nilai terbaik di harga tersebut | DeepSeek | 81.1 | US$0,029 |
| GPT-5.6 Sol Max Effort | OpenAI | 81.0 | US$0,515 |
| GPT-5.5 Thinking xHigh Effort | OpenAI | 80.2 | US$0,435 |
| Claude 5 Opus Thinking Max Effort | Anthropic | 80.1 | US$0,699 |
| Kimi K3 | Moonshot AI | 79.2 | US$0,348 |
| Gemini 3.7 Flash High | 78.8 | US$0,157 | |
| Qwen 3.8 Max | Alibaba | 78.5 | US$0,275 |
| Grok 4.6 | xAI | 78.0 | US$0,207 |
| Muse Spark 1.2 xHigh Effort | Meta | 78.0 | US$0,375 |
| GPT-5.4 Thinking xHigh Effort | OpenAI | 78.0 | US$0,387 |
| GPT-5.6 Terra Max Effort | OpenAI | 77.9 | US$0,352 |
| DeepSeek V4 Pro 0813 | DeepSeek | 77.4 | US$0,044 |
| Grok 4.7 xHigh | xAI | 77.4 | US$0,718 |
| Gemini 3.1 Pro Preview High | 77.0 | US$0,286 | |
| DeepSeek V4 Flash Vision Exp | DeepSeek | 76.8 | US$0,051 |
| Claude 4.7 Opus Thinking xHigh Effort | Anthropic | 76.5 | US$0,528 |
| Qwen 3.8 Flash Next | Alibaba | 76.2 | US$0,042 |
| Claude 4.8 Opus Thinking Max Effort | Anthropic | 76.2 | US$0,983 |
| GLM-5.3 | Z.AI | 76.1 | US$0,450 |
| Claude Sonnet 5 xHigh Effort | Anthropic | 76.0 | US$0,505 |
| Grok 4.5 | xAI | 75.8 | US$0,131 |
| Gemini 3.8 Flash High | 75.8 | US$0,307 | |
| Qwen3.8 27B | Alibaba | 75.3 | US$0,094 |
| Muse Spark 1.1 xHigh Effort | Meta | 75.3 | US$0,198 |
| GPT-5.2 High | OpenAI | 74.6 | US$0,234 |
| Gemini 3.5 Flash High | 74.6 | US$0,249 | |
| Claude 4.6 Opus Thinking High Effort | Anthropic | 74.5 | US$0,404 |
| DeepSeek V4 Flash 0731 | DeepSeek | 74.2 | US$0,060 |
| GPT-5.2 Codex | OpenAI | 74.0 | US$0,187 |
| GPT-5.6 Luna Max Effort | OpenAI | 73.6 | US$0,169 |
| Gemini 3.6 Flash High | 73.6 | US$0,235 | |
| GLM-5.2 | Z.AI | 73.2 | US$0,225 |
| Qwen 3.7 Max | Alibaba | 73.1 | US$0,182 |
| Claude 4.6 Sonnet Thinking Medium Effort | Anthropic | 73.0 | US$0,306 |
| Claude 4.5 Opus Thinking High Effort | Anthropic | 72.6 | US$0,610 |
| Inkling xHigh Effort | Thinking Machines | 71.9 | US$0,310 |
| GLM-5.3 Flash | Z.AI | 71.6 | US$0,031 |
| Kimi K2.6 Thinking | Moonshot AI | 70.5 | US$0,169 |
| GPT-5.4 Nano xHigh | OpenAI | 69.6 | US$0,091 |
| Qwen 3.6 Plus | Alibaba | 68.9 | US$0,227 |
| Kimi K2.7 Code | Moonshot AI | 68.4 | US$0,100 |
| Grok Build 0.1#1 · Nilai terbaik di harga tersebut | xAI | 67.8 | US$0,024 |
| Nemotron 3 Ultra 550B A55B | NVIDIA | 67.4 | US$0,371 |
| Minimax M3 | Minimax | 67.3 | US$0,060 |
| GPT-5.4 Mini xHigh | OpenAI | 66.4 | US$0,334 |
| Qwen 3.6 27B | Alibaba | 64.0 | US$0,202 |
| Gemini 3.5 Flash-Lite High | 63.9 | US$0,069 | |
| Grok 4.3 | xAI | 62.3 | US$0,061 |
Pilih hingga tiga model dari papan peringkat untuk membandingkan skor tujuh kategori mereka.
Setiap jalur berjalan dari 0 hingga 100. Semakin ke kanan semakin kuat; angka tebal adalah pemimpin di baris itu.
| Kategori | Claude Fable 5.1 Max Effort | Union Alpha | DeepSeek V4.1 Flash Max Effort |
|---|---|---|---|
| Penalaran | 91.7 | 80.8 | 86.7 |
| Pemrograman | 86.4 | 82.1 | 80.0 |
| Pemrograman agen | 66.1 | 54.7 | 77.3 |
| Matematika | 97.0 | 95.3 | 93.3 |
| Analisis data | 80.3 | 74.6 | 79.3 |
| Bahasa | 89.5 | 85.9 | 81.2 |
| Mengikuti instruksi | 73.0 | 59.5 | 70.0 |
Urutkan setiap kategori, persempit daftar, dan buka model untuk melihat rincian subtugas yang dipublikasikan.
Claude Fable 5.1 Max Effort
Anthropic
83.4
Keseluruhan
US$1,212
per tugas berhasil
Claude Fable 5 Max Effort
Anthropic
83.0
Keseluruhan
US$1,439
per tugas berhasil
GPT-6 Astra Max Effort
OpenAI
82.2
Keseluruhan
US$0,736
per tugas berhasil
Muse Spark 1.3 xHigh Effort
Meta
81.6
Keseluruhan
US$0,219
per tugas berhasil
DeepSeek V4.1 Flash Max Effort
TerbukaDeepSeek
81.1
Keseluruhan
US$0,029
per tugas berhasil
GPT-5.6 Sol Max Effort
OpenAI
81.0
Keseluruhan
US$0,515
per tugas berhasil
GPT-5.5 Thinking xHigh Effort
OpenAI
80.2
Keseluruhan
US$0,435
per tugas berhasil
Claude 5 Opus Thinking Max Effort
Anthropic
80.1
Keseluruhan
US$0,699
per tugas berhasil
Kimi K3
TerbukaMoonshot AI
79.2
Keseluruhan
US$0,348
per tugas berhasil
Gemini 3.7 Flash High
78.8
Keseluruhan
US$0,157
per tugas berhasil
Qwen 3.8 Max
TerbukaAlibaba
78.5
Keseluruhan
US$0,275
per tugas berhasil
Muse Spark 1.2 xHigh Effort
Meta
78.0
Keseluruhan
US$0,375
per tugas berhasil
GPT-5.4 Thinking xHigh Effort
OpenAI
78.0
Keseluruhan
US$0,387
per tugas berhasil
Grok 4.6
xAI
78.0
Keseluruhan
US$0,207
per tugas berhasil
GPT-5.6 Terra Max Effort
OpenAI
77.9
Keseluruhan
US$0,352
per tugas berhasil
DeepSeek V4 Pro 0813
TerbukaDeepSeek
77.4
Keseluruhan
US$0,044
per tugas berhasil
Grok 4.7 xHigh
xAI
77.4
Keseluruhan
US$0,718
per tugas berhasil
Gemini 3.1 Pro Preview High
77.0
Keseluruhan
US$0,286
per tugas berhasil
DeepSeek V4 Flash Vision Exp
TerbukaDeepSeek
76.8
Keseluruhan
US$0,051
per tugas berhasil
Claude 4.7 Opus Thinking xHigh Effort
Anthropic
76.5
Keseluruhan
US$0,528
per tugas berhasil
Qwen 3.8 Flash Next
TerbukaAlibaba
76.2
Keseluruhan
US$0,042
per tugas berhasil
Claude 4.8 Opus Thinking Max Effort
Anthropic
76.2
Keseluruhan
US$0,983
per tugas berhasil
GLM-5.3
TerbukaZ.AI
76.1
Keseluruhan
US$0,450
per tugas berhasil
Union Alpha
Stealth
76.1
Keseluruhan
US$0,000
per tugas berhasil
Claude Sonnet 5 xHigh Effort
Anthropic
76.0
Keseluruhan
US$0,505
per tugas berhasil
Gemini 3.8 Flash High
75.8
Keseluruhan
US$0,307
per tugas berhasil
Grok 4.5
xAI
75.8
Keseluruhan
US$0,131
per tugas berhasil
Qwen3.8 27B
TerbukaAlibaba
75.3
Keseluruhan
US$0,094
per tugas berhasil
Muse Spark 1.1 xHigh Effort
Meta
75.3
Keseluruhan
US$0,198
per tugas berhasil
GPT-5.2 High
OpenAI
74.6
Keseluruhan
US$0,234
per tugas berhasil
Gemini 3.5 Flash High
74.6
Keseluruhan
US$0,249
per tugas berhasil
Claude 4.6 Opus Thinking High Effort
Anthropic
74.5
Keseluruhan
US$0,404
per tugas berhasil
DeepSeek V4 Flash 0731
TerbukaDeepSeek
74.2
Keseluruhan
US$0,060
per tugas berhasil
GPT-5.2 Codex
OpenAI
74.0
Keseluruhan
US$0,187
per tugas berhasil
Gemini 3.6 Flash High
73.6
Keseluruhan
US$0,235
per tugas berhasil
GPT-5.6 Luna Max Effort
OpenAI
73.6
Keseluruhan
US$0,169
per tugas berhasil
GLM-5.2
TerbukaZ.AI
73.2
Keseluruhan
US$0,225
per tugas berhasil
Qwen 3.7 Max
Alibaba
73.1
Keseluruhan
US$0,182
per tugas berhasil
Claude 4.6 Sonnet Thinking Medium Effort
Anthropic
73.0
Keseluruhan
US$0,306
per tugas berhasil
Claude 4.5 Opus Thinking High Effort
Anthropic
72.6
Keseluruhan
US$0,610
per tugas berhasil
Inkling xHigh Effort
TerbukaThinking Machines
71.9
Keseluruhan
US$0,310
per tugas berhasil
GLM-5.3 Flash
TerbukaZ.AI
71.6
Keseluruhan
US$0,031
per tugas berhasil
Kimi K2.6 Thinking
TerbukaMoonshot AI
70.5
Keseluruhan
US$0,169
per tugas berhasil
GPT-5.4 Nano xHigh
OpenAI
69.6
Keseluruhan
US$0,091
per tugas berhasil
ox-alpha-max
Stealth
69.2
Keseluruhan
US$0,000
per tugas berhasil
Qwen 3.6 Plus
Alibaba
68.9
Keseluruhan
US$0,227
per tugas berhasil
Kimi K2.7 Code
TerbukaMoonshot AI
68.4
Keseluruhan
US$0,100
per tugas berhasil
Grok Build 0.1
xAI
67.8
Keseluruhan
US$0,024
per tugas berhasil
Nemotron 3 Ultra 550B A55B
TerbukaNVIDIA
67.4
Keseluruhan
US$0,371
per tugas berhasil
Minimax M3
Minimax
67.3
Keseluruhan
US$0,060
per tugas berhasil
GPT-5.4 Mini xHigh
OpenAI
66.4
Keseluruhan
US$0,334
per tugas berhasil
Qwen 3.6 27B
TerbukaAlibaba
64.0
Keseluruhan
US$0,202
per tugas berhasil
Gemini 3.5 Flash-Lite High
63.9
Keseluruhan
US$0,069
per tugas berhasil
Grok 4.3
xAI
62.3
Keseluruhan
US$0,061
per tugas berhasil
Cost is the source-provided cost-per-successful-task metric for the selected scope.
These scores say which model is strongest at a set of held-out tasks. They don’t say which one fits your workflow, your budget, or the tools you already pay for — and the cheapest model that clears your bar usually beats the highest scorer.
What changed recently
Dated releases from every major lab, with the vendor's own claims labelled as claims.
OpenWhat it costs to run
Our tool catalogue carries the price we last verified and the date we checked it.
OpenWhat to actually build
A deployment plan picks the stack for one workflow at your budget — models included.
Open