Dapols
プランAI vs 採用料金ツールベンチマークブログ
サインインAI業務フローを見つける
Dapols

The job big companies pay a forward-deployed engineer six figures to do — as a tool, for two figures.

価格ウォッチリストに参加

ベンダーページで確認済み。まだメールは配信していません。アドレスを残すと、開始時にいち早くお知らせします。

製品

  • ビジネスAIプラン
  • AIプランファインダー
  • AIスキルライブラリ
  • AIツール
  • 料金

探索

  • AI vs 採用
  • スタック費用チェッカー
  • ROI計算機
  • あなたのツールと連携
  • 比較
  • 代替案
  • 業界別
  • 予算別

提供内容

  • AI導入プラン
  • Bigger or more complex? Tell us.
  • もっと大きく、複雑ですか?

会社

  • ブログ
  • AIモデルトラッカー
  • ツールを提出する
  • お問い合わせ
  • プライバシー
  • 利用規約

信頼と方法論

  • 概要
  • 方法論
  • AIツールのランキング方法
  • 現場展開型AI
  • アフィリエイト開示
  • AIツールの価格更新
  • AI価格インデックス
  • セキュリティとデータプライバシー

© 2026 Dapols. All rights reserved.

support@dapols.comX

AIを仕事に。成果の測れる業務フローをひとつずつ。

AIモデルトラッカー

すべての主要なAIモデルリリース。追跡、検証、解説。

最前線は毎週動きます。実際に市場を動かす8つの研究室から、重要なモデルを追跡 — 出典付きの数値と、わかりやすい言葉で書いた「なぜ重要か」を添えて。

毎週更新· 最新リリース: 2026年8月14日

現在のベスト

公開ベンチマークの当社の読み解き、毎週更新。意見であり明確にラベル付け — 各モデルをクリックして出典を確認。

Best overall (general use)Claude Fable 5 Tops the LiveBench leaderboard (overall 83.0, 2026-07-31 snapshot). Grok 4.6 is not on that snapshot yet.Best frontier per dollarGrok 4.6 AA Intelligence Index 61, tied with GPT-5.6 Sol Max, at $2/$6 per 1M tokens (xAI, 12 Aug 2026). Weak on Terminal-Bench v3 (26%).Best for codingClaude Opus 5 Leads LiveBench agentic coding (65.2, 2026-07-31 snapshot) at half Fable 5's price; Fable 5 still holds the raw coding score (86.0).Best value / open-weightsDeepSeek V4 Flash (0731) Still the volume pick at $0.14/$0.28. V4 Pro-0813 is new weights on the same API id and promo price — not a LiveBench update yet.

最近のフロンティアモデル

モデル研究室リリース日コンテキスト主要ベンチマーク
GLM-5.3Z.ai2026年8月14日—Terminal-Bench 3.0 (Z.ai self-reported): 28.3, up from GLM-5.2's 4.6 (GPT-5.6 Sol 34.6)
Grok 4.6xAI2026年8月12日500KAPI price (per 1M in/out tokens, <200k prompt): $2 / $6
DeepSeek V4 Pro (0813)オープンウェイトDeepSeek2026年8月12日1MAPI price (per 1M in/out tokens, promo): $0.435 / $0.87
Seedance 2.5動画ByteDance2026年7月31日—Single-pass clip length: 30 seconds (up from 15)
DeepSeek V4 Flash (0731)オープンウェイトDeepSeek2026年7月31日1MAPI price (per 1M in/out tokens): $0.14 / $0.28
Claude Opus 5Anthropic2026年7月24日—API price (per 1M in/out tokens): $5 / $25
Kimi K3オープンウェイトMoonshot AI2026年7月16日1MTotal parameters (sparse MoE): 2.8T
GPT-5.6 (Sol · Terra · Luna)OpenAI2026年7月9日—API price, Sol (per 1M in/out tokens): $5 / $30

タイムライン

2026年8月

ZA
GLM-5.3Z.ai· 2026年8月14日

Z.ai's 14 Aug 2026 release reuses GLM-5.2's base model — every gain comes from scaled post-training on long-horizon RL environments. Large jumps on agentic coding (Terminal-Bench 3.0 4.6 → 28.3, SWE-Marathon 19.4 → 42.5) and an unplanned cyber result (CyberGym 84.5). All figures are Z.ai's own; no independent evaluation exists yet. Weights promised two weeks after launch.

なぜ重要か: It reaches Opus 4.8's coding score on roughly 2.4× fewer output tokens, so the saving is on the bill rather than the leaderboard — but check your API calls first: `thinking.type: "disabled"` is no longer supported and will fail on glm-5.3.

Terminal-Bench 3.0 (Z.ai self-reported): 28.3, up from GLM-5.2's 4.6 (GPT-5.6 Sol 34.6)Z.ai Code Bench, High effort — output tokens per task: 31.4% at ~50K vs Claude Opus 4.8's 29.5% at ~120KGLM Coding Plan off-peak quota rate: 50% of standard points outside 14:00–18:00 UTC+8, Mon–Fri出典
XA
Grok 4.6xAI· 2026年8月12日

xAI's 12 Aug 2026 successor to Grok 4.5: same $2/$6 API price, same 500K context, a longer post-training run aimed at long-running agents and visual/interactive work. A faster variant costs double. Knowledge cut-off is 1 Feb 2026.

なぜ重要か: Frontier-class intelligence at the old 4.5 price — migrate off 4.5 with a one-line model swap. It is a knowledge-work pick, not the cheapest volume model and not the strongest terminal coding agent.

500K コンテキストAPI price (per 1M in/out tokens, <200k prompt): $2 / $6Artificial Analysis Intelligence Index: 61 (tied GPT-5.6 Sol Max; Fable 5 Max 62)Terminal-Bench v3.0: 26% (Sol Max 34.6%, Fable 5 Max 34.1%)出典
DS
DeepSeek V4 Pro (0813)DeepSeek· 2026年8月12日

The deepseek-v4-pro API id now serves DeepSeek-V4-Pro-0813 weights. Calling method and promo pricing are unchanged. DeepSeek's changelog still has no Pro GA entry — the last note (31 Jul) said an official Pro release would follow soon — so this is a weight drop, not a new product.

なぜ重要か: Harder reasoning stays cheap on the same endpoint. Keep Flash ($0.14/$0.28) for volume; do not treat 0813 as a priced relaunch until DeepSeek posts a changelog.

1M コンテキストAPI price (per 1M in/out tokens, promo): $0.435 / $0.87Model version on the existing API id: DeepSeek-V4-Pro-0813出典

2026年7月

BD
Seedance 2.5動画ByteDance· 2026年7月31日

ByteDance's new video model generates 30-second audio-and-video clips in one pass, extends them over multiple rounds for multi-minute pieces, and edits by timestamp. At launch it runs inside Jimeng AI and Doubao Pro, with API access announced as coming via BytePlus ModelArk rather than available.

なぜ重要か: The clearest jump yet in one-take video, but there is no general API to wire into a workflow — treat it as something to try inside a consumer video app, not a tool to rebuild your content process around until the API actually ships.

Single-pass clip length: 30 seconds (up from 15)Reference material per pass: 30 images, 10 videos, 10 audio clips出典
DS
DeepSeek V4 Flash (0731)DeepSeek· 2026年7月31日

The official V4 Flash checkpoint entered public beta on the existing deepseek-v4-flash API id, replacing the April preview. The legacy deepseek-chat and deepseek-reasoner ids were retired on 2026-07-24.

なぜ重要か: Near-frontier results at roughly a tenth of typical API prices, with MIT-licensed open weights — the budget and privacy-sensitive pick just got better without changing price.

1M コンテキストAPI price (per 1M in/out tokens): $0.14 / $0.28Terminal Bench 2.1: 82.7SWE-bench Verified (thinking max): 79.0%出典
AN
Claude Opus 5Anthropic· 2026年7月24日

Anthropic's new everyday flagship: close to Fable 5 on most benchmarks at half the price ($5/$25 vs $10/$50 per MTok), with a low/medium/high effort toggle to trade cost against capability.

なぜ重要か: Frontier-level output at half the flagship price — re-check which Claude tier your workflows actually need before renewing.

API price (per 1M in/out tokens): $5 / $25LiveBench overall (2026-07-31 snapshot): 80.1出典
KM
Kimi K3Moonshot AI· 2026年7月16日

Moonshot's 2.8-trillion-parameter sparse MoE — reportedly the largest open-weight model yet, with a 1M-token context window. Weights announced for late July; until then benchmark claims are vendor-reported.

なぜ重要か: Open-weight models keep closing on the paid frontier — if you pay per-token for API work, the cheap tier just got stronger again.

1M コンテキストTotal parameters (sparse MoE): 2.8T出典
OA
GPT-5.6 (Sol · Terra · Luna)OpenAI· 2026年7月9日

OpenAI's GPT-5.6 family in three tiers — Sol (frontier), Terra (balanced), Luna (fast/cheap) — general availability across ChatGPT, Codex, and the API on Jul 9 after a government-vetted limited preview in late June.

なぜ重要か: Three clear price tiers make it easier to match the model to the job — most small-business tasks belong on the cheap tier, not the flagship.

API price, Sol (per 1M in/out tokens): $5 / $30API price, Terra (per 1M in/out tokens): $2 / $12出典
XA
Grok 4.5xAI· 2026年7月8日

xAI's first model built specifically for coding and agentic work, priced aggressively under Anthropic and OpenAI flagships with a 500K context window.

なぜ重要か: Agentic coding on a budget is now a three-way price war — worth re-testing your coding stack before renewing anything.

500K コンテキストAPI price (per 1M in/out tokens): $2 / $6Terminal-Bench 2.1: 83.3%出典
AN
Claude Fable 5Anthropic· 2026年7月1日

A Mythos-class model made safe for general use — Anthropic's most capable generally available model, sitting above the Opus tier.

なぜ重要か: The frontier of general-purpose reasoning just moved again; capable assistants keep getting cheaper to match.

出典

2026年6月

AN
Claude Mythos 5Anthropic· 2026年6月24日

The Mythos-class model available to approved organizations without the general-use safety measures applied to Fable 5.

なぜ重要か: Signals how fast the top tier is advancing — the same capability reaches everyone shortly after.

出典

2026年5月

AN
Claude Opus 4.8Anthropic· 2026年5月1日

A strong all-round released model, widely cited as a top performer through mid-2026.

なぜ重要か: A dependable default for hard reasoning, coding, and long-document work.

200K コンテキスト出典
OA
GPT-5.5OpenAI· 2026年5月1日

OpenAI's mid-2026 frontier update, trading the top spot with Claude Opus on many benchmarks.

なぜ重要か: Keeps the price-for-capability race moving — good news for anyone paying per token.

出典

2026年3月

OA
GPT-5.4OpenAI· 2026年3月4日

A March 2026 frontier release with a 1M-token context window and strong computer-use scores.

なぜ重要か: Million-token context means it can read whole manuals, contracts, or codebases at once.

1M コンテキストOSWorld-Verified: 75.0%出典
DS
DeepSeek V4DeepSeek· 2026年3月3日

An open-weights frontier model with a 1M+ token context window and strong coding scores.

なぜ重要か: Open weights + very low cost make it the value pick for budget-conscious and privacy-sensitive setups.

1M コンテキストHumanEval: 94.7%出典

2026年2月

GO
Gemini 3.1 ProGoogle· 2026年2月1日

Google's February 2026 frontier update to the Gemini 3 line, with a very large context window.

なぜ重要か: Deep integration with Google Workspace makes it a natural fit if you live in Docs and Gmail.

1M コンテキスト出典
OA
GPT-5.3OpenAI· 2026年2月1日

A February 2026 iteration in the GPT-5 line ahead of the March 5.4 release.

なぜ重要か: Part of the steady cadence keeping the mainstream assistant sharp.

出典
AN
Claude Opus 4.6Anthropic· 2026年2月1日

A February 2026 Opus update (alongside Sonnet 4.6), continuing Anthropic's rapid iteration.

なぜ重要か: Reliability gains at the same price point — worth re-testing your prompts on each bump.

200K コンテキスト出典

2025年11月

GO
Gemini 3 ProGoogle· 2025年11月1日

The Gemini 3 flagship that opened the current generation for Google.

なぜ重要か: Set the bar for long-context multimodal work heading into 2026.

1M コンテキスト出典
AN
Claude Opus 4.5Anthropic· 2025年11月1日

A late-2025 Opus release that anchored Anthropic's top tier into 2026.

なぜ重要か: The baseline many businesses standardized on before the 2026 wave.

200K コンテキスト出典

追跡方法

これはツール価格を最新に保つ同じルーチンの一環として毎週更新されます。市場を動かすラボからの真に主要なリリースのみを掲載します。フロンティアテキストモデル、そして現在はビデオモデルも同様です。すべての統計は出典にリンクしています。出典がなければ数値もありません。APIからまだ利用できないモデルについては、その旨を明記します。まだ購入できないモデルは推奨ではないからです。

モデルは毎週変わります。あなたのプランも追いつきます。

プランは市場のある時点を写したものです。任意の月額サブスクリプションは、こうした変化が起きるたびに内容を見直すため、おすすめのツールと価格が今の市場に合った状態のまま保たれます。

AIプランを入手 プランを見る