Dapols
PlansAI employeesAI integrationPricingToolsBenchmarksBlog
Sign inFind my AI workflow
Dapols

Find one valuable AI workflow, choose the right stack, deploy it with human approval, and measure the result.

AI price watch — free weekly email

Which AI tools changed price, and what launched — for small businesses. Verified from vendor pages. No spam, unsubscribe anytime.

Product

  • AI Deployment Plans
  • AI Plan Finder
  • AI skills library
  • AI tools
  • Pricing

Offerings

  • AI Deployment Plans · $29
  • AI integration
  • Human-reviewed FDE Blueprint
  • One-workflow FDE Pilot

Company

  • Blog
  • AI model tracker
  • Submit a tool
  • Contact
  • Privacy
  • Terms of Service

Trust & methodology

  • About
  • Methodology
  • How we rank AI tools
  • Forward-deployed AI
  • Affiliate disclosure
  • AI tool pricing updates
  • Security & data privacy

© 2026 Dapols. All rights reserved.

support@dapols.comX

Put AI to work—one measurable workflow at a time.

Dapols
AI benchmarksLatest selected release · Jun 25, 2026

AI models, compared on quality and cost.

A source-attributed view of LiveBench scores that helps you compare capability, category strengths, and run cost without losing the detail.

Explore leaderboardLiveBench source

Score leader

GPT-5.6 Sol · Max effort

82.4 · Overall

Lowest run cost

DeepSeek V4 Pro

$0.050 · per successful task

Top open weights

Kimi K3

78.5 · Overall

01

Leaderboard

Sort every category, narrow the list, and open a model for its published subtask breakdown.

ModelCompare
OpenAI82.491.783.965.696.279.887.771.8$0.589
Anthropic80.889.786.046.996.080.590.775.8$1.573
Anthropic80.390.283.261.394.877.987.367.5$0.487
OpenAI79.890.678.268.094.979.382.964.6$0.497
Moonshot AI78.590.781.457.684.478.785.571.4$0.379
Google77.184.076.545.491.078.585.479.1$0.262
xAI76.387.268.659.890.873.082.871.5$0.128
DeepSeek71.682.770.042.690.774.578.162.4$0.050

Cost is the source-provided cost-per-successful-task metric for the selected LiveBench scope.

02

Insights

Scores are clearer beside cost and a model’s category profile. Use comparison to keep the trade-off visible.

Cost view

Quality vs. cost

Higher is stronger. Lower cost sits further left. Points use the source-provided cost metric.

GPT-5.6 Sol: 82.4, $0.589Claude Fable 5: 80.8, $1.573Claude 5 Opus Thinking: 80.3, $0.487GPT-5.6 Terra: 79.8, $0.497Kimi K3: 78.5, $0.379Gemini 3.1 Pro Preview: 77.1, $0.262Grok 4.5: 76.3, $0.128DeepSeek V4 Pro: 71.6, $0.050Cost per successful taskOverall
GPT-5.6 Sol: 82.4, $0.589. Claude Fable 5: 80.8, $1.573. Claude 5 Opus Thinking: 80.3, $0.487. GPT-5.6 Terra: 79.8, $0.497. Kimi K3: 78.5, $0.379. Gemini 3.1 Pro Preview: 77.1, $0.262. Grok 4.5: 76.3, $0.128. DeepSeek V4 Pro: 71.6, $0.050.

Compare

Category profile

Pick up to three models from the leaderboard to compare their seven category scores.

Use the compare buttons in the table to add up to three models.

Scores are a snapshot of the public LiveBench leaderboard; Dapols does not alter the benchmark values.
View source