Dapols
PlanesIA vs contrataciónPreciosHerramientasPuntos de referenciaBlog
Iniciar sesiónEncontrar mi flujo con IA
Dapols

The job big companies pay a forward-deployed engineer six figures to do — as a tool, for two figures.

Únete a la lista de vigilancia de precios

Verificado en páginas de proveedores. Todavía no enviamos correos; deja tu dirección y serás de los primeros en enterarte cuando empecemos.

Producto

  • Planes de IA para Negocios
  • Buscador de plan de IA
  • Biblioteca de habilidades de IA
  • Herramientas de IA
  • Precios

Explorar

  • IA vs contratación
  • Comprobador de coste del stack
  • Calculadora de ROI
  • Funciona con tus herramientas
  • Comparativas
  • Alternativas
  • Por industria
  • Por presupuesto

Ofertas

  • Planes de implementación de IA
  • Bigger or more complex? Tell us.
  • ¿Algo más grande o más complejo?

Empresa

  • Blog
  • Rastreador de modelos de IA
  • Enviar una herramienta
  • Contacto
  • Privacidad
  • Términos del servicio

Confianza y metodología

  • Acerca de
  • Metodología
  • Cómo clasificamos las herramientas de IA
  • IA desplegada en primera línea
  • Divulgación de afiliados
  • Actualizaciones de precios de herramientas de IA
  • Índice de precios de IA
  • Seguridad y privacidad de datos

© 2026 Dapols. Todos los derechos reservados.

support@dapols.comX

Pon la IA a trabajar, un flujo de trabajo medible a la vez.

Rastreador de modelos de IA

Cada lanzamiento importante de modelo de IA. Rastreado, verificado, explicado.

La frontera se mueve semanalmente. Rastreamos los modelos que importan (de los ocho laboratorios que realmente mueven el mercado) con números con fuentes y un 'por qué importa' en lenguaje sencillo.

Actualizado semanalmente· Último lanzamiento: 14 ago 2026

Mejores ahora mismo

Nuestra lectura de los puntos de referencia públicos, actualizada semanalmente. Opinión claramente etiquetada: haga clic en cualquier modelo para ver su fuente.

Best overall (general use)Claude Fable 5 Tops the LiveBench leaderboard (overall 83.0, 2026-07-31 snapshot). Grok 4.6 is not on that snapshot yet.Best frontier per dollarGrok 4.6 AA Intelligence Index 61, tied with GPT-5.6 Sol Max, at $2/$6 per 1M tokens (xAI, 12 Aug 2026). Weak on Terminal-Bench v3 (26%).Best for codingClaude Opus 5 Leads LiveBench agentic coding (65.2, 2026-07-31 snapshot) at half Fable 5's price; Fable 5 still holds the raw coding score (86.0).Best value / open-weightsDeepSeek V4 Flash (0731) Still the volume pick at $0.14/$0.28. V4 Pro-0813 is new weights on the same API id and promo price — not a LiveBench update yet.

Modelos de frontera recientes

ModeloLaboratorioLanzadoContextoPunto de referencia principal
GLM-5.3Z.ai14 ago 2026—Terminal-Bench 3.0 (Z.ai self-reported): 28.3, up from GLM-5.2's 4.6 (GPT-5.6 Sol 34.6)
Grok 4.6xAI12 ago 2026500KAPI price (per 1M in/out tokens, <200k prompt): $2 / $6
DeepSeek V4 Pro (0813)pesos abiertosDeepSeek12 ago 20261MAPI price (per 1M in/out tokens, promo): $0.435 / $0.87
Seedance 2.5vídeoByteDance31 jul 2026—Single-pass clip length: 30 seconds (up from 15)
DeepSeek V4 Flash (0731)pesos abiertosDeepSeek31 jul 20261MAPI price (per 1M in/out tokens): $0.14 / $0.28
Claude Opus 5Anthropic24 jul 2026—API price (per 1M in/out tokens): $5 / $25
Kimi K3pesos abiertosMoonshot AI16 jul 20261MTotal parameters (sparse MoE): 2.8T
GPT-5.6 (Sol · Terra · Luna)OpenAI9 jul 2026—API price, Sol (per 1M in/out tokens): $5 / $30

La línea de tiempo

agosto de 2026

ZA
GLM-5.3Z.ai· 14 ago 2026

Z.ai's 14 Aug 2026 release reuses GLM-5.2's base model — every gain comes from scaled post-training on long-horizon RL environments. Large jumps on agentic coding (Terminal-Bench 3.0 4.6 → 28.3, SWE-Marathon 19.4 → 42.5) and an unplanned cyber result (CyberGym 84.5). All figures are Z.ai's own; no independent evaluation exists yet. Weights promised two weeks after launch.

Por qué importa: It reaches Opus 4.8's coding score on roughly 2.4× fewer output tokens, so the saving is on the bill rather than the leaderboard — but check your API calls first: `thinking.type: "disabled"` is no longer supported and will fail on glm-5.3.

Terminal-Bench 3.0 (Z.ai self-reported): 28.3, up from GLM-5.2's 4.6 (GPT-5.6 Sol 34.6)Z.ai Code Bench, High effort — output tokens per task: 31.4% at ~50K vs Claude Opus 4.8's 29.5% at ~120KGLM Coding Plan off-peak quota rate: 50% of standard points outside 14:00–18:00 UTC+8, Mon–FriFuente
XA
Grok 4.6xAI· 12 ago 2026

xAI's 12 Aug 2026 successor to Grok 4.5: same $2/$6 API price, same 500K context, a longer post-training run aimed at long-running agents and visual/interactive work. A faster variant costs double. Knowledge cut-off is 1 Feb 2026.

Por qué importa: Frontier-class intelligence at the old 4.5 price — migrate off 4.5 with a one-line model swap. It is a knowledge-work pick, not the cheapest volume model and not the strongest terminal coding agent.

500K contextoAPI price (per 1M in/out tokens, <200k prompt): $2 / $6Artificial Analysis Intelligence Index: 61 (tied GPT-5.6 Sol Max; Fable 5 Max 62)Terminal-Bench v3.0: 26% (Sol Max 34.6%, Fable 5 Max 34.1%)Fuente
DS
DeepSeek V4 Pro (0813)DeepSeek· 12 ago 2026

The deepseek-v4-pro API id now serves DeepSeek-V4-Pro-0813 weights. Calling method and promo pricing are unchanged. DeepSeek's changelog still has no Pro GA entry — the last note (31 Jul) said an official Pro release would follow soon — so this is a weight drop, not a new product.

Por qué importa: Harder reasoning stays cheap on the same endpoint. Keep Flash ($0.14/$0.28) for volume; do not treat 0813 as a priced relaunch until DeepSeek posts a changelog.

1M contextoAPI price (per 1M in/out tokens, promo): $0.435 / $0.87Model version on the existing API id: DeepSeek-V4-Pro-0813Fuente

julio de 2026

BD
Seedance 2.5vídeoByteDance· 31 jul 2026

ByteDance's new video model generates 30-second audio-and-video clips in one pass, extends them over multiple rounds for multi-minute pieces, and edits by timestamp. At launch it runs inside Jimeng AI and Doubao Pro, with API access announced as coming via BytePlus ModelArk rather than available.

Por qué importa: The clearest jump yet in one-take video, but there is no general API to wire into a workflow — treat it as something to try inside a consumer video app, not a tool to rebuild your content process around until the API actually ships.

Single-pass clip length: 30 seconds (up from 15)Reference material per pass: 30 images, 10 videos, 10 audio clipsFuente
DS
DeepSeek V4 Flash (0731)DeepSeek· 31 jul 2026

The official V4 Flash checkpoint entered public beta on the existing deepseek-v4-flash API id, replacing the April preview. The legacy deepseek-chat and deepseek-reasoner ids were retired on 2026-07-24.

Por qué importa: Near-frontier results at roughly a tenth of typical API prices, with MIT-licensed open weights — the budget and privacy-sensitive pick just got better without changing price.

1M contextoAPI price (per 1M in/out tokens): $0.14 / $0.28Terminal Bench 2.1: 82.7SWE-bench Verified (thinking max): 79.0%Fuente
AN
Claude Opus 5Anthropic· 24 jul 2026

Anthropic's new everyday flagship: close to Fable 5 on most benchmarks at half the price ($5/$25 vs $10/$50 per MTok), with a low/medium/high effort toggle to trade cost against capability.

Por qué importa: Frontier-level output at half the flagship price — re-check which Claude tier your workflows actually need before renewing.

API price (per 1M in/out tokens): $5 / $25LiveBench overall (2026-07-31 snapshot): 80.1Fuente
KM
Kimi K3Moonshot AI· 16 jul 2026

Moonshot's 2.8-trillion-parameter sparse MoE — reportedly the largest open-weight model yet, with a 1M-token context window. Weights announced for late July; until then benchmark claims are vendor-reported.

Por qué importa: Open-weight models keep closing on the paid frontier — if you pay per-token for API work, the cheap tier just got stronger again.

1M contextoTotal parameters (sparse MoE): 2.8TFuente
OA
GPT-5.6 (Sol · Terra · Luna)OpenAI· 9 jul 2026

OpenAI's GPT-5.6 family in three tiers — Sol (frontier), Terra (balanced), Luna (fast/cheap) — general availability across ChatGPT, Codex, and the API on Jul 9 after a government-vetted limited preview in late June.

Por qué importa: Three clear price tiers make it easier to match the model to the job — most small-business tasks belong on the cheap tier, not the flagship.

API price, Sol (per 1M in/out tokens): $5 / $30API price, Terra (per 1M in/out tokens): $2 / $12Fuente
XA
Grok 4.5xAI· 8 jul 2026

xAI's first model built specifically for coding and agentic work, priced aggressively under Anthropic and OpenAI flagships with a 500K context window.

Por qué importa: Agentic coding on a budget is now a three-way price war — worth re-testing your coding stack before renewing anything.

500K contextoAPI price (per 1M in/out tokens): $2 / $6Terminal-Bench 2.1: 83.3%Fuente
AN
Claude Fable 5Anthropic· 1 jul 2026

A Mythos-class model made safe for general use — Anthropic's most capable generally available model, sitting above the Opus tier.

Por qué importa: The frontier of general-purpose reasoning just moved again; capable assistants keep getting cheaper to match.

Fuente

junio de 2026

AN
Claude Mythos 5Anthropic· 24 jun 2026

The Mythos-class model available to approved organizations without the general-use safety measures applied to Fable 5.

Por qué importa: Signals how fast the top tier is advancing — the same capability reaches everyone shortly after.

Fuente

mayo de 2026

AN
Claude Opus 4.8Anthropic· 1 may 2026

A strong all-round released model, widely cited as a top performer through mid-2026.

Por qué importa: A dependable default for hard reasoning, coding, and long-document work.

200K contextoFuente
OA
GPT-5.5OpenAI· 1 may 2026

OpenAI's mid-2026 frontier update, trading the top spot with Claude Opus on many benchmarks.

Por qué importa: Keeps the price-for-capability race moving — good news for anyone paying per token.

Fuente

marzo de 2026

OA
GPT-5.4OpenAI· 4 mar 2026

A March 2026 frontier release with a 1M-token context window and strong computer-use scores.

Por qué importa: Million-token context means it can read whole manuals, contracts, or codebases at once.

1M contextoOSWorld-Verified: 75.0%Fuente
DS
DeepSeek V4DeepSeek· 3 mar 2026

An open-weights frontier model with a 1M+ token context window and strong coding scores.

Por qué importa: Open weights + very low cost make it the value pick for budget-conscious and privacy-sensitive setups.

1M contextoHumanEval: 94.7%Fuente

febrero de 2026

GO
Gemini 3.1 ProGoogle· 1 feb 2026

Google's February 2026 frontier update to the Gemini 3 line, with a very large context window.

Por qué importa: Deep integration with Google Workspace makes it a natural fit if you live in Docs and Gmail.

1M contextoFuente
OA
GPT-5.3OpenAI· 1 feb 2026

A February 2026 iteration in the GPT-5 line ahead of the March 5.4 release.

Por qué importa: Part of the steady cadence keeping the mainstream assistant sharp.

Fuente
AN
Claude Opus 4.6Anthropic· 1 feb 2026

A February 2026 Opus update (alongside Sonnet 4.6), continuing Anthropic's rapid iteration.

Por qué importa: Reliability gains at the same price point — worth re-testing your prompts on each bump.

200K contextoFuente

noviembre de 2025

GO
Gemini 3 ProGoogle· 1 nov 2025

The Gemini 3 flagship that opened the current generation for Google.

Por qué importa: Set the bar for long-context multimodal work heading into 2026.

1M contextoFuente
AN
Claude Opus 4.5Anthropic· 1 nov 2025

A late-2025 Opus release that anchored Anthropic's top tier into 2026.

Por qué importa: The baseline many businesses standardized on before the 2026 wave.

200K contextoFuente

Cómo rastreamos esto

Actualizamos esto semanalmente como parte de la misma rutina que mantiene actualizados los precios de nuestras herramientas. Solo listamos lanzamientos realmente importantes de los laboratorios que mueven el mercado: modelos de texto de vanguardia, y ahora también modelos de vídeo, y cada estadística enlaza a su fuente. Sin fuente, no hay número. Cuando un modelo se anuncia pero aún no está disponible a través de una API, lo decimos, porque un modelo que aún no puedes comprar no es una recomendación.

Los modelos cambian cada semana. Tu plan se mantiene al día.

Tu plan es una foto fechada del mercado, y la suscripción mensual opcional la vuelve a revisar a medida que ocurren estos cambios, para que las herramientas y los precios recomendados sigan reflejando el mercado real.

Obtén mi plan de IA Ver planes