DeepSeek-V4-Pro-0813
Released Aug 13, 2026 · knowledge cutoff unpublished
- Status
- Superseded
- Location
- China
- Modality
- Text
- Context window
- 1M
- Max output
- 384K
- Speed (max effort)
- 78 tok/s · 26s to first answer token
- Price ($/MTok in / out)
- $1.32 / $3.96
- Cost per task
- —
- Open weights
- Yes
Benchmarks6 of 10 reported
GPQA Diamond
SWE-bench Verified
Terminal-Bench 2.1
Humanity's Last Exam
LMArena Elo
ARC-AGI-2
MMLU-Proretired
SWE-bench Pro
AIME (latest)retired
GDPval-AA v2
Strengths & weaknesses
Strengths
- Same 1.6T/49B MoE structure as April's preview with a DSpark speculative-decoding module bolted on — the agentic gains arrive without a new architecture to re-integrate
- Cyber-defense and terminal work are where the re-post-training shows most; independent evaluators single out CTF-style tasks as its clearest jump over the preview
- Thinking and non-thinking modes with three effort levels behind one endpoint, so a quick lookup and a long agent run share the same model id
- MIT-licensed weights mirrored on Fireworks, DeepInfra and OpenRouter within days — no dependence on DeepSeek's capacity-constrained first-party API
Weaknesses
- Headline agent scores come from DeepSeek's own unreleased minimal-mode harness — on the Terminal-Bench reference harness it lands below its cheaper Flash sibling, and nobody has reproduced the card's numbers
- Multi-turn thinking mode requires replaying the previous turn's reasoning verbatim — the most common cause of 400 errors when migrating existing agent loops
- Landed as a stealth update with no launch post (the brief notice on DeepSeek's site was pulled within a day), so documentation is thin and behavior is still being characterised
The production release of DeepSeek-V4-Pro, superseding April's preview per DeepSeek's own model card: same 1.6T/49B MoE with a DSpark speculative-decoding module, all gains from re-post-training. Pricing is the peak-hour rate that took effect 2026-08-16 (peak = 01:00–04:00 and 06:00–10:00 UTC); off-peak is 50% lower at $0.66/$1.98 and cache-hit input is cheaper still — roughly a 3–4x rise from the preview's $0.435/$0.87. Terminal-Bench 2.1 87.9 and HLE 42.7% (60.0% with tools) are DeepSeek's own figures at max effort in its minimal-mode harness; independent runs on the Terminus 2 reference harness report 54.68 (V4-Flash-0731: 67.04). GPQA Diamond 93 is Artificial Analysis' measurement as displayed on its charts, not lab-reported — DeepSeek's own 0813 card figure could not be confirmed directly (one unofficial site cites 92.8; the 90.1 many aggregators list is April's preview). SWE-bench Verified 96.4 is third-party (Vals AI, #2 on its board behind Claude Opus 5) — DeepSeek published no SWE-bench figure and the 80.6 some trackers carry is the preview's. Also reported by DeepSeek: DeepSWE 62.7, CyberGym 83.3, NL2Repo 61.5, Toolathlon 74.1, Agents' Last Exam 25.7, AutomationBench 31.8. Artificial Analysis scores it 53 on its Intelligence Index — 8 points above April's preview but only 1 above V4-Flash-0731 on the current index version — measuring 78.1 tok/s and 25.67 s to first answer token on DeepSeek's endpoint at max effort; its per-task cost figure could not be retrieved. MMLU-Pro and AIME are unpublished; the arena.ai listing and ARC Prize result arrived after launch (below). Max output is the 384K DeepSeek recommends for high/max effort. The API went live 2026-08-12 under the unchanged `deepseek-v4-pro` id; the Hugging Face card and DeepSeek's release notice followed on 2026-08-13. Updated 2026-09-15: DeepSeek's V4.1-Flash launch post (2026-09-10) said all `deepseek-v4-pro` requests would route to V4.1-Flash from 2026-09-14 04:00 UTC until V4.1-Pro ships, but the API change log and pricing page then reversed this "in response to user demand" — V4 Pro service continues after 2026-09-14 with billing unchanged, so no retirement is recorded. ARC-AGI-2 61.3% is ARC Prize-verified at max effort ($0.60/task; high 59.7%, low 56.3%), published 2026-08-13 at arcprize.org/results/deepseek-v4-pro-0813. LMArena Elo 1463 is the "deepseek-v4-pro-high-20260813" text-arena listing (arena.ai, 2026-09-13 update); the undated "deepseek-v4-pro" listing sits at 1457.
For developers
API model strings
- deepseek
deepseek-v4-pro
Licence
MIT License · Permissive · commercial use permitted
Retirement
No retirement announced
Lineage
Replaces DeepSeek-V4