← Back to directory
DeepSeek
DeepSeek-V4
Released Apr 24, 2026 · knowledge cutoff unpublished
- Status
- Superseded
- Location
- China
- Modality
- Text
- Context window
- 1M
- Max output
- 384K
- Speed
- —
- Price ($/MTok in / out)
- $0.43 / $0.87
- Cost per task (max effort)
- $0.045
- Open weights
- Yes
Benchmarks4 of 8 reported
MMLU-Pro
87.5%
GPQA Diamond
90.1%
SWE-bench Verified
80.6%
Humanity's Last Exam
48.2%
Terminal-Bench 2.1
—
AIME (latest)
—
LMArena Elo
—
ARC-AGI-2
—
Strengths & weaknesses
Strengths
- Cheap enough to change what you are willing to spend tokens on at all
- Whole codebases fit in context, so cross-file work needs no chunking
- The small variant is fast and good enough for most work; the large one is there when it isn't
Weaknesses
- Reads as benchmark-tuned — stronger on tests than in day-to-day use
- Serving capacity is chip-constrained, so the top variant is frequently unavailable
- No dedicated reasoning line to escalate to
V4-Pro is a 1.6T/49B-active MoE; V4-Flash costs $0.14/$0.28. MMLU-Pro 87.5 / GPQA 90.1 are DeepSeek-reported for V4-Pro in Max reasoning mode; HLE 48.2 is third-party (LLM-Stats), same mode.
For developers
API model strings
Not researched
Licence
Not researched
Retirement
No retirement announced
Lineage
No recorded predecessor
News
In the news
- Jul 17, 2026
Artificial Analysis puts Kimi K3 third on its Intelligence Index
- Jul 5, 2026
"Price per 1M tokens is meaningless" argues for comparing models on cost per task
- May 1, 2026
DeepSeek V4 triggers scramble for Huawei Ascend 950 chips
- Apr 24, 2026
DeepSeek releases open-weights V4 family, tying Gemini 3.1 Pro on SWE-bench