Wait Which Model?
← Back to directory
DeepSeek

DeepSeek-V4

Released Apr 24, 2026 · knowledge cutoff unpublished

Status
Superseded
Location
China
Modality
Text
Context window
1M
Max output
384K
Speed
Price ($/MTok in / out)
$0.43 / $0.87
Cost per task (max effort)
$0.045
Open weights
Yes
Benchmarks4 of 8 reported

MMLU-Pro

87.5%

GPQA Diamond

90.1%

SWE-bench Verified

80.6%

Humanity's Last Exam

48.2%

Terminal-Bench 2.1

AIME (latest)

LMArena Elo

ARC-AGI-2

Strengths & weaknesses

Strengths

  • Cheap enough to change what you are willing to spend tokens on at all
  • Whole codebases fit in context, so cross-file work needs no chunking
  • The small variant is fast and good enough for most work; the large one is there when it isn't

Weaknesses

  • Reads as benchmark-tuned — stronger on tests than in day-to-day use
  • Serving capacity is chip-constrained, so the top variant is frequently unavailable
  • No dedicated reasoning line to escalate to

V4-Pro is a 1.6T/49B-active MoE; V4-Flash costs $0.14/$0.28. MMLU-Pro 87.5 / GPQA 90.1 are DeepSeek-reported for V4-Pro in Max reasoning mode; HLE 48.2 is third-party (LLM-Stats), same mode.

For developers

API model strings

Not researched

Licence

Not researched

Retirement

No retirement announced

Lineage

No recorded predecessor

News