Wait Which Model?
← Back to directory
Moonshot AI

Kimi K2 Thinking

Released Nov 6, 2025 · knowledge cutoff 2025-04

Status
Superseded
Location
China
Modality
Text
Context window
262K
Max output
66K
Speed
Price ($/MTok in / out)
$0.60 / $2.50
Cost per task
Open weights
Yes
Benchmarks6 of 8 reported

MMLU-Pro

84.6%

GPQA Diamond

84.5%

SWE-bench Verified

71.3%

AIME (latest)

94.5%

Humanity's Last Exam

23.9%

LMArena Elo

1417

Terminal-Bench 2.1

ARC-AGI-2

Strengths & weaknesses

Strengths

  • Chains hundreds of tool calls without losing the goal — rare stamina for an open model
  • Research-style depth close to closed models of its moment
  • Weights available, so the entire agent loop can run in-house

Weaknesses

  • Slow, and it deliberates on nearly everything whether or not it needs to
  • Needs low-precision serving to be practical, which is fiddly to set up correctly

The strongest open reasoning model as of late 2025.

For developers

API model strings

Not researched

Licence

Not researched

Retirement

No retirement announced

Lineage

No recorded predecessor

News