Wait Which Model?
← Back to directory
Anthropic

Claude Sonnet 4.5

Released Sep 29, 2025 · knowledge cutoff 2025-07

Status
Superseded
Location
United States
Modality
Multimodal
Context window
200K
Max output
64K
Speed
Price ($/MTok in / out)
$3 / $15
Cost per task
$0.41
Open weights
No
Benchmarks6 of 8 reported

GPQA Diamond

83.4%

SWE-bench Verified

77.2%

AIME (latest)

87%

Humanity's Last Exam

17.3%

LMArena Elo

1449

ARC-AGI-2

13.6%

MMLU-Pro

Terminal-Bench 2.1

Strengths & weaknesses

Strengths

  • Sustains day-long autonomous runs without drifting off the original goal
  • Drives a real desktop — clicks, scrolls and reads the screen — reliably enough to trust with a workflow
  • Manages its own context, summarising and discarding what it no longer needs

Weaknesses

  • Hits a ceiling on genuinely novel reasoning; more thinking doesn't produce a better answer
  • Declares a task finished before verifying that it is
For developers

API model strings

Not researched

Licence

Proprietary — weights not released

Retirement

No retirement announced

Lineage

No recorded predecessor

News