← Back to directory
Anthropic
Claude Sonnet 4.5
Released Sep 29, 2025 · knowledge cutoff 2025-07
- Status
- Superseded
- Location
- United States
- Modality
- Multimodal
- Context window
- 200K
- Max output
- 64K
- Speed
- —
- Price ($/MTok in / out)
- $3 / $15
- Cost per task
- $0.41
- Open weights
- No
Benchmarks6 of 8 reported
GPQA Diamond
83.4%
SWE-bench Verified
77.2%
AIME (latest)
87%
Humanity's Last Exam
17.3%
LMArena Elo
1449
ARC-AGI-2
13.6%
MMLU-Pro
—
Terminal-Bench 2.1
—
Strengths & weaknesses
Strengths
- Sustains day-long autonomous runs without drifting off the original goal
- Drives a real desktop — clicks, scrolls and reads the screen — reliably enough to trust with a workflow
- Manages its own context, summarising and discarding what it no longer needs
Weaknesses
- Hits a ceiling on genuinely novel reasoning; more thinking doesn't produce a better answer
- Declares a task finished before verifying that it is
For developers
API model strings
Not researched
Licence
Proprietary — weights not released
Retirement
No retirement announced
Lineage
No recorded predecessor