MiniMax-M2.5
Released Feb 12, 2026 · knowledge cutoff unpublished
- Status
- Superseded
- Location
- China
- Modality
- Text
- Context window
- 205K
- Max output
- —
- Speed
- 95 tok/s · 1.67s to first answer token
- Price ($/MTok in / out)
- $0.15 / $1.20
- Cost per task
- —
- Open weights
- Yes
Benchmarks4 of 8 reported
GPQA Diamond
SWE-bench Verified
AIME (latest)
Humanity's Last Exam
MMLU-Pro
Terminal-Bench 2.1
LMArena Elo
ARC-AGI-2
Strengths & weaknesses
Strengths
- SWE-bench Verified jumped sharply over M2 in under four months, the fastest capability gain in the series so far
- Cheap enough to run continuously — MiniMax markets a roughly $1/hour cost at 100 output tokens/sec
- A separate 'Lightning' throughput mode trades some cost efficiency for materially faster output when latency matters more than price
Weaknesses
- Still text-only and capped at the same ~205K context as M2 — the long-context and multimodal jump waited for M3
- MMLU-Pro wasn't part of MiniMax's own benchmark table at launch, an unusual omission next to its detailed coding and agentic suite
- Superseded within four months by M3, the usual pace for this series
Officially announced 2026-02-12 (minimax.io/news/minimax-m25); MiniMax's own release material frames this as an iteration on M2/M2.1 rather than an explicit replacement of either, so predecessorId is left null. GPQA-Diamond 85.2, AIME25 86.3 and HLE 19.4% (without tools) are from the official Hugging Face README's comparison table; SWE-bench Verified 80.2%, which MiniMax describes as a new record for the series, is from the same announcement. MMLU-Pro is absent from MiniMax's own table; a third-party 74% figure (oddly lower than M2's official 82) could not be corroborated and looks inconsistent with M2.5's gains elsewhere, so left null. Context window confirmed at 204,800 tokens via MiniMax's own API docs — the same as M2; the 1M-token window is M3-only, despite some third-party trackers conflating the two. Pricing shown is the standard (non-Lightning) tier; the Lightning variant runs faster (100 tok/s vs 50) at roughly double the output price ($0.30/$2.40 vs $0.15/$1.20).
For developers
API model strings
- minimax
MiniMax-M2.5
Licence
MiniMax Modified MIT License · Permissive · commercial use permitted
Retirement
No retirement announced
Lineage
No recorded predecessor