Muse Spark 1.3
Released Sep 2, 2026 · knowledge cutoff unpublished
- Status
- Frontier
- Location
- United States
- Modality
- Multimodal
- Context window
- 1M
- Max output
- —
- Speed (xhigh effort)
- 179 tok/s · 42s to first answer token
- Price ($/MTok in / out)
- $1.25 / $4.25
- Cost per task (xhigh effort)
- $0.55
- Open weights
- No
Benchmarks5 of 10 reported
GPQA Diamond
Terminal-Bench 2.1
Humanity's Last Exam
LMArena Elo
GDPval-AA v2
MMLU-Proretired
SWE-bench Verified
SWE-bench Pro
AIME (latest)retired
ARC-AGI-2
Strengths & weaknesses
Strengths
- Trained to stop and ask — checks in before irreversible actions and surfaces conflicting inputs rather than picking one silently
- Terser in agent loops than 1.2: roughly a fifth fewer tool calls and a quarter fewer tokens on the same coding work
- The million-token window is genuinely usable — near-ceiling recall on Meta's own 512K–1M retrieval evals rather than degrading past the halfway mark
- Streams fast for a reasoning model, so the cost of the long thinking phase is paid up front rather than throughout the answer
Weaknesses
- Long silence before the first answer token — the reasoning phase runs well over half a minute at the shipping effort setting
- Answer verbosity climbed sharply over 1.2, so the unchanged per-token price does not mean unchanged bills
- The headline agentic numbers come from a max reasoning mode that was still in limited preview at launch — what developers can call is the weaker xhigh configuration
Figures here describe the generally available xhigh reasoning configuration; a stronger `max` variant was still in safety testing at launch and only Artificial Analysis' partner preview had access. Terminal-Bench 2.1 conflicts: Artificial Analysis measured 85% (xhigh) and 86% (max) in its own harness, while Meta reports 88.8% for max — the AA xhigh figure is recorded as the one matching the shipping model. GPQA Diamond, HLE and GDPval-AA v2 are Artificial Analysis measurements of xhigh, not lab-reported. Meta lists a Contributor tier at $0.10/$0.20 per M tokens in exchange for training rights, and cached input at $0.15/M. Meta's roadmap promises a future 'Muse Spark open weights release' but 1.3 itself is closed. LMArena Elo 1493 is the "muse-spark-1.3-max" text-arena listing (arena.ai, 2026-09-13 update, rank 8).
For developers
API model strings
- meta
muse-spark-1.3 - openrouter
meta/muse-spark-1.3
Licence
Proprietary — weights not released
Retirement
No retirement announced
Lineage
No recorded predecessor