Model
Model explorer

Muse Spark 1.3

CLOSED
Meta · Muse Spark family · released Sep 2, 2026

Meta's September 2, 2026 developer release in Muse Code and the Meta Model API — no consumer surface named. Pricing is unchanged since July at $1.25/$4.25 per Mtok with a 1M context, which is what makes it interesting: it posts frontier long-context and agentic-coding numbers at roughly an eighth of the Fable/Astra rate. Meta reports ~20% fewer tool calls and ~25% fewer tokens than 1.2 in its own engineers' use, and trained it to ask clarifying questions, invoke help when stuck and confirm before consequential actions. TWO CAVEATS THE LAUNCH TABLE DOES NOT MAKE OBVIOUS, both recorded here rather than papered over. First, the benchmarked configuration is Muse Spark 1.3 at MAX reasoning, which is still in limited preview pending safety testing; the generally-available build is xhigh, which Artificial Analysis scores 61 on Intelligence Index v4.1.1 against max's 62. Second, Meta's methodology states it ran 1.3, Claude Opus 5 and GPT-5.6 Sol at max but Muse Spark 1.2 at xhigh — so the eye-catching gains over its own predecessor are partly an effort-setting change that Meta does not break out, while the comparison against Opus 5 and Sol is like-for-like. The table also predates Claude Fable 5.1 by a day and omits Gemini 3.8 Flash, released the same day. On Meta's own scorecard it wins coding and long context and loses all six agent rows to Opus 5 or Sol. Terminal-Bench 2.1 pairs each model with its own coding-agent product rather than a common harness, so that row measures products, not models — noted per row. Its high overall position therefore describes the limited-preview max configuration, not the xhigh build a developer gets on the API today.

ReasoningCodingVisionFunction callingTool useAgentic
3269.8
Elo · rank #2
Parameters
Undisclosed
Active params
Context
1M tokens
Architecture
Proprietary multimodal reasoning model (undisclosed architecture); 1M-token context, trained across a diverse set of agent harnesses for long-horizon collaborative work
License
Proprietary
Languages
API price (in/out)
$1.25 / $4.25
Modalities
text · vision · audio · video
Benchmark results
Bar shows position within the tracked field; marker = field best
AutomationBenchAgents49.4%#2
best: Qwen3.8-Max-0902 · 50.8%
DeepSearchQAAgents89.4%#3
best: GPT-5.6 Sol · 93.0%
DeepSWECoding75.4%#1
best: this model · 75.4%
GDPval-AAAgents1754#5
best: Claude Fable 5 · 1932
JobBenchAgents64.9%#2
best: Claude Opus 5 · 65.7%
best: this model · 98.1%
best: GPT-6 Astra · 100.0%
OSWorld 2.0Agents66.9%#3
best: Claude Fable 5.1 · 77.9%
best: this model · 59.4%
Terminal-Bench 2.0Coding88.8%#3
best: Gemini 3.8 Flash · 89.4%
Run it locally
Closed weights — available via API only. No local deployment.
Input / M tok
$1.25
Output / M tok
$4.25
API price $1.25/$4.25 · each benchmark row carries its own source badge (see methodology)