Models / Meta/ Muse Spark 1.2

Muse Spark 1.2

Meta · released Aug 5, 2026

Input: text, images, audio, video and documents. Output: text.InputOutput
Type
Closed
Input
$1.25
Output
$4.25
Cached
$0.15

List price · per 1M tokens · Meta at 1M context · source ↗

Our take

Written Sep 17, 2026

Muse Spark 1.2 is a hosted-only model from Meta, so using it means choosing a host rather than running it yourself. Its measured strengths sit in maths and reasoning, and it takes text, images, files, audio and video in the same request.

Who should pick it

Reach for it on maths and reasoning work, where it scores 91.2% and 90% on the LiveBench categories measured 2026-06-25, or when a long document has to go into one request without being split first. Skip it if you need to run the model on your own machine, or if you need agent-driven coding fixes landed without review.

The case for it

  • Maths and reasoning are its highest measured categories, at 91.2% and 90% on LiveBench (evaluated 2026-06-25), so it suits analytical work you can check.
  • Text, images, files, audio and video all go into the same request, so a screenshot or a recording does not have to be described in words first.
  • The request capacity takes a long report or a stack of documents beside the question, though reliable recall across all of it is unverified in our data.

The case against it

  • Agentic coding is its weakest measured area at 57.58% on LiveBench (2026-06-25), against 77.54% on its general coding category, so set-piece code generation is the safer use.
  • We list no download for it, so using it means choosing a host; no licence is supplied in our data, so permissions are unverified.
  • The parameter count is not disclosed, so its size and memory footprint cannot be assessed.
00

How good is it?

A closed text model for everyday questions, coding and drafting, though it can be slow to change course when you give new instructions.

Good at
  • getting answers to everyday questionsArena Text (overall) · 5th of 168
  • drafts, rewrites and editingArena Creative Writing · 28th of 168
  • writing and completing codeArena Coding · 8th of 168
  • getting back on track after a step failsArena Agent · Recovery · 12th of 55
Less good at
  • changing course when you give new instructionsArena Agent · Steerability · 43rd of 55

EverydayGeneral questions and everyday reasoning

4.5 of 5

Arena Text (overall)5th of 168 · 1496

Arena Hard Prompts 10th of 168LiveBench Reasoning 10th of 58LiveBench Data Analysis 24th of 58LiveBench Mathematics 25th of 58

CodingWriting and fixing code on its own

4.5 of 5

Arena Coding8th of 168 · 1535

Arena Code (WebDev) 32nd of 95LiveBench Coding 32nd of 58

AgenticPlanning, calling tools, staying on task

2 of 5

Arena Agent37th of 55 · −0.03

LiveBench Agentic Coding 17th of 58

WritingDrafting and rewriting prose

3.5 of 5

Arena Creative Writing28th of 168 · 1448

LiveBench Language 32nd of 58
How it behaves in an agent loop
Tool usereaches for the right one, and does not invent one28th of 55
Steerabilitydoes what it was asked, and changes course when told43rd of 55
Recoverygets back on track after a command fails12th of 55
Task outcomefinishes what the session set out to do37th of 55

Placings on Arena's agent boards, from live sessions people ran themselves. A model can lead on one of these and sit mid-field on the others.

Other boards it appears on
Arena Instruction Following 17th of 168LiveBench Instruction Following 10th of 58LiveBench 17th of 58Arena Agent · Recovery 12th of 55Arena Agent · Tool use 28th of 55Arena Agent · Task outcome 37th of 55Arena Agent · Steerability 43rd of 55

Boards this model appears on that none of the ratings above are built on.

Every published score for this model19 scoresEvery figure we hold, from 19 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
77.95source ↗
57.58source ↗
77.54source ↗
76.46source ↗
74.33source ↗
78.57source ↗
91.2source ↗
90source ↗
−0.03source ↗
0.063source ↗
−0.056source ↗
−0.045source ↗
0.003source ↗
1535source ↗
1448source ↗
1511source ↗
1475source ↗
1496source ↗
1531source ↗
01

Where to rent it

Prices checked 1 hour ago — each listing carries its own date.

Cheapest published offer

Meta, through OpenRouter

The lab is also the cheapest we hold. The strip above and this offer are the same one, so nothing on this page undercuts Meta on 1M of context.

per 1M tokens
$1.25 in / $4.25 out
Context served
1M
Throughput
~181 tok/s
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenRouterOpenRouter's own listing$1.25 / $4.25checked 1 hour ago1Mnot measuredUnknownUnknownUnknown
MetaThrough OpenRouter$1.25 / $4.25checked 1 hour ago1M944K max reply181 tok/sNoYes30 daysUnknown

Across the 2 listings we hold: 1 says it does not train on prompts, 0 say they do and 1 does not say. 0 appear in the zero-retention registry we check; the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
OpenRouterOpenRouter's own listing✓✓✓
MetaThrough OpenRouter✓✓✓

Tool calling: 2 of 2 listings say yes. JSON output: 2 of 2 listings say yes. Strict schema: 2 of 2 listings say yes.

02

Models people weigh against Muse Spark 1.2

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored −0.03 via xHigh on Arena Agent
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.063 via xHigh on Arena Agent · Recovery
What movedleaderboard
Sep 25, 2026BenchmarkScored −0.056 via xHigh on Arena Agent · Steerability
What movedleaderboard
Sep 25, 2026BenchmarkScored −0.045 via xHigh on Arena Agent · Task outcome
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.003 via xHigh on Arena Agent · Tool use
What movedleaderboard
Sep 25, 2026BenchmarkScored 1535 via xHigh on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1448 via xHigh on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1511 via xHigh on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1475 via xHigh on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1496 via xHigh on Arena Text (overall)
What movedleaderboard

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 2 listings does not say whether it trains on prompts.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images, audio, video and documents in, text out
Catalogue slug
meta-muse-spark-1-2

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us