Muse Spark

Meta · Released Apr 8, 2026
DFO score #44 / 694
51.6 leaderboard score
Speed
– Not reported
Input Price
– Not priced
Output Price
– Not priced
Context
– Not reported

Analysis Summary

Meta's Muse Spark has strong measured coding performance alongside good reasoning, instruction following, and long-context results. Its agentic capability is less pronounced than its coding result, so it appears better suited to assisted work than to complex autonomous execution. The supplied entry does not include pricing, a context-window size, modality, or tool and function-calling details, which limits assessment of deployment economics and workflow integration.

For an agency, the benchmark profile supports testing it on coding assistance, structured content, and analysis of longer inputs. Its instruction-following result is useful for work with detailed briefs, but client-facing use still needs review for factual accuracy and brand voice. Without listed pricing or interface details, it is difficult to recommend it as a volume default or an integrated agent without additional product evaluation.

Pilot it on tasks that match its measured strengths and compare outputs with the agency's current models. Confirm cost, available integrations, and context limits before assigning recurring production traffic.

Assessed September 23, 2026

Editorial notes

Muse Spark combines strong coding and instruction-following results with high measured reasoning and long-context performance; the supplied entry does not specify pricing, context size, or multimodal and tool capabilities.

The DFO score ranks models on the Artificial Analysis Intelligence Index and is refreshed as new models launch. The write-up and verdict are our own. Feedback?

DFO Verdict

Muse Spark combines strong coding and instruction-following results with high measured reasoning and long-context performance; the supplied entry does not specify pricing, context size, or multimodal and tool capabilities.

#44 of 694 overall Down 2 this week

Benchmarks

Intelligence Index 31.3 #44 of 433 · Artificial Analysis
Coding Index 58.6 #42 of 198 · Artificial Analysis
Agentic Index 28.7 #44 of 181 · Artificial Analysis
Creative Writing 1,465 #16 of 64 · Arena rating
Instruction Following 1,465 #32 of 65 · Arena rating
GPQA Diamond 88.4%
HLE 39.9%
SciCode 51.5%
TerminalBench Hard 45.5%
τ²-Bench 91.5%
IFBench 75.9%
LCR 69.7%

Artificial Analysis data refreshed Oct 6, 2026. Arena ratings from the Oct 2, 2026 text leaderboard, with style control. Bars: magenta = reasoning, ink = coding and agents, cyan = instructions and long context.

How Muse Spark compares

Muse Spark ranks #44 of 433 AI models we track for overall intelligence, #42 of 198 for coding, #44 of 181 for agentic tasks. Muse Spark is currently free to use via OpenRouter.

Position in the field
DFO score: ahead of 94% of models #44
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Muse Spark $0.00 in $0.00 out
Anthropic: Claude Opus 5.5 $4.00 in $20.00 out
Anthropic: Claude Sonnet 5.5 $2.00 in $10.00 out
Anthropic: Claude Fable 5.1 $10.00 in $50.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Frequently asked questions about Muse Spark

How much does Muse Spark cost?

Muse Spark is currently available for free via OpenRouter.

Is Muse Spark good for coding?

On our coding benchmark index, Muse Spark ranks #42 of 198 models, placing it in the top quartile of the field for code generation and debugging.

Who created Muse Spark?

Muse Spark is developed by Meta and was released on April 8, 2026.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.