Meta’s closed reasoning model for agent work, announced 5 August 2026. It is mostly a coding upgrade on Spark 1.1, the model at the top of Scale AI’s MCP-Atlas leaderboard.
Paste a server URL, pick a model, and watch it call your tools in a real conversation. No install.
No server of your own? Leave it blank and use one of the public mock servers.
MCP Evals
Can Muse Spark 1.2 actually complete tasks with your server?
One chat shows the tool calls work. An eval writes real tasks from your tool schemas, has Muse Spark 1.2 drive each one, and reports where it got the wrong answer even though every call succeeded. Pick Muse Spark 1.2 as the driver.
Muse Spark 1.2 is the third Muse Spark release in four months: 1.0 in April, 1.1 in July, 1.2 in August. Artificial Analysis scores it 54 on its Intelligence Index, up from 51 for 1.1 and 43 for 1.0. It has a 1M-token window and takes text, image, video and PDF input. Unlike Llama, the weights are not public. Meta’s open model in this line is Muse Glimmer 30B, which is distilled from it.
These are the vendor’s own claims, not measurements of ours. Run the model against your own server to find out whether they hold for your tools.
Run the same four-step task on Muse Spark 1.2 and Muse Glimmer 30B against the complex-schema mock: create_user_profile, process_order, analyze_data, then configure_workflow. Glimmer is distilled from Spark, so the steps where they differ show what the open model lost. All the mock servers →
| Model | Model ID |
|---|---|
| Muse Spark 1.2 | meta/muse-spark-1.2 |