I wanted to try something quick with Muse Spark 1.3, so I had it build this pirate platformer. It was incredibly fast and worked flawlessly at one-shot! Now I’m enhancing the artwork, sprites and adding new levels 😊This is such a fun coding model, it just works, fast and cheap!
DataCamp's Josep Ferrer ran Muse Spark 1.3 on three real coding tasks. Two used 23–32% fewer completion tokens, but a refactor used 70% more, for a net 12% cost increase.
$META just went from 3.5% to 45.4% token share on OpenCode in just over two weeks
Muse Spark 1.3 being good + free is enough to become the default for most users
Default gets you usage → usage gets you data → data makes the next model better
Anthropic and OpenAI can’t afford
Dilmer Valecillos uses Muse Code with Muse Spark 1.2 and the Unity CLI to run tests, validate changes, port a Mini Golf game to other platforms and convert it to VR.
Muse Spark 1.2 (xHigh) by @AIatMeta is #14 in the Code Arena: WebDev, with 1,545 pts!
This is an improvement from Muse Spark 1.1 at #18. See its biggest gains by category in the post below.
Congrats to the @AIatMeta team on this release!
BREAKING: Muse Spark 1.3 (xhigh) takes 1st overall on Website Arena with an Elo of 1362!
This is a jump of 5 positions from Muse Spark 1.2, establishing a new Pareto frontier for Speed and Price.
Only a month after the release of Muse Spark 1.2, @AIatMeta has topped this
Muse Spark 1.1 has entered the Code Arena: Frontend at #9!
Muse Spark 1.1 reshapes the cost-performance Pareto Frontier by scoring 1541 at a blended $3.5M ($1.25 per input MToken, $4.25 per output MToken). This is frontier performance at a fraction of the price.
Congrats to
Muse Spark 1.3 is now available on OpenRouter!
Built for long-running agentic, multi-agent, and coding workflows. It tracks what it learns, handles conflicting inputs, and asks for clarification or confirmation when needed.
Use it now: openrouter.ai/meta/muse-spar…
Box added Muse Spark 1.3 to Box AI, reporting it runs 42% faster than Muse Spark 1.2 with roughly a third fewer tokens, and lifts financial services accuracy from 66% to 75% on Box's eval.
Meta has released Muse Spark 1.3, their fourth Muse Spark model release in five months. Muse Spark 1.3 (max), which is in limited preview for Meta’s partners, scores 62 on the Artificial Analysis Intelligence Index, behind only Claude Fable 5.1 and Claude Opus 5. The variant
✨ Three New Models, Sharper Lore Awareness & Smoother Storytelling Across Complex Worlds
This update introduces three new storytelling models, improves how Lore is recognized and prioritized during a story, and brings another round of reliability improvements to the model experience.
🎭 Three New Voices 💎 Gemini 3.8 Flash — Pro+ Google’s newest fast storyteller, with polished prose, responsive pacing, and a steady grip on complex scenes.
Context support:
🔹 16K on Pro
🔹 48K on Ultra
🔹 80K on Legendary
Context limits subject to change.
✨ Muse Spark 1.3 — Plus+ Vivid character interplay, strong world-state awareness, and deliberate continuity as stories evolve.
Context support:
🔹 16K on Plus
🔹 32K on Pro
🔹 64K on Ultra
🔹 100K on Legendary
🌐 Hunyuan 4 Preview — Pro+ Built for ambitious living worlds, shifting relationships, and large casts with lasting consequences.
Context support:
🔹 16K on Pro
🔹 32K on Ultra
🔹 48K on Legendary
📚 Sharper Lore Awareness Lore is getting better at recognizing what matters in the current scene.
🏷️ Smarter Lore activation — Lore cards can now activate when their title or character name appears naturally in the story, even
Muse Spark 1.2 (xHigh) by @AIatMeta is now in Agent Arena, with a net improvement of +2.1%!
Agent Arena measures models on millions of real-world, long-horizon agentic tasks. We use causal tracing methodology to measure a model's net improvement, indicating how much it improves
Meta just released Muse Spark 1.1 and is the new SOTA on MedScribe and TaxEval, taking the top spot from Fable 5 while being 10x cheaper and twice as fast. Meta currently holds the top 2 spots on TaxEval
It is also the new #1 on Harvey's Legal Agent Bench, dethroning Grok 4.5
1/ we just publicly released Muse Spark 1.3 max!
we see significantly stronger coding and agentic performance on muse spark 1.3 max, so would strongly recommend trying it out even if you've already tried muse spark 1.3 high or muse spark 1.3 xhigh.
Muse Spark 1.1 just launched and it's their most capable coding agent model yet.
On Terminal-Bench 2.1 it scores 80.0%, in the same cluster as Opus 4.8 (82.7%) and GPT 5.5 (83.4%).
Use it in Cline with the Meta API!
Muse Spark 1.3 vs 1.2: sculpt viking figurines
@AIatMeta dropped Muse Spark 1.3 today — we ran it against Muse Spark 1.2
both models got the same brief: three collectible 3D figurines — a viking helmet, a diamond-studded axe, a longship with a crew — one self-contained HTML
Last week the Intelligence Index vs Cost Pareto frontier moved out substantially. Claude Fable 5.1, Muse Spark 1.3, and GPT-6 Astra each set a new point in efficient intelligence
Link to analysis: artificialanalysis.ai/#intelligence-…
Muse Spark 1.3 Max by @AIatMeta has reshaped the Pareto frontier for Code Arena: WebDev!
Meta's latest model at Max reasoning is doing something interesting on the Arena Pareto frontier: it's the only model holding down the wide price band between Qwen3.8-max ($5/MToken) and
Box's Complex Work Eval finds Muse Spark 1.1 up to 5-6 points above the top-tier composite on structured work, and nearly 30 points ahead on cost-optimization analysis.
Meta has released Muse Spark 1.2. It's their third release in four months and scores 54 on the Artificial Analysis Intelligence Index, significantly improving agentic knowledge work capabilities over prior releases and putting Meta next to SpaceXAI in a tie for third place
I swear Muse Spark 1.3 is a really good model. Like I don't know if it's just my setup or something. But it just constantly surprises me. So in Synara there's a thread/agents tool where it can create new threads and communicate and manage agents. And literally always, like
Muse Spark 1.3 is rolling out today with frontier performance almost too cheap to meter. This is the biggest jump we've made so far on coding and agentic work. Try it in Muse Code and our API.
Next up 🍉 and Muse Spark open weights releases coming soon.
Vercel added Muse Spark 1.3 (meta/muse-spark-1.3) to AI Gateway on launch day, with 1M-token context, text/image/PDF input, and both standard and contributor pricing tiers.
So, I finally tested Muse Spark 1.3. 2 real repos, 105 planted bugs, find and fix what you can. Original harness and API.
Big surprise:
Muse Spark 1.3 (max): 33
Fable 5.1 (high): 33
Grok 4.6 (xhigh): 27
Opus 5 (max): 27
Muse Spark 1.3 (high): 19
Meta joined the frontier.
Matt Johnston's live gauntlet puts Muse Spark 1.2 at 95 and #5 on his board, at $1.25/$4.25 per M tokens and 171 tok/s on OpenRouter; the full bench ran in 17 minutes.
Bijan Bowen runs Muse Spark 1.3 through a browser OS, C++ skate game, subway FPS, 3D CAD Hot Wheels, a Blender/Godot wrestling game and more, then tallies the total cost.
Dilmer Valecillos converted a minigolf prototype to VR for Meta Quest 3 using Muse Code with the Muse Spark 1.2 Contributor model, the Unity CLI and the MetaVR CLI, and published the plans and prompts.
Custom firmware and a Go server that turn an M5Stack StackChan into Tarquin, a voice-driven robot butler with wake word, face tracking and on-device speech, using Muse Spark 1.3 via the Meta Model API as its brain.
In addition to agents and coding, Muse Spark 1.1 is also really strong at answering health questions, a steadily growing use case for AI. On HealthBench-Pro, Muse Spark 1.1 achieves +5% better performance than Muse Spark 1.0 and beats all competitor models except Fable/Mythos.
Meta's Muse Spark 1.1 scores 51 on the Artificial Analysis Intelligence Index and is cost and token efficient compared to its peers
Muse Spark 1.1 (xhigh) improves 8 points over Muse Spark 1.0 (43) in three months. It is effectively tied with GLM-5.2 (max), GPT-5.4 (xhigh), and
We just measured Meta Muse Spark 1.3 on our internal writing benchmark hoping for new SOTA. Unfortunately, it isn't...
It enters at #24 of the 87 models we track, up from #31 for Muse Spark 1.1. As @alexandr_wang highlighted, it beats every Gemini configuration we have,
BREAKING: Muse Spark 1.2 by @AIatMeta takes 1st for Video-to-Website with an Elo rating of 1279 and impressive scores across all of our multimodal code categories.
Muse Spark 1.2 also takes 2nd for Image-to-HTML with an Elo rating of 1252 and 3rd for Image-to-Frontend with Elo
Muse Spark 1.3 (xHigh) just landed @AIatMeta back in the top 10 models for Code Arena: WebDev!
This release is ~#10 in Code Arena: WebDev with 1623 pts (AutoEval). It’s performance is par with Claude Fable 5 at #8 (1628 pts) and GPT-5.6 Sol at #11 (1616 pts).
Muse Spark 1.3
Exciting news: Muse Spark 1.2 (xHigh) by @AIatMeta is #4 in the Text Arena (1498 pts), and has reshaped the Pareto frontier!
It is priced at $1.25/$4.25 per MToken.
Congrats again to the @AIatMeta team on this release!
Muse Spark 1.2 from @AIatMeta is live on OpenRouter alongside expanded global access to both Muse Spark models.
At $1.25/M in and $4.25/M out, the model builds its position as one of the most price-efficient, high-intelligence models on OpenRouter.
A gallery of 100 self-contained generative-art, physics-toy and typography pages generated with Muse Spark 1.3, each shown next to its original prompt.
muse spark 1.1 in AdaL Engineer beats Opus4.8 in Claude Code with 20% of the cost
loop engineering, when done right, is beyond just running longer, its delivering better results when contexts are managed well and when workers are better prompted to stay honest.
how GANs had
Muse Spark 1.1 seems like a very good model. I tested it on Stata Benchmark and it ranks 4th (!) beating Opus 4.7/4.8, tied with GPT 5.5.
I like this benchmark because it often catches seemingly good model performing poorly on OOD tasks
Meta’s Muse Spark 1.3 Contributor model is now available for free in Cline.
Benchmarks show it’s similar in performance to Opus 5, but considerably more affordable.
Motion Labs shows how to connect a product catalog to Advantage+ Catalog Ads and the Meta Model API with Muse Spark 1.1, plus compliance checks for agencies.
Muse Spark 1.1 is used across Meta in coding and research workflows, scoring competitively with leading models on Meta's internal coding benchmark.
Our researchers are now automating model development and evaluation tasks by leveraging Muse Spark 1.1 in their workflows.
Tested Muse Spark 1.3 on a ~10k LOC codebase and found it sometimes beat Gemini 3.8 Flash on hard multi-part backend bugs, though Gemini produced better-looking frontends.
🎮 This entire Avo Lawn game was built with Muse Spark 1.2 + Muse Code. 🥑
Now it’s your turn—use Muse Spark 1.2 to build your own game and see what you can create.
To play: research.meta.ai/artifacts/intr…
Check out our blog post: research.meta.ai/blog/introduci… h
We’re excited to bring @AIatMeta’s Muse Spark 1.2 contributor tier to OpenRouter.
At $0.10/M input and $0.20/M output, It’s meaningfully cheaper than Muse Spark 1.2 and beats other comparable models on real cost, providing frontier intelligence-per-dollar.
A reproducible evaluation of a stochastic defect where Muse Spark 1.1, under some agent-like request envelopes, wrote to claude-smoke.txt when asked for muse-smoke.txt; a neutral control kept 80/80 filenames intact.
Releasing Muse Code in beta today. It's a terminal coding agent that takes on complete software engineering tasks across large repos: planning changes, writing code, validating the results. Powered by Muse Spark 1.2, a coding-focused model update.
We ran Muse Spark 1.3 on our Cybersecurity benchmark, and it's actually not that good (yet) 😬
- At pass@1, it rediscovers an average of 19/32 CVEs. In comparison, Grok 4.6 scores 23.3/32
- When pooling the results of 3 runs (pass@3), Muse scores 24/32. DeepSeek V4 Pro gets
Muse Spark 1.2 supports a broad range of multimodal tasks, from turning visuals into working code to translating perception into physical action. It also brings robust audio-visual understanding to enable video-heavy workflows common in real-world enterprise use.
Today, we’re
A same-prompt bakeoff site that starts with a playable penguin Tetris generated by Muse Spark 1.3 in OpenCode, with a fixed prompt for comparing other models.
Siamese Cat Dev pits Muse Spark 1.3 against Claude Fable 5.1 on a browser OS, a Call of Duty-style shooter, a neon arena shooter and a Mario Kart-style racer.
Muse Spark 1.2 just cracked the top 5 on the Vals Index, at just $0.69 per test. This is 3x cheaper than Kimi and 10x or more cheaper than Fable, Opus, and 5.6 Sol.
We're going to have to work much, MUCH harder to spend all those wonderful @AIatMeta tokens! Muse Spark 1.3 Contributor is unbelievably frugal. I've been chugging tokens at translations to get to 50 languages. Barely made a dent. Thanks again @alexandr_wang and team 🙏
Meta’s Muse Spark 1.3 just scored 62 on Artificial Analysis’ Intelligence Index.
That ties Claude Fable 5, but Muse costs 8x less for input and nearly 12x less for output.
It also scores above GPT-5.6 Sol, Grok 4.6, Kimi K3, and Gemini 3.8 Flash.
Meta is suddenly in the top