Artificial Analysis' independent benchmark write-up of Muse Spark 1.3.

Resource · Benchmarks & research
17 builds · page 1 of 1
Artificial Analysis' independent benchmark write-up of Muse Spark 1.3.

Resource · Benchmarks & research
Your product
Sponsored
Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.
Shown every 12 builds · on every catalog page
Independent benchmarks and analysis of Muse Spark 1.2, released alongside Muse Code.

Resource · Benchmarks & research
Artificial Analysis
@ArtificialAnlys
Meta has released Muse Spark 1.2. It's their third release in four months and scores 54 on the Artificial Analysis Intelligence Index, significantly improving agentic knowledge work capabilities over prior releases and putting Meta next to SpaceXAI in a tie for third place

X post · Benchmarks & research· ♥ 1.2K
Artificial Analysis
@ArtificialAnlys
Meta's Muse Spark 1.1 scores 51 on the Artificial Analysis Intelligence Index and is cost and token efficient compared to its peers Muse Spark 1.1 (xhigh) improves 8 points over Muse Spark 1.0 (43) in three months. It is effectively tied with GLM-5.2 (max), GPT-5.4 (xhigh), and

X post · Benchmarks & research· ♥ 708
Artificial Analysis
@ArtificialAnlys
At 30B parameters, Muse Glimmer sits near the Intelligence vs Parameters frontier for open weights models: 5 points above Gemma 4 31B (Reasoning) at the same size, effectively matching Kimi K2.5 (Reasoning) at 33x fewer total parameters, and just behind Qwen3.6 27B (Reasoning),

X post · Benchmarks & research· ♥ 60
Motion Labs explains Muse Spark 1.2 setup, the cheaper Contributor tier at $0.10/M input tokens, Artificial Analysis scores and using it for content.

Resource · Content & creative
Your product
Sponsored
Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.
Shown every 12 builds · on every catalog page
Glimmer obtient 92 % du score d'intelligence de Qwen3.6 (35/38), mais Qwen a généré environ 2,9× plus de tokens sur l'ensemble de l'Intelligence Index. Et sur les endpoints mesurés par Artificial Analysis, Glimmer génère environ 1,8× plus vite. Et le context de glimmer et bien plus efficace ! C est une belle avancer architecture tout de meme , je pense que si il sorte une version 1.1 (surtout pour améliorer terminal benchmark ) ont pourrai être très surpris !
Reddit post · Benchmarks & research
eesel AI reports Muse Spark 1.3 ranks #6 on the Artificial Analysis Intelligence Index, leads long-context and coding rows, but trails Claude Opus 5 on four of six agent evals.
Resource · Benchmarks & research
Artificial Analysis
@ArtificialAnlys
Meta has released Muse Spark 1.3, their fourth Muse Spark model release in five months. Muse Spark 1.3 (max), which is in limited preview for Meta’s partners, scores 62 on the Artificial Analysis Intelligence Index, behind only Claude Fable 5.1 and Claude Opus 5. The variant

X post · Benchmarks & research· ♥ 2.6K
Independent benchmarks and analysis of Meta's open-weight Muse Glimmer.

Resource · Benchmarks & research
Artificial Analysis puts Muse Spark 1.1 at 51 on its Intelligence Index, 8 points above 1.0, and calls it cost and token efficient versus peers.

Resource · Benchmarks & research
Artificial Analysis
@ArtificialAnlys
Meta is back! Muse Spark scores 52 on the Artificial Analysis Intelligence Index, behind only Gemini 3.1 Pro, GPT-5.4, and Claude Opus 4.6. Muse Spark is the first new release since Llama 4 in April 2025 and also Meta's first release that is not open weights Muse Spark is a new

X post · Benchmarks & research· ♥ 2.4K
Artificial Analysis
@ArtificialAnlys
Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence Index. It is a 30B-parameter model, and the first from Meta to be released under Apache 2.0 Muse Glimmer (high) arrives 16 months after

X post · Benchmarks & research· ♥ 778
Artificial Analysis
@ArtificialAnlys
Muse Glimmer's gaps against its class concentrate in agentic evaluations: 953 Elo on GDPval-AA v2 against 1141 for Qwen3.6 27B (Reasoning), 1141 for Gemini 3.5 Flash-Lite, and 1004 for Kimi K2.5 (Reasoning), with Terminal-Bench v2.1 (52%) also behind Qwen3.6 27B (61%). The

X post · Benchmarks & research· ♥ 34
Artificial Analysis
@ArtificialAnlys
Last week the Intelligence Index vs Cost Pareto frontier moved out substantially. Claude Fable 5.1, Muse Spark 1.3, and GPT-6 Astra each set a new point in efficient intelligence Link to analysis: artificialanalysis.ai/#intelligence-…
X post · Benchmarks & research· ♥ 732
Dan
@DanDr1s
Meta’s Muse Spark 1.3 just scored 62 on Artificial Analysis’ Intelligence Index. That ties Claude Fable 5, but Muse costs 8x less for input and nearly 12x less for output. It also scores above GPT-5.6 Sol, Grok 4.6, Kimi K3, and Gemini 3.8 Flash. Meta is suddenly in the top

X post · Benchmarks & research· ♥ 223
MindStudio notes Muse Spark 1.3 topped DeepSWE at 75.4 and placed third on Artificial Analysis, yet in a hands-on game-clone test produced "a cube shooting at other cubes."

Resource · Benchmarks & research