Using @AIatMeta's Muse Glimmer all locally to process personal monthly credit card statements.
Your data belongs to you!
Try different agent tasks using your favorite apps / harnesses with Ollama.
Meta has released Muse Spark 1.3, their fourth Muse Spark model release in five months. Muse Spark 1.3 (max), which is in limited preview for Meta’s partners, scores 62 on the Artificial Analysis Intelligence Index, behind only Claude Fable 5.1 and Claude Opus 5. The variant
Meta is back! Muse Spark scores 52 on the Artificial Analysis Intelligence Index, behind only Gemini 3.1 Pro, GPT-5.4, and Claude Opus 4.6. Muse Spark is the first new release since Llama 4 in April 2025 and also Meta's first release that is not open weights
Muse Spark is a new
Meta returns to open weights: Muse Glimmer, its first open-weights release since Llama 4, scores 35 on the Artificial Analysis Intelligence Index. It is a 30B-parameter model, and the first from Meta to be released under Apache 2.0
Muse Glimmer (high) arrives 16 months after
Muse Glimmer's gaps against its class concentrate in agentic evaluations: 953 Elo on GDPval-AA v2 against 1141 for Qwen3.6 27B (Reasoning), 1141 for Gemini 3.5 Flash-Lite, and 1004 for Kimi K2.5 (Reasoning), with Terminal-Bench v2.1 (52%) also behind Qwen3.6 27B (61%). The
Last week the Intelligence Index vs Cost Pareto frontier moved out substantially. Claude Fable 5.1, Muse Spark 1.3, and GPT-6 Astra each set a new point in efficient intelligence
Link to analysis: artificialanalysis.ai/#intelligence-…
I'm surprised more people aren't talking about this because Muse provides deep Instagram connectivity that other AI agents are unable to match, especially for free. I have been using Muse as an Instagram sidekick for a while. Here some things ive been playing around with and doing through Muse.
These are especially useful for anyone running any kind of business through IG:
Muse can
- Read your insights and tell you what is working: Ask which posts drove the most reach, what time your audience is actually online, or why a reel flopped.
- Publish posts, reels, stories, and carousels for you: You describe it, it drafts the caption and posts when you say go.
- Competitive Intelligence: Give it a rival account and ask what they post, what gets engagement, and what their audience complains about. Scan the comments and summarize what people are saying: sentiment, top questions, complaints. Ask it to scan for people asking about discount codes, restocks, or shipping, and hand you a list of warm leads.
Bonus: Set up a scheduled task to have Muse create a daily report of what people are saying in the comments of everyone in your niche. Who comments the most, what people keep asking for
An oldie post from 2024, but I've been curious about the data side. I had @Muse pull in my whole angel portfolio via MCP from @signedco of ~150 startups; the interesting bit:
- Startups who sent investor regular updates: 8% died
- Startups who didn't send updates: 28% died
Meta’s Muse Spark 1.3 just scored 62 on Artificial Analysis’ Intelligence Index.
That ties Claude Fable 5, but Muse costs 8x less for input and nearly 12x less for output.
It also scores above GPT-5.6 Sol, Grok 4.6, Kimi K3, and Gemini 3.8 Flash.
Meta is suddenly in the top
eesel AI reports Muse Spark 1.3 ranks #6 on the Artificial Analysis Intelligence Index, leads long-context and coding rows, but trails Claude Opus 5 on four of six agent evals.
Meta has released Muse Spark 1.2. It's their third release in four months and scores 54 on the Artificial Analysis Intelligence Index, significantly improving agentic knowledge work capabilities over prior releases and putting Meta next to SpaceXAI in a tie for third place
Meta's Muse Spark 1.1 scores 51 on the Artificial Analysis Intelligence Index and is cost and token efficient compared to its peers
Muse Spark 1.1 (xhigh) improves 8 points over Muse Spark 1.0 (43) in three months. It is effectively tied with GLM-5.2 (max), GPT-5.4 (xhigh), and
At 30B parameters, Muse Glimmer sits near the Intelligence vs Parameters frontier for open weights models: 5 points above Gemma 4 31B (Reasoning) at the same size, effectively matching Kimi K2.5 (Reasoning) at 33x fewer total parameters, and just behind Qwen3.6 27B (Reasoning),
i find muse spark is very good at data analysis—both finding relevant open-source data and analyzing it.
for example, here's my results for analyzing global share of GDP over past century:
meta.ai/share/cw54skLB…
Box's Complex Work Eval finds Muse Spark 1.1 up to 5-6 points above the top-tier composite on structured work, and nearly 30 points ahead on cost-optimization analysis.
Glimmer obtient 92 % du score d'intelligence de Qwen3.6 (35/38), mais Qwen a généré environ 2,9× plus de tokens sur l'ensemble de l'Intelligence Index. Et sur les endpoints mesurés par Artificial Analysis, Glimmer génère environ 1,8× plus vite. Et le context de glimmer et bien plus efficace !
C est une belle avancer architecture tout de meme , je pense que si il sorte une version 1.1 (surtout pour améliorer terminal benchmark ) ont pourrai être très surpris !
MindStudio notes Muse Spark 1.3 topped DeepSWE at 75.4 and placed third on Artificial Analysis, yet in a hands-on game-clone test produced "a cube shooting at other cubes."