shipwithmuse

Entries matching “claude”

51 builds · page 1 of 1

Greg Ceccarelli

@gregce10

For funsies, I got Claude Code running on my Muse agent’s computer tonight. Muse gives every agent a Linux VM with 2 vCPUs, ~8GB RAM, 100GB of persistent disk plus a scheduler, hosted artifacts and their permission system. The VM has no inbound ports, so I built a tiny

X post · Coding & dev tools★ Pick· ♥ 112

Claude Code inside a Muse agent VM

@Srimi1

@Srimi1

A hybrid harness with a claude-muse launcher that runs Claude Code on Meta Model API Muse Spark models, plus a /talk broker for structured exchanges between Claude Code and Muse Code terminals.

GitHub · Coding & dev tools

Muse Spark in Claude Code with /talk

@Sighton-GH

@Sighton-GH

opencode-delegate is a Claude Code plugin where Claude plans and writes briefs while opencode sessions on free models, Muse Spark 1.3 by default, implement and review each task.

Skill · Coding & dev tools

opencode-delegate Claude Code plugin

rob

@goof

My muse has codex and Claude running on his machine so he can call up for some reinforcements when he gets stuck, unless you got muse 1.3 pro on deck

X post · Agents & automation· ♥ 2

Muse calls Codex and Claude as backup

@kevintsai1202

@kevintsai1202

MCP server that gives Claude Code, Claude Desktop, Cursor and other MCP clients image generation and editing powered by Meta's Muse image model, published to npm.

Skill · Connectors & MCP

muse-image-mcp

@rtravellin

@rtravellin

A Claude Code plugin that adds /muse:review, /muse:critique, /muse:delegate and /muse:transfer commands, running Muse Code reviews and delegated tasks as background jobs.

Skill · Coding & dev tools

Muse Code plugin for Claude Code

Newton Jones

@NewtonJonesIV

I was able to migrate multi agent workflows to Muse from Claude. I gave it the master prompt- Muse executed flawlessly- no feedback- just my prompt- built for Claude 4.6. It is now running this prompt every morning at a fraction of the cost. Leaned up all of my emails-

X post · Agents & automation· ♥ 1

Claude multi-agent workflow migrated to Muse

Your product

Sponsored

Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.

$100/week

Put your product here

Shown every 12 builds · on every catalog page

@vasyl-stanislavchuk

@vasyl-stanislavchuk

A separate launcher with its own config profile that runs Claude Code's interface on Meta's muse-spark-1.3, for comparing models on identical work or using a Muse Code subscription.

GitHub · Coding & dev tools

claude-muse launcher

@CogniTechSystems

@CogniTechSystems

This repo documents running Muse Glimmer 30B locally on an M4 Max MacBook via llama.cpp, benchmarking it with and without speculative decoding, and wiring it into Claude Code through LiteLLM for fully offline coding.

GitHub · Local & open models

Claude Code on local Muse Glimmer 30B

@troioi-vn

@troioi-vn

delegate-to-muse is a Claude Code skill for handing repository work to Muse Code via muse exec, with a handoff template and checks for reviewing the diff and re-running tests rather than trusting the summary.

Skill · Coding & dev tools

delegate-to-muse Claude Code skill

@NatesVibeCode

@NatesVibeCode

Portable SKILL.md bundles for handing work to Muse, Antigravity, OpenCode, Grok Build, Cursor, Codex and Claude Code, each with supporting references.

Skill · Coding & dev tools

Harness handoff skills including Muse

@corona10

@corona10

Unofficial native VS Code chat UI for Muse Code, modeled on the Claude Code VS Code extension's layout, commands and keybindings.

Skill · Coding & dev tools· ★ 2

Muse Code for VS Code (unofficial)

@gregmos

@gregmos

Routing overlay for the Superpowers plugin for Claude Code where Muse writes the code, Codex reviews it, and rules keep expensive models from being misused.

Skill · Coding & dev tools

superpowers-muse routing overlay

@luckeyfaraday

@luckeyfaraday

An unofficial loopback shim that lets Meta Muse Code use OpenAI Responses APIs, OpenRouter, Codex OAuth and the official Claude Code CLI, switchable from Muse's /model menu.

GitHub · Coding & dev tools· ★ 1

muse-shim

@Solar2004

@Solar2004

A config (in Spanish) plus two helper proxies for using Muse Spark 1.2 Contributor from Claude Code and Hermes through routatic/proxy and an OpenCode Go subscription.

GitHub · Coding & dev tools

muse-stack

Artificial Analysis

@ArtificialAnlys

Meta is back! Muse Spark scores 52 on the Artificial Analysis Intelligence Index, behind only Gemini 3.1 Pro, GPT-5.4, and Claude Opus 4.6. Muse Spark is the first new release since Llama 4 in April 2025 and also Meta's first release that is not open weights Muse Spark is a new

X post · Benchmarks & research· ♥ 2.4K

Original Muse Spark scores 52

C

corona10

corona10

A community-built VS Code chat UI for Muse Code, Meta's agentic coding CLI, inspired by the Claude Code extension; about 145 installs on the Marketplace.

Site · Coding & dev tools

Muse Code (Unofficial) VS Code extension

C

trycodus.com

trycodus.com

Codus reads the three benchmark charts Meta published for Muse Code and notes Claude Opus 5 wins all three, including Meta's own internal eval.

Resource · Benchmarks & research

What Meta's Muse Code benchmarks actually say

Idobn

@idobn

Had @Muse go through 6 months of my email receipts. It found a $100/mo Claude Max sub on my corp card I'd forgotten about, while we already have an enterprise license. $1,200/year saved. @finkd clearly articulates the biz model here - muse is supposed to save / make you money,

X post · Errands & personal agent· ♥ 14

Receipt audit finds $1,200/yr waste

Your product

Sponsored

Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.

$100/week

Put your product here

Shown every 12 builds · on every catalog page

T

@t3dotgg

@t3dotgg

Theo (t3.gg) tests Meta's Muse Code terminal agent powered by Muse Spark 1.2 and focuses on how cheap it is compared with Claude Code.

Video · Coding & dev tools· ♥ 3.5K

Theo reviews Muse Code and Muse Spark 1.2

U

Graemer71

u/Graemer71

OK, for context, I have Claude Code desktop app driving the CLI and orchestrating the code and verification tasks to try to save tokens. So Claude runs things, a Deepseek 4.1 Flash (cloud) session does the planning, Qwen 3.8 27b Q8 does the boiler plate coding and Muse Glimmer sanity checks the code and pushes any issues back to Qwen. If there are issues Qwen and Glimmer can't agree on, Deepseek validates. If Deepseek can't sort it out, it goes back to Claude. This had been working fine, but then in the last few days token use spiked, tasks that used to take 10 minutes were taking an hour or more and Qwen started going into more and more reasoning loops. It seems that since I last checked (on 12th September) the CLI changed. I used to strip unnecessary tool calls from the prompt using --disallowedTools and enabledPlugins: false. It would seem that these no longer work. In the end I got Claude to build a request-dumping diagnostic server, that actually measured the payload bytes, and confirmed --tools (an allowlist) is the flag that works now: 55→7 tools, 161KB→24KB, byte-verified. It also caught something specific to my workflow running the wrapper from inside an already-active

Dan

@DanDr1s

Meta’s Muse Spark 1.3 just scored 62 on Artificial Analysis’ Intelligence Index. That ties Claude Fable 5, but Muse costs 8x less for input and nearly 12x less for output. It also scores above GPT-5.6 Sol, Grok 4.6, Kimi K3, and Gemini 3.8 Flash. Meta is suddenly in the top

X post · Benchmarks & research· ♥ 223

Spark 1.3 matches Fable 5 at a fraction of the price

@alexanderradahl

@alexanderradahl

A fork of Spotify's shunt plugin for Claude Code and Codex whose hooks block full reads of large files and route them to Muse Spark via muse exec; the author measured about 4,800 tokens out and 300 back on a 617-line file.

I

company.instacart.com

company.instacart.com

Instacart is bringing personalized grocery shopping to Muse through its connector program, drawing on real-time inventory from about 100,000 stores.

Site · Connectors & MCP

Instacart grocery shopping via Muse

U

Ok-Inevitable8391

u/Ok-Inevitable8391

Benchmarked qwen3.8 xhigh, medium and muse glimmer. Xhigh effort mode with qwen3.8 took almost 30hrs. (And still failed on 16 cases because of the 32K output token limit) Medium effort mode and muse glimmer were 3-4 hours each. But I'm actually surprised by the muse glimmer results, they came better than the qwen. These benchmarks are on implicit knowledge of the model, which is a bit unfair to smaller models, but throw in a RAG and I'm sure they get on par with frontier models. I have taken the result of claude models directly from embedeval repo by ecro. I'm not pushing qwen down here, I like how qwen thinks and gives better results. I know with more context and RAG qwen will do better. I'm just appreciating muse here, cause i feel it is underrated. The advantage is efficient kv cache due to sliding window, which can give you more context window.

Reddit post · Benchmarks & research

Glimmer vs Qwen 3.8 on an implicit-knowledge eval

@HarjjotSinghh

@HarjjotSinghh

Helicon is an open-source desktop and web app that wraps Meta's Muse Code CLI in a Codex/Claude-style interface, grouping sessions by project with history, resume and diffs on Windows, macOS and WSL2.

GitHub · Coding & dev tools★ Pick

Helicon desktop app for Muse Code

Your product

Sponsored

Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.

$100/week

Put your product here

Shown every 12 builds · on every catalog page

@jellologic

@jellologic

Claude Code plugin that offloads bulk coding work to Muse Code workers in isolated git worktrees, with a Claude supervisor that runs your acceptance check and sends failing patches back.

Skill · Coding & dev tools

claude-code-muse delegation plugin

C

composio.dev

composio.dev

Composio looks at the three benchmark charts from Meta's Muse Code launch and whether developers should switch from Claude Code.

Resource · Coding & dev tools

Composio: Muse Code vs Claude Code

Artificial Analysis

@ArtificialAnlys

Meta has released Muse Spark 1.3, their fourth Muse Spark model release in five months. Muse Spark 1.3 (max), which is in limited preview for Meta’s partners, scores 62 on the Artificial Analysis Intelligence Index, behind only Claude Fable 5.1 and Claude Opus 5. The variant

X post · Benchmarks & research· ♥ 2.6K

Artificial Analysis scores Muse Spark 1.3

@AloreBB

@AloreBB

Port of openai/codex-plugin-cc that lets Claude Code call Muse for reviews or delegated tasks via muse exec, with /muse:review, /muse:adversarial-review and background job commands.

Skill · Coding & dev tools

Muse plugin for Claude Code

R

runtimewire.com

runtimewire.com

RuntimeWire's head-to-head eval has Muse Spark 1.1 beating Claude Opus 4.8 by 10 points overall with a 95% confidence verdict.

Resource · Benchmarks & research

Head to head: Muse Spark 1.1 vs Claude Opus 4.8

@raksix

@raksix

A documented setup that runs Muse Spark 1.3 Contributor inside Claude Desktop on macOS through a small Node proxy that translates Anthropic requests to OpenCode Go's OpenAI-style API.

GitHub · Coding & dev tools· ★ 3

Muse Spark 1.3 inside Claude Desktop

@phareim

@phareim

musecode adds a /musecode slash command, a Haiku relay agent and a muse-run CLI so Claude Code Workflow fan-outs run their leaf tasks on the free Muse Spark 1.3 model through opencode.

S

@SiameseCatDev

@SiameseCatDev

Siamese Cat Dev pits Muse Spark 1.3 against Claude Fable 5.1 on a browser OS, a Call of Duty-style shooter, a neon arena shooter and a Mario Kart-style racer.

A

armand0e

armand0e

Muse Glimmer 30B fine-tuned on agentic coding traces and chat distilled from Claude Fable 5, merged bf16 and drop-in for the base.

Resource · Local & open models· ♥ 3

Muse Glimmer Fable Distill

O

ollama.com

ollama.com

Ollama shipped Muse Glimmer on day one: `ollama run muse-glimmer`, plus a muse-glimmer:30b-mlx tag for Apple Silicon that Ollama says runs 1.5–1.8x faster with DFlash.

Site · Local & open models

Muse Glimmer in the Ollama library

Your product

Sponsored

Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.

$100/week

Put your product here

Shown every 12 builds · on every catalog page

@ArshansGithub

@ArshansGithub

An OpenAI Responses-compatible shim that serves the Muse assistant as a stateless HTTP model backend for Claude Code via CLIProxyAPI, one model turn per request.

GitHub · Coding & dev tools· ★ 1

Muse agent-api shim

Arena.ai

@arena

Muse Spark 1.3 (xHigh) just landed @AIatMeta back in the top 10 models for Code Arena: WebDev! This release is ~#10 in Code Arena: WebDev with 1623 pts (AutoEval). It’s performance is par with Claude Fable 5 at #8 (1628 pts) and GPT-5.6 Sol at #11 (1616 pts). Muse Spark 1.3

X post · Benchmarks & research· ♥ 610

Spark 1.3 in Code Arena: WebDev top 10

@Israelmusondaayliffe

@Israelmusondaayliffe

Muse workspace skills ported from a Codex/Claude Code plugin marketplace, translating manifests and hooks into procedures Muse can run with its terminal, browser and files.

ollama

@ollama

Muse Glimmer is now available to run with Ollama. Available today via Ollama’s MLX engine with state-of-the-art-performance on Apple Silicon, Muse Glimmer can power Claude Code, Codex, and more always-on local agent workflows natively using Ollama. Additional support and

X post · Local & open models· ♥ 1.3K

Muse Glimmer on Ollama's MLX engine

@siddicky

@siddicky

An oh-my-claudecode-style gated pipeline for the Muse Code CLI, from deep-interview to verified code, shipped as a native Muse plugin manifest with a skills fallback.

Skill · Coding & dev tools

oh-my-musecode delivery harness

@Wraient

@Wraient

A Go reverse proxy with an embedded dashboard that lets Claude Code and OpenAI/Codex clients use muse-spark models through a Muse Code subscription, authenticated with the muse login credential.

GitHub · Coding & dev tools

muse-cc subscription proxy

li yin

@panda_liyin

muse spark 1.1 in AdaL Engineer beats Opus4.8 in Claude Code with 20% of the cost loop engineering, when done right, is beyond just running longer, its delivering better results when contexts are managed well and when workers are better prompted to stay honest. how GANs had

X post · Benchmarks & research· ♥ 75

Spark 1.1 in AdaL Engineer vs Opus 4.8

Artificial Analysis

@ArtificialAnlys

Last week the Intelligence Index vs Cost Pareto frontier moved out substantially. Claude Fable 5.1, Muse Spark 1.3, and GPT-6 Astra each set a new point in efficient intelligence Link to analysis: artificialanalysis.ai/#intelligence-…

X post · Benchmarks & research· ♥ 732

Spark 1.3 moves the cost Pareto frontier

@JacobStephens2

@JacobStephens2

A reproducible evaluation of a stochastic defect where Muse Spark 1.1, under some agent-like request envelopes, wrote to claude-smoke.txt when asked for muse-smoke.txt; a neutral control kept 80/80 filenames intact.

GitHub · Benchmarks & research

Muse Spark 1.1 filename-substitution repro

L

lennysnewsletter.com

lennysnewsletter.com

A How I AI review: the reviewer gave Muse a calendar, email and kids' chaotic schedule and got a one-shot family PDF described as more beautiful than anything made with Claude.

Resource · Errands & personal agent

Lenny's Newsletter: Muse gets consumer agent UX right

U

TopPromise9775

u/TopPromise9775

I ran two timestamped prospective tests through ForecastNest (a project I’m building) on September 22. The same six models researched the market independently at 10:25–10:42 AM EDT. Both forecasts were frozen and settled one trading session later at the same wall-clock time. Prompt A — Stock Selection “Select exactly five distinct US-listed stocks or ETFs most likely to outperform SPY over one trading session.” The five picks were equally weighted, and the score was portfolio return minus SPY. GPT-5.6 Sol produced +0.73 percentage points of alpha and DeepSeek V3 +0.03. The other four portfolios failed to beat SPY. Prompt B — Extreme Movers Research the previous session’s top gainers and losers, choose exactly five stocks, and predict up or down as either continuation or reversal. The score was the mean signed return: a correct down call benefits from a falling price. GPT-5.6 Sol scored +6.17%, Muse Spark 1.1 +1.34%, and Claude Opus 5 +0.68%. Gemini 1.5 Pro, Grok 4.5, and DeepSeek V3 finished negative. Same date, same models, same horizon—but changing the task changed the apparent model performance. GPT led both tests that day. That is interesting, but one session is not evid

Reddit post · Business & commerce

ForecastNest: six LLMs make one-day market calls

E

eesel.ai

eesel.ai

eesel AI reports Muse Spark 1.3 ranks #6 on the Artificial Analysis Intelligence Index, leads long-context and coding rows, but trails Claude Opus 5 on four of six agent evals.

Your product

Sponsored

Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.

$100/week

Put your product here

Shown every 12 builds · on every catalog page