shipwithmuse

Entries matching “enterprise”

4 builds · page 1 of 1

I

int21.ai

int21.ai

INT21 ran Muse Spark 1.1 in its SwarmOS orchestration layer and argues the durable enterprise advantage is orchestration, not any single model.

Resource · Agents & automation

Running Muse Spark 1.1 inside SwarmOS

B

blog.box.com

blog.box.com

Box's Complex Work Eval finds Muse Spark 1.1 up to 5-6 points above the top-tier composite on structured work, and nearly 30 points ahead on cost-optimization analysis.

Resource · Benchmarks & research

Box eval: Muse Spark 1.1 on real enterprise work

Abdul

@AbdulTheBuilder

Big fan of @Muse Took grok bot & ran with it, haven't tried any serious work with it yet but it's working on getting me a refund from enterprise car rentals, booked me a flight for a Christmas trip & found me some pretty sick deals on recliners all within less than an hour of

X post · Errands & personal agent· ♥ 8

Car-rental refund, flight and recliner deals

Your product

Sponsored

Put your logo, a line of copy and an image right here, between the builds Muse developers come to read. Same size as a post.

$100/week

Put your product here

Shown every 12 builds · on every catalog page

U

PathfinderTactician

u/PathfinderTactician

I'm guessing that many people have been waiting for this comparison. For clarity, both models are running at full FP16 KV-cache. Due to VRAM limitations, Muse Glimmer is running full 262,144 context, whilst Qwen3.6 27B can only run at 147,500 context - full GPU offload in both cases. Both models have been coding on an enterprise-grade web application. Detailed report of each model (warning - includes AI generated content): Diagnostic quality - comparable. Both have shown genuinely good root-cause work when they apply themselves. Qwen found coding issue and worked to fix things cleanly. Muse Glimmer correctly traced bugs and even caught something that a Frontier model missed after more than 10 rounds of review. Neither one is weak at diagnosis. Implementation reliability - Qwen ahead. Qwen did introduce real regressions into the coding along the way (eg. severe zone-scope refactor regression, and case-sensitivity regression) but each one eventually got fixed properly once caught, usually within one or two corrective rounds. Muse Glimmer did land fixes that were clean and verified true to spec. However, when working in a complex environment exceeding 200k context, Muse Glimmer fa

Reddit post · Benchmarks & research

BF16 Muse Glimmer vs Qwen3.6 27B on real code