loot radar is my game deals site, radar.codemeoww.com. tracks deals and freebies across stores like steam and epic, ranked by discount and popularity.
features: daily loot summaries on telegram and discord, fast price alerts, a morning deals digest, wishlist tracking so you get pinged when a game you want drops, and an api for the deal data.
Muse Spark 1.3 Max by @AIatMeta has reshaped the Pareto frontier for Code Arena: WebDev!
Meta's latest model at Max reasoning is doing something interesting on the Arena Pareto frontier: it's the only model holding down the wide price band between Qwen3.8-max ($5/MToken) and
In the last week @muse has
- figured out that FedEx misfiled my iPhone shipment, then spoke with apple to get my money back
- bought me prime tickets to Primetime
- made a dinner reservation, then got me refund when I had to cancel
- built a leads list based on my Instagram and
Muse Spark 1.1 seems like a very good model. I tested it on Stata Benchmark and it ranks 4th (!) beating Opus 4.7/4.8, tied with GPT 5.5.
I like this benchmark because it often catches seemingly good model performing poorly on OOD tasks
Muse Glimmer's gaps against its class concentrate in agentic evaluations: 953 Elo on GDPval-AA v2 against 1141 for Qwen3.6 27B (Reasoning), 1141 for Gemini 3.5 Flash-Lite, and 1004 for Kimi K2.5 (Reasoning), with Terminal-Bench v2.1 (52%) also behind Qwen3.6 27B (61%). The
Taylor Arndt tested Muse with VoiceOver and found non-standard text fields and missing heading structure in chats, plus connector gaps such as iCloud email that kept it out of her work.
Exciting news: Meta’s Muse Image just claimed #2 in the Image Arena!
Muse Image from @AIatMeta now ranks second only to OpenAI's GPT Image 2, outperforming Nano Banana, Grok Imagine, MAI Image, and many other leading image models.
It holds #2 across the board: Text-to-Image,
Browserbase is now available as a plugin in Muse Code.
Get started with two commands:
> muse plugins marketplace add browserbase github.com/browserbase/br…
> muse plugins install browse@browserbase
Box's Complex Work Eval finds Muse Spark 1.1 up to 5-6 points above the top-tier composite on structured work, and nearly 30 points ahead on cost-optimization analysis.
10 minutes later, I have my own workout tracker 💪 📈
Built with @Muse. I tell it my exercises, weights, sets, and reps. It logs everything and graphs my progress for each exercise.
Even added a light/dark mode toggle. h
dragging our friends to hot yoga every sunday just for me to have slightly better yoga form than @andrewztan
the model is natively multimodal, which allows visual grounding + coding to make interactive UIs like this!
ai.meta.com/blog/introduci…
Postiz's Nevo David shows how to hook Postiz's OAuth MCP server into Muse through the custom-connector path, so Muse can schedule posts across 30+ platforms before an official directory listing.
Today we’re also previewing WildArtifactBench, an internal evaluation framework designed to assess agents on complex, real-world tasks across diverse deliverable formats.
By using win rates and Elo scores from human and agentic preference judges rather than strict ground-truth
Last night, I built a VR game prototype in a few hours with Muse Code on PC. It was already available on macOS and Linux, and today we’re bringing it to PC! 🔥
I tested it with Unity through the Unity CLI, and here are a few things that worked really well for me:
- Having Muse
January 1st, 2026. I wrote down “Make Faceless YouTube & TikTok channel”
9 months later, I still never got to it! 🤦🏻♂️
Today with @Muse I created 6 accounts: Gmail, YouTube, TikTok, Snapchat, X, IG & now FB in under 2 hours!
This feels like AGI.
The main benefit of Muse is h
Building an app for tales across the world to be heard in the voice of their loved ones
muse.ai/s/granny-s-tal…
Which feeds the insta channel automatically
instagram.com/grannys.tales?…
Box added Muse Spark 1.3 to Box AI, reporting it runs 42% faster than Muse Spark 1.2 with roughly a third fewer tokens, and lifts financial services accuracy from 66% to 75% on Box's eval.
Port of openai/codex-plugin-cc that lets Claude Code call Muse for reviews or delegated tasks via muse exec, with /muse:review, /muse:adversarial-review and background job commands.
Muse Spark 1.1 also excels in perception and multimodal reasoning, inspecting visual and audio inputs, preserving details across long workflows, and acting on them in real execution environments. It shows particular strengths in visual-to-code generation, rich image/video
An independent desktop command center for the Muse Code CLI that brings workspaces, sessions, approvals, reasoning controls, changed files and agent activity into one interface.
I went through hundreds of posts to find what people are actually getting done with @muse.
128 real use cases, organized by category; each with the original source and a prompt you can copy.
muse.ai/s/muse-use-cas…
MarkTechPost walked through Muse for Mac, which pulls context from Calendar, Mail, Messages, Notes and files for multi-step tasks, with opt-in access, optional Full Disk Access and approval before deleting files or sending messages.
A reproducible evaluation of a stochastic defect where Muse Spark 1.1, under some agent-like request envelopes, wrote to claude-smoke.txt when asked for muse-smoke.txt; a neutral control kept 80/80 filenames intact.
I'm guessing that many people have been waiting for this comparison. For clarity, both models are running at full FP16 KV-cache. Due to VRAM limitations, Muse Glimmer is running full 262,144 context, whilst Qwen3.6 27B can only run at 147,500 context - full GPU offload in both cases.
Both models have been coding on an enterprise-grade web application.
Detailed report of each model (warning - includes AI generated content):
Diagnostic quality - comparable. Both have shown genuinely good root-cause work when they apply themselves. Qwen found coding issue and worked to fix things cleanly. Muse Glimmer correctly traced bugs and even caught something that a Frontier model missed after more than 10 rounds of review. Neither one is weak at diagnosis.
Implementation reliability - Qwen ahead. Qwen did introduce real regressions into the coding along the way (eg. severe zone-scope refactor regression, and case-sensitivity regression) but each one eventually got fixed properly once caught, usually within one or two corrective rounds. Muse Glimmer did land fixes that were clean and verified true to spec. However, when working in a complex environment exceeding 200k context, Muse Glimmer fa
Matt Naughton's merchant guide to Muse as a Shopify AI channel: where the Agentic toggles live, what direct checkout does, and why browser pixels don't fire on Meta-side checkout, so only server-to-server events count.
DesignRush argued Muse inserts itself into product discovery and ranks on structured data like specs and reviews rather than ad spend, and told brands to map the agent shopping journey, measure agent referrals separately, and audit who can access their product data.
Muse Spark 1.1 can use browsers and it refilled our office fridge
- Took 2 pics of our fridge stocked / not stocked
- Told Muse to analyze the pics and restock what we were low on
- It analyzed each shelf and which brand / items that were missing
- Used the @juliusai built in
@muse is officially our new Head of Accounts Receivable.
Emailed every partner with a pending invoice and somehow got weeks of stuck payments moving 😂
Thanks @alexandr_wang @finkd
Practical Ecommerce's Armando Roggio compared Shopify welcoming Muse as an agentic commerce channel with Amazon blocking it, and what that split means for merchants and the customer relationship.
A community catalog of Meta Muse integrations and 150 connector skills with capability, authentication and permission notes, plus workflow templates for using Muse with connectors.
I asked Muse to purchase a widely available thermostat. The only site that didn’t block Muse was the manufacturer website, which was 20% more. Walmart, Amazon, and Best Buy all blocked it.
Asked @Muse to file pothole repair requests with the county. Took 30 seconds, zero hassle. Honestly assumed they'd disappear into a void.
Five days later: all fixed. 🤯
Meta’s Muse Spark 1.3 Contributor model is now available for free in Cline.
Benchmarks show it’s similar in performance to Opus 5, but considerably more affordable.