№ 0027Resource
Introducing Muse Glimmer
Meta's launch post for Muse Glimmer, an Apache 2.0 30B model for local agents that fits in ~20GB at 4-bit and runs on M4/M5 Max Macs, RTX 5090s or 24–32GB GPUs.
Meta compares it with Gemma4-31B and Qwen3.6-27B on agentic, coding, multimodal, safety and reasoning tasks, and says speculative decoding speeds generation 1.5–3.1x depending on hardware. Launch partners include Ollama, LM Studio, Unsloth, llama.cpp, MLX, vLLM, SGLang, Together, Fireworks, OpenRouter, NVIDIA, AMD, Arm, Dell and Intel.







ChatForm
Tgmlabs