shipwithmuse

Entries matching “runpod”

1 build · page 1 of 1

@cezaronx

@cezaronx

A reference CUDA worker that serves Meta's official Muse Glimmer 30B GGUF through llama-server on Runpod Serverless load-balancing endpoints or manual Pods, exposing a real OpenAI-compatible API.

GitHub · Local & open models

Muse Glimmer Runpod serverless worker