Meta releases Muse Glimmer 30B, available on Fireworks for always-on agents
Original titleMuse Glimmer from Meta on Fireworks: Ideal for your Always-On Agents
AISummary
Meta's Muse Glimmer is a 30B dense model with a 128K+ token context window, now available on Fireworks in serverless and on-demand deployments. Meta reports it leads its size class on MCP Atlas (75.5) and DeepSearch QA (74.6) against Gemma 4 31B and Qwen 3.6 27B, with its sliding-window attention and two KV heads keeping the cache small for concurrent agent sessions.
AIWhy it matters
The post pairs an architecture explained through KV cache size with benchmark tables against two rival models, which helps readers judge whether it fits their agent workload.
Source: Fireworks AI Blog · fireworks.ai