Meta released Muse glimmer, a 30B agentic model with open weights under apache 2.0. muse glimmer can run on 24GB of VRAM without losing agentic reliability
Just like much larger models, muse glimmer can operate as a fully capable agent via planning, tool calls, checking its own results, and failure recovery.
Muse glimmer was developed with its own architecture and recipe, optimized for its size and agentic performance requirements.
Weights on hugging face now. running this week through ollama, LM Studio, vllm, sglang, together, fireworks, and openrouter, with llama.cpp, MLX, and executorch.
Post #4403
641