Meta has released Muse Glimmer, a 30B open-weight dense model with a 120K+ context window designed for local AI agentic work. Optimized to run across NVIDIA edge, desktop, and workstation platforms, the model delivers 20K tokens/sec on a single GPU.
- Model size: 30B parameters, dense architecture.
- Context window: 120K+ tokens.
- Performance: 20K tokens/sec on a single GPU.
- Target hardware: NVIDIA edge, desktop, and workstation AI platforms.
This release enables always-on agents to process data locally and execute complex workflows without relying on cloud infrastructure.