Muse Glimmer 30B
Muse Glimmer 30B (August 10, 2026) is Meta Superintelligence Labs’ first open model under the Apache 2.0 license, running without an EU exclusion clause for local deployment. The dense 29.6-billion-parameter model with an additional 1.8-billion-parameter vision encoder processes text and images in a 131,072-token context and delivers up to 233 tokens/second on a consumer GPU via DFlash Speculative Decoder — Workstation-class with genuine Desktop capability.
- Open Weights
- Desktop
- vLLM
- Text
- Vision
- Long Context
- Unusable
Sovereign Risk: LOW Meta is a US company and subject to the CLOUD Act. However, Muse Glimmer 30B is released as fully open weights under the Apache 2.0 license — Meta’s first model ever under this license. When running entirely locally on your own hardware, any dependency on US cloud infrastructure is eliminated, which is why the risk is rated as low despite US jurisdiction.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 71.56
- Routine
- 43.59
- Reasoning
- 27.97
- LLM Judge Avg
- 3.51 / 5
- 100 Coverage
- Avg Task Duration
- 190.4s
- Unusable
- Token Rate
- 10.59tok/s
- Output Rate
- P95 Latency
- 392.46s
- Top 5 %
- Total Tokens
- 124900
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 124900 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median