Muse Glimmer 30B

Muse Glimmer 30B (August 10, 2026) is Meta Superintelligence Labs’ first open model under the Apache 2.0 license, running without an EU exclusion clause for local deployment. The dense 29.6-billion-parameter model with an additional 1.8-billion-parameter vision encoder processes text and images in a 131,072-token context and delivers up to 233 tokens/second on a consumer GPU via DFlash Speculative Decoder — Workstation-class with genuine Desktop capability.

Meta Version Glimmer Commercial use permitted Dense 29.6 B (29.6 B active) 131 K Context 01/2026 locally tested

  • Open Weights
  • Desktop
  • vLLM
  • Text
  • Vision
  • Long Context
  • Unusable

Sovereign Risk: LOW Meta is a US company and subject to the CLOUD Act. However, Muse Glimmer 30B is released as fully open weights under the Apache 2.0 license — Meta’s first model ever under this license. When running entirely locally on your own hardware, any dependency on US cloud infrastructure is eliminated, which is why the risk is rated as low despite US jurisdiction.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
71.56
Routine
43.59
Reasoning
27.97

Rank #58

LLM Judge Avg
3.51
100 Coverage
Avg Task Duration
190.4
Unusable
Token Rate
10.59
Output Rate
P95 Latency
392.46
Top 5 %
Total Tokens
124900
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 124900 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Muse Glimmer 30B Best model Ø All models
Code Quality 77
CLI Benchmark 90
Logical Reasoning 60.09
UX Writing 66.85
Documentation 66.87
Content Transform. 68.12
Cultural Intelligence 80.3
Synthesis Quality 60
Tool Execution 85.83
ToolUse Score 72.5
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile