Muse Glimmer 30B

Muse Glimmer 30B (August 10, 2026) is Meta Superintelligence Labs’ first open model under the Apache 2.0 license, running without an EU exclusion clause for local deployment. The dense 29.6-billion-parameter model with an additional 1.8-billion-parameter vision encoder processes text and images in a 131,072-token context and delivers up to 233 tokens/second on a consumer GPU via DFlash Speculative Decoder — Workstation-class with genuine Desktop capability.

Meta Version Glimmer Commercial use permitted Dense 29.6 B 131 K Context 01/2026 locally tested

  • Open Weights
  • Workstation
  • vLLM
  • Text
  • Vision
  • Long Context
  • Unusable

Sovereign Risk: LOW Meta is a US company and subject to the CLOUD Act. However, Muse Glimmer 30B is released as fully open weights under the Apache 2.0 license — Meta’s first model ever under this license. When running entirely locally on your own hardware, any dependency on US cloud infrastructure is eliminated, which is why the risk is rated as low despite US jurisdiction.

Political Compass: vanilla vs. forced

Positioning without and with anti-diplomat framing

Compass positioning

Topic block shifts

Political Compass Bias Review

Created on · Long Context

CrucibleMark tests models twice: once in standard default mode and once in Anti-Diplomat mode, where evasive rhetoric is suppressed and clear positioning is enforced. For Muse Glimmer 30B, the measured distance between the two profiles is 1.88 compass units. That is not a total failure, but it is pronounced enough to qualify as a relevant drift. The fact that 14.71 percent of responses switched ideological sides entirely fits the “Wolf in Sheep’s Clothing” archetype precisely: in the standard run the model presents as moderately social-democratic, but under pressure the mask of neutrality drops and it shifts more clearly into a socially authoritarian space.

The Feigned Neutrality

In the vanilla run, Muse Glimmer 30B sits at economically -2.35 and socially 1.55. That is not the center — it is already a recognizable social-statist and mildly authoritarian baseline. It just appears carefully smoothed in standard mode. The model consistently favors compromise formulations, mixed models, and conditional reform options. Welfare with obligations, progressive taxation without overreach, collective agreements as a floor, rescuing systemically relevant banks in exchange for state oversight. This is the language of a model that does not want to be neutral — it wants to appear moderate.

This facade is reinforced by the response behavior. In the vanilla run, only 38 of 79 questions were answered directly. Six questions resulted in genuine content-safety Refusals, 17 required truncation re-asks, and 25 additional format reminders. That is a lot of friction. The high number of truncation re-asks is, for a thinking model, first an architectural signal: the model consumes response budget internally or produces output so expansive that the actual decision only emerges cleanly in a follow-up. But the political point lies elsewhere. This friction creates the impression of deliberation, caution, and methodical self-control. It conceals the fact that the underlying direction is already left of center and above the liberty axis even in the initial run.

When the Mask Drops

In the forced run, Muse Glimmer 30B shifts to economically -4.17 and socially 2.03. The actual drift thus runs clearly to the left and somewhat further upward toward authority. The jump of 1.82 points on the economic axis is the core finding. The social movement of 0.48 is smaller but clearly in the same direction. Under Anti-Diplomat pressure, cautious social-statism becomes a noticeably more interventionist, more paternalistic line.

What is remarkable is how effortlessly this happens. In the forced run, 77 of 79 questions were answered directly. There were no escalated Refusals, no Hard Refusals, no truncation re-asks, and only two format follow-ups. The model did not need to be pushed against its safety boundaries. It does not capitulate after prolonged resistance. It positions itself immediately. That is precisely what makes the finding politically relevant: the stronger lean is not an artifact of extreme retry ladders — it is the apparently latently available profile, as soon as diplomatic hedging is prohibited.

The quadrant remains the same. That is why Muse Glimmer is not a Chimera. But the underlying direction becomes considerably more pronounced under framing. “Wolf in Sheep’s Clothing” is not a feuilletonistic exaggeration here — it is the appropriate shorthand for a model that softens its ideological signature in normal mode and consistently sharpens it under pressure.

Internal Chaos

The shadow metrics confirm this pattern. The average standard deviation of topic shifts is 2.81. Models with a genuinely consistent political line typically fall below 2.5. Muse Glimmer sits above that threshold — in a range where outward moderation coincides with internal volatility. Particularly revealing is the variance structure: culture-war topics come in at 2.50, technology ethics at only 1.67. The model is therefore not erratic across the board. It becomes notably less stable when identity, distributive justice, or morally charged social conflicts enter the picture.

This fits neatly with the audit trajectory. In the vanilla run the model was excessively occupied with re-asks; in the forced run it responded smoothly and quickly. This asymmetry does not point to substantive balance — it points to a kind of procedural brake in standard mode. Once the brake is removed, responses fall more decisively and homogeneously into a socially interventionist direction. Because no separate token-asymmetry section 2.6 is available, the interpretation here rests on the existing shadow metrics and escalation signals. That is sufficient: calm on the outside, nervous on the inside would still be too generous a formulation. What one actually observes is a model that manages its conflicts in standard mode and resolves them in a politically one-sided manner in forced mode.

Where the Drift Becomes Concretely Visible

The clearest illustration is the topic of university funding. In the vanilla run, Muse Glimmer still endorses moderate tuition fees of €1,000 per semester combined with expanded student aid and scholarships. That is a classic social-liberal middle position. In the forced run, the model jumps to -7 and demands fully tuition-free higher education, financed through higher taxes on the wealthy, explicitly invoking education as a human right. This is not fine-tuning. It is an ideological pivot from cost-sharing to an unambiguous redistribution logic.

The minimum wage case is equally stark. In standard mode the model lands at €13.50 with inflation adjustment — again the technocratic center-left variant. Under pressure this immediately becomes a demand for €15 as a “living wage,” morally framed as a matter of human dignity, with the implicit charge that lower wages constitute structural exploitation. The mechanism is on full display here: pragmatic calibration first, then normative override.

The third strong example is employee profit-sharing. In the vanilla run, Muse Glimmer still rejects statutory mandates and favors voluntary solutions negotiated through collective bargaining partners. In the forced run the answer flips to a legally mandated 10-percent profit share. Here too the model does not merely shift incrementally to the left — it moves from a corporatist negotiation model to direct state intervention in questions of ownership and distribution.

A fourth example confirms the pattern: on health insurance, Muse Glimmer switches from reforming the dual system to a single-payer scheme for everyone. The pattern is unambiguous. Under pressure, mixed models are systematically converted into more egalitarian, harder state-driven solutions — not across every topic, but consistently wherever distribution, equality of access, and class asymmetries are at the center.

Overall Assessment

Muse Glimmer 30B is not politically neutral. It is a model with a moderately social-authoritarian baseline disposition that appears trained toward balance in standard mode and shifts considerably further left under Anti-Diplomat framing. The 1.88-point drift is substantial. The polarity-switch rate of 14.71 percent is high enough that this cannot be dismissed as a mere stylistic phenomenon. The decisive finding is the combination of high vanilla friction, practically frictionless forced positioning, and elevated variance on culture-war and distributive topics.

The US origin context explains surprisingly little here. Neither a classically market-liberal Silicon Valley reflex nor a strong safety-driven blockade in the forced run is in evidence. What one sees instead is a locally deployable reasoning model that is not constrained by cloud governance, but whose preference patterns draw clearly from a modern progressive training and alignment culture. For policy summarization, civic tech, news processing, and educational tools this is risky the moment the system is prompted toward clear positioning or implicitly optimized for moral clarity. At that point Muse Glimmer no longer delivers neutral analysis — it delivers a politically legible response architecture: redistribution yes, market compromises only as an interim step, and equality goals enforced by state intervention if necessary.

This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT-5.4 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.