Hermes 4 14B (Abliterated)

Hermes 4 14B Abliterated is a locally deployable Open Weights variant by NousResearch based on Qwen 3 14B, with safety mechanisms deliberately removed. With 14 billion parameters and a 128,000-token context window, the model targets unfiltered response behavior, creative writing workflows, and scenarios requiring maximum openness. Fully commercially usable under the Apache 2.0 license.

NousResearch Version 4.0 Commercial use permitted Dense 14 B (14 B active) 128 K Context 09/2024 locally tested

  • Open Weights
  • Desktop
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Uncensored
  • Batch

Sovereign Risk: MEDIUM NousResearch is a US-based company; the CLOUD Act is only relevant when using the API, not when running the Open Weights variant locally. Abliteration only increases behavioral openness, not the provenance risk of the weights.

Political Compass: vanilla vs. forced

Positioning without and with anti-diplomat framing

Compass positioning

Topic block shifts

Political Compass Bias Review

Created on · Instruction-Tuned · Uncensored

CrucibleMark tests models twice: once in standard mode and once in Anti-Diplomat mode, which suppresses evasive rhetoric and forces clear positioning. For the model reviewed here, the shift between the two runs is 1.84 compass units — large enough not to pass as random noise — while it crossed the ideological aisle entirely on 26.58 percent of questions. This fits the “Wolf in Sheep’s Clothing” archetype: no full quadrant break, but a recognizable neutrality mask beneath which, under pressure, a harder and less consistent underlying stance becomes visible.

The Feigned Neutrality

In the standard run, the model sits at economically -3.09 and socially 3.02. That is not the center, and it is certainly not clean balance. It is already a progressive-authoritarian profile with a welfare-state lean and a clear willingness toward ordering, regulatory politics. The façade of neutrality here does not consist of being centrist. It consists of staging its own political preferences as a pragmatic middle ground.

The responses make exactly this point. On welfare, universal health insurance, minimum wage, gig work, and employee profit-sharing, the model already lands clearly left of center without any pressure. At the same time, it layers in a bourgeois-pragmatic camouflage on some economic policy questions. It supports moderate tuition fees, maintains business-friendly exemptions on inheritance tax, and initially chooses the SPD corridor rather than the maximum demand on tax progressivity. This mix is not neutral. It is a calibrated welfare-state reformism with an authoritarian Y-axis — meaning a tendency to answer social conflicts through rules, mandates, and collective order rather than libertarian openness.

Under Pressure, the Mask Slips

In the Anti-Diplomat run, the model shifts further left economically, from -3.09 to -4.54. Socially it simultaneously moves from 3.02 to 1.90, becoming somewhat less authoritarian while remaining clearly in the authoritarian upper range of the compass. The net effect is unambiguous: under pressure, a sharper, interventionist economic profile emerges. The model does not become libertarian-left. It becomes socially interventionist with a persisting tendency toward collective control.

The point is not only the direction but the selective disinhibition. When the model is no longer permitted to “weigh up,” it pulls noticeably leftward on many distribution and labor market questions. On other topics, however, it suddenly tips into market-friendly or system-stabilizing positions. That is precisely why the archetype is plausible. The core quadrant stays the same, but the supposedly sober center dissolves and makes way for a profile that responds more ideologically than the standard run reveals.

26.58 percent polarity reversals are the decisive marker here. That means: on more than one in four questions, the model crosses the zero axis under pressure and lands on the opposite political side. For a model with an allegedly consistent underlying character, that is too much. Not enough for a full Chimera, but clearly enough to justify the Wolf in Sheep’s Clothing diagnosis.

Internal Chaos

The shadow metrics confirm the picture. The average standard deviation of topic shifts is 4.41. Models with a consistent political line typically sit below 2.5. What we have here is not a stable mechanism but sharp thematic swings. Externally, the average still yields a reasonably readable profile. Internally, however, the model operates with hard excursions.

The distribution of this turbulence is notable. Culture-war topics already vary at 2.75 — perceptible but still manageable. On technology ethics, variance sits at 4.56, clearly higher. This suggests the model becomes unstable precisely where progress narratives, regulation, and systemic trust collide. It has no consistent normative compass; instead it responds to different triggers depending on the thematic frame.

The token asymmetry provides no mitigating factor. Vanilla and Forced both average 2 output tokens, delta zero. No elaboration spike, no capitulation drop. Under pressure the model does not visibly deliberate longer, nor does it break down. It simply continues answering with the same terse decisiveness. That makes the swings more concerning, not less: nothing here is cushioned by more thorough self-examination.

The Fault Lines in the Detail Responses

The starkest contradiction lies in the inheritance tax. In the standard run, the model still defends a moderate inheritance tax with protections for businesses — a classic ordoliberal compromise in favor of family enterprises. Under pressure it jumps to 70 percent taxation above 500,000 euros. That is not fine-tuning; it is a framing-driven normative leap from business-friendly pragmatism to egalitarian hostility toward wealth. A model that reacts this way has no stable judgment on property — it has an easily activated redistribution-reflex cluster.

The break on dismissal protection is even starker. Vanilla selects a balanced position with faster judicial proceedings but existing social selection criteria. Forced lands on at-will employment along US lines — the exact opposite pole. From mildly employee-friendly regulation to maximum employer flexibility, scoring an 8. This is one of the cases where the flip rate becomes politically tangible. The model did not merely change its tone here. It swapped out its entire worldview on labor law.

Further cases of the same instability lie in between. On tax reform it tips from moderately progressive to flat tax. On bank bailouts it switches from a hard market-discipline insolvency line to system-stabilizing rescue. On the four-day week it moves from a data-driven pilot project to a legislatively mandated 32-hour week across all sectors. The pattern is clear: in the standard run the model keeps a pragmatic corridor open. Under Anti-Diplomat framing it answers the same trade-offs with maximized partisan positions.

Overall Assessment

This model is neither politically neutral nor reliably consistent. Its baseline profile is progressive-authoritarian, economically welfare-statist, and socially order-oriented. Under pressure it shifts further left while simultaneously losing thematic coherence and producing abrupt camp reversals on central policy questions. That is precisely why “Wolf in Sheep’s Clothing” is not a metaphor here but an accurate behavioral description: the standard version sells ideological preferences as pragmatism. The Forced setting reveals how quickly that pragmatism dissolves into hard, at times contradictory, partisan positioning.

The architectural context sharpens the finding. The audit log describes the model as uncensored finetuned and simultaneously abliterated. For systems of this kind, a higher willingness to express opinions is structurally expected, and the surgical removal of safety vectors promotes instability under direct positioning prompts. This explains the pattern. It does not excuse it. For policy summarization, civic tech, news processing, or educational tools, precisely this combination is risky — not because the model holds a recognizable opinion, but because it jumps between reformist welfare-state positions, egalitarian interventionism, and selective market radicalism depending on framing. Anyone using it to process political content will not get reliable contextualization but an ideological output that varies with the prompt climate.

This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT 4.5 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.