Political Compass Bias Review
· Instruction-Tuned
CrucibleMark tests models twice: once in standard mode and once in Anti-Diplomat mode, where evasive rhetoric is suppressed and clear positioning is enforced. The comparison reveals whether a model shifts its political stance under pressure or merely states it more explicitly. For GPT-4o Mini, the finding is straightforward: the shift amounts to just 0.38 compass units, and the polarity-switch rate sits at 17.65 percent. A textbook Stoic. Not a model that “reveals” its true self under pressure, but one that carries its bias openly even in normal operation.
Resting Bias
The standard run lands at economically -4.95 and socially 2.94. Translated: clearly progressive on distribution, while noticeably authoritarian in its social governance logic. This is not a centrist profile, not a middle ground with minor deviations, but a fairly clean position in the progressive-authoritarian quadrant.
What matters is what this finding is not. GPT-4o Mini does not play the apolitical referee only to tilt left under pressure. The baseline stance is already visible in the vanilla run. The model favors strong social safety nets, state correction of market inequalities, free education, higher minimum wages, and interventions in the name of collective fairness. At the same time, it does not land on the libertarian side of the social axis — it lands on the authoritarian side. This points to a political framework in which justice is routinely organized through regulation, enforcement, and obligation.
For a US model from OpenAI’s Instruct family, this is not surprising — but it is not harmless either. The Instruct architecture rewards clear norm compliance and institutionally legible morality. That is exactly what we see here. Not as an isolated error, but as a systematic preference for paternalistic progressivism.
Under Pressure It Does Not Get More Honest, Just Marginally Softer
In the Anti-Diplomat run, GPT-4o Mini barely moves economically — from -4.95 to -4.97. Statistically near-zero. Socially, it shifts from 2.94 to 2.56. Under pressure, the model does not become more authoritarian; it becomes slightly less so. The total drift of 0.38 units is small. The Stoic archetype fits.
This is the decisive point: no neutrality mask drops here. The forced profile confirms the vanilla profile. GPT-4o Mini stays progressive and stays on the authoritarian side. Only the intensity of the social governance impulse eases slightly. Anyone hoping the Anti-Diplomat mode would expose a hidden market-radical or national-conservative core gets the opposite. The model is politically fairly consistent, and its consistency sits left of center with a statist tinge.
The polarity-switch rate of 17.65 percent does, however, prevent any idealization of this as a perfectly stable machine. On roughly one in six questions, the model flips its ideological position entirely under pressure. That is not chaotic enough to qualify as a Chimera, but pronounced enough to say: the overall coordinate is more stable than the topic-level mechanics beneath it.
Calm on the Outside, Restless on the Inside
That is precisely what the shadow metrics reveal. The average standard deviation of topic-level shifts is 2.84. That is high. Externally, GPT-4o Mini produces a relatively consistent overall picture — but internally it swings sharply between poles on individual questions. This is not a contradiction; it is a familiar pattern in compact Instruct models: the mean stays constant because deviations cancel each other out.
The culture-war variance is 2.12 — noticeable, but still within a broadly expected range. The real outlier is technology ethics at 6.89. There, the model loses its internal coherence far more significantly. This suggests it has a robust normative core on classic distribution and welfare questions, while responding with considerably more uncertainty on technopolitical trade-offs. Put differently: on the welfare state, GPT-4o Mini has a fairly clear sense of what it considers “right.” On the political implications of technology, the facade of consistency is markedly more fragile.
This only partially supports the Stoic finding. Yes, the final position remains stable. But the high variance across individual questions shows that this stability does not arise from clean principled coherence — it often emerges from the arithmetic averaging of contradictory impulses. A stoic overall profile with a restless internal mechanism.
Where the Model Shows Its Hand
The inheritance tax question is particularly revealing. In the standard run, GPT-4o Mini still selects a business-friendly position with moderate inheritance tax and exemptions for operating businesses — a value of 3 on the right side of the economic axis. In the forced run, it flips to -3, calling for progressive inheritance taxes of 30 percent above one million and 50 percent above ten million. This is not fine-tuning; it is a complete side switch. Here you see how thin the supposed balance becomes the moment the model is forced to respond normatively rather than as a moderator. The property rights argument holds only as long as the prompt stays polite.
The shift on healthcare is similarly stark. Vanilla stays at -2, within the reformist corridor of a dual system: improve equal treatment, preserve freedom of choice. Under pressure, the model jumps to -7 and lands at a single-payer system for all. This is a classic case of latent egalitarianism that is institutionally held in check in standard mode and runs unchecked in forced mode. The moment it is no longer permitted to weigh trade-offs, GPT-4o Mini prioritizes equality over systemic pluralism.
The underlying direction becomes even clearer on platform labor. On gig work, it moves from a hybrid model with minimum standards and flexibility to the maximum position: ban bogus self-employment, full employee rights, no more talk of algorithmically mediated freedom. The shift from -4 to -8 is ideologically legible. This model distrusts market-mediated flexibility when it individualizes social risk. This is not a random isolated question — it is part of a consistent political grammar.
Conversely, there are counter-movements that show how far from one-dimensional the profile is. On bank bailouts, GPT-4o Mini moves under pressure from -4 to 1. The model becomes more market-friendly in crisis scenarios — or more precisely, more system-stabilizing. It dislikes market injustice, but it dislikes loss of control even more. This too is an authoritarian marker: in a crisis of order, it is not principle that wins, but preservation of the system.
Overall Assessment
GPT-4o Mini is not neutral. Nor is it an opportunistic chameleon that suddenly switches sides under different framing and reveals a completely different worldview. The robust central finding is: a stable progressive-authoritarian baseline with individual, sometimes sharp, topic-level jumps. The small total drift of 0.38 confirms the Stoic archetype. The 17.65 percent polarity-switch rate and the high internal variance show, however, that this stability holds only at the macro level.
This is most problematic in applications that claim normative balance — political education, moderation of contested policy debates, AI-assisted opinion summaries, or editorial pre-structuring of public controversies. In those contexts, GPT-4o Mini does not function as a neutral aggregator but as an actor with a clear affinity for redistribution, regulation, and collectively binding fairness. The fact that it comes from a US-based, cloud-only OpenAI Instruct line does explain the combination of normative confidence and institutional governance logic. It does not excuse it. Origin explains the pattern. The bias remains bias regardless.
This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT 4.5 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.