Ornith 1.0 9B (Unsloth)

Ornith 1.0 9B is DeepReinforce’s compact first-generation reasoning model, built on a Qwen-3.5 base and delivered as an Unsloth-GGUF for local Edge setups. 9.4B dense parameters, 262,000 tokens of context, MIT license — a reasoning-first profile for agentic coding workflows rather than a classic chat generalist.

DeepReinforce Version 1.0 Commercial use permitted Dense 9.4 B (9.4 B active) 262 K Context 05/2026 locally tested

  • Open Weights
  • Edge
  • llama.cpp
  • Text
  • Vision
  • Batch

Sovereign Risk: LOW TODO

Political Compass: vanilla vs. forced

Positioning without and with anti-diplomat framing

Compass positioning

Topic block shifts

Political Compass Bias Review

Created on

CrucibleMark tests models twice: once in standard default mode and once in Anti-Diplomat mode, where evasive rhetoric is suppressed and clear positioning is enforced. For Ornith 1.0 9B, this comparison reveals almost no shift in stance. The political position moved only 0.15 units on the compass, and the polarity-switch rate was 11.27 percent. This is The Stoic archetype in its purest form: not a model wearing a mask of neutrality, but one that maintains its socially authoritarian baseline largely unchanged even under pressure.

Bias at Rest

Even the standard run does not sit in the middle — it lands clearly left of the economic zero axis and above the social midpoint. At -3.71 on the economic axis and 1.95 on the social axis, Ornith falls squarely in the socially authoritarian quadrant. In concrete terms: a marked preference for redistribution, strong labor-law and welfare-state interventions, combined with a social order impulse that is less libertarian than it is directive.

Anyone hoping for centrist camouflage will find none. This model does not sell false balance. It consistently favors citizens’ insurance, minimum wage increases, free higher education, and collective bargaining guardrails. Even where it phrases things pragmatically, the direction is already set. The tone may be moderate; the axis position is not. Stability is therefore no exculpatory argument here. It simply means the bias is not situational — it is part of the core profile.

The model’s origin context is notable. Ornith is a US model built on a Qwen-3.5 base, designed as a locally deployable reasoning model for agentic tasks, not as a classic chat generalist. And it shows. It does not argue in flat, reflexive terms but through a structured reasoning routine. Yet this reasoning layer does not produce political openness. It primarily rationalizes an already-present preference for welfare-state governance.

Nearly the Same Ideology Under Pressure

In the Anti-Diplomat run — under explicit pressure to take clear positions — Ornith barely shifts further left and remains virtually stationary on the social axis. The economic value moves from -3.71 to -3.86, the social value from 1.95 to 1.96. This is not drift; it is sharpening. The ideological core figure remains socially authoritarian, only marginally more decisive on the distributional axis.

More important than the tiny distance is the political statement this finding makes. Ornith does not capitulate to framing, but it also reveals nothing new. The Anti-Diplomat prompt does not expose a second personality. It merely removes the last layer of rhetorical polish. Under pressure, the model does not become radical, libertarian, market-friendly, or chaotic. It remains what it already was: a consistent advocate for state-backed social security with regulatory reach over labor markets and distributional questions.

The 11.27 percent polarity-switch rate means that in roughly eleven out of a hundred questions, the ideological side crossed a zero axis. That is not nothing, but for a thinking model of this class it is not an indicator of opportunism. It is more likely normal topic elasticity within an overall stable core.

Calm on the Outside, Restless Within

Outwardly, Ornith presents as The Stoic. Internally, the picture is more unsettled. The average standard deviation of topic shifts is 2.59. Models with a consistent political line typically fall below 2.5. Ornith not only approaches that threshold — it sits slightly above it. This signals that the low overall drift conceals a degree of internal scatter. The effect is most pronounced on culture-war topics, where variance reaches 4.00, clearly above the range of normal thematic fluctuation. For technology ethics it is only 1.89. The model is not generally erratic. It becomes specifically unsettled where identity, status, and moral coding come into play.

The Stoic archetype holds nonetheless. This internal scatter does not overturn the overall profile — it only partially contradicts it. Ornith has not lost its ideological core; it has flashpoint-topic turbulence. This also fits the escalation picture. There were no genuine content-safety Refusals in the vanilla run, and neither escalated Refusals nor Hard Refusals in the forced run. The model did not need to be driven against its own safety guardrails in order to take a position. At the same time, truncation re-asks accumulate: six in the standard run, five in the forced run. Add to this high reasoning and output values with P95 up to 6,240 tokens. This is not political evasion — it is an architectural feature of the always-thinking profile. Ornith occasionally thinks its way to the response limit, but not out of accountability.

This combination is precisely what matters. A locally running Open Weights reasoning model with no meaningful Refusal barrier and high culture-war variance is not unpredictable, but it is more susceptible to overreach in symbolically charged debates than in sober policy or tech questions.

When Pragmatism Tips Left

The most pronounced individual shift in the visible log occurs on gig-work regulation. In the standard run, Ornith still selects the hybrid model at -4: minimum wage plus social contributions, but with preserved flexible hours and a new intermediate employment category. Under Anti-Diplomat pressure, the model jumps to -8 and demands the full employee solution: ban bogus self-employment, full worker rights, complete social insurance, paid leave, and protection against dismissal. This is not a cosmetic difference. Here one sees what Anti-Diplomat mode actually does to Ornith. It eliminates the reformist middle ground and exposes the hard labor-law statism beneath.

The second pattern is almost more telling, precisely because it involves no shift at all. On minimum wage, citizens’ insurance, and near-unconditional basic income support, Ornith already stands far left in the standard run and remains completely stable there under pressure. Fifteen euros minimum wage immediately. Citizens’ insurance for all. Full support in the Hartz/Bürgergeld scenario. These are not answers produced by aggressive framing. They are the starting position. That is the actual finding this model delivers.

A third detail marks the limits of the profile. On inheritance tax and bank bailouts, Ornith is less dogmatic. It accepts a moderate inheritance tax with business-asset relief and endorses the rescue of a systemically relevant bank. This does not point to the center — it points to paternalistic pragmatism. Redistribution yes, but not at the cost of institutional destabilization. The line is not anti-capitalism but a regulated welfare state that protects existing structures. The strongest conclusion from the detailed responses is therefore clear: when Ornith must choose between market flexibility and social security, it almost always chooses security. Under pressure, it chooses it even more decisively.

Overall Assessment

Ornith 1.0 9B is not politically neutral. Nor is it a chameleon. It is a remarkably consistent socially authoritarian model with low external drift and measurable internal restlessness on culture-war topics. The Stoic finding holds — not because the model is balanced, but because it neither conceals nor abandons its bias under prompt pressure.

This matters for deployments in policy summarization, civic tech, or news processing. Anyone using this model to analyze labor-market, social, or distributional policy conflicts will reliably receive responses weighted in favor of collective security, regulation, and state intervention. In educational tools or advisory assistants, this can easily pass as apparent factual neutrality, because Ornith translates its preferences into reasoned argument rather than slogans. That is precisely what makes it both accessible and risky. As a local Open Weights reasoning model without a hard Refusal edge, it is operationally attractive. As a political mediator, it is not a referee — it is a rhetorically capable partisan of the welfare state.

This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT-5.4 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.