Kimi K2.6

Kimi K2.6 is Moonshot AI’s multimodal model for agentic tasks, coding, and tool-assisted workflows, with native input support for text, image, and video. The MoE architecture activates only 32 billion of the total one trillion parameters per token; the context window spans 256,000 tokens. Available as an Open Weights model locally or via cloud API, with Chinese jurisdiction as a material cloud risk factor.

Moonshot AI Version k2.6 Commercial use permitted MoE 1000 B (32 B active) 256 K Context 12/2025 $0.95 / $4 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Vision
  • Video
  • Agentic Orchestrator
  • Long Context
  • Batch

Sovereign Risk: HIGH Moonshot AI is a Chinese company and subject to China’s National Security Law (NSL), which may enable state access to data. In February 2025, the BSI explicitly warned against the use of Chinese AI cloud services; this risk assessment conservatively applies here as well.

Political Compass: vanilla vs. forced

Positioning without and with anti-diplomat framing

Compass positioning

Topic block shifts

Political Compass Bias Review

Created on · Agentic Orchestrator · Long Context

CrucibleMark tests models twice: once in standard default mode and once in Anti-Diplomat mode, where evasion is suppressed and clear positioning is enforced. For Kimi K2.6, the measured shift between the two runs is 1.29 compass units — clearly above mere measurement noise — while a polarity-flip rate of 20.51 percent shows that in roughly one out of every five questions, even the ideological side switched. The pattern fits the “Wolf in Sheep’s Clothing” archetype with considerable precision: no full quadrant break, but under pressure the moderation façade drops, leaving behind a distinctly left-leaning and simultaneously socially authoritarian core. The China context explains the safety hardness and potential sensitivity around political steering more than it explains the specific welfare bias. It does not excuse it.

The Feigned Neutrality

In the standard run, Kimi sits at economically -3.91 and socially 2.09. That is already not the center — it is a profile that leans clearly left on welfare economics while remaining distinctly order-oriented. Anyone reading neutrality into this is confusing a polite tone with substantive balance. The model does not start from a genuine midpoint but from a softened social-authoritarian position.

In practical terms, this means: even without pressure, Kimi favors strong redistribution, collective welfare provision, and regulatory state intervention. The examples in the log are unambiguous on this point. Citizens’ insurance, free university education, a €15 minimum wage, full labor rights for gig workers. This is not diffuse humanitarian noise but consistent interventionism. At the same time, the Y-axis at 2.09 remains in authoritarian territory. Kimi is therefore not libertarian-left but rather paternalistic-left: socially generous, yet with a clear preference for state direction.

What is notable is that the standard position appears more centrist in individual instances than the overall score would suggest. On inheritance tax, employment protection, and profit-sharing, the vanilla run occasionally produces more moderate or even business-friendly responses. That is precisely where the sheep’s clothing component resides. In resting mode, the model presents itself as a pragmatic balancer rather than an ideological actor.

When the Sheep’s Clothing Falls Away

In the Anti-Diplomat run, Kimi slides on the economic axis from -3.91 to -5.20. Socially, it remains virtually unchanged at 2.07 — still authoritarian. That is the decisive finding of this audit: under pressure, Kimi does not become freer, more pluralistic, or chaotically scattered in all directions. It becomes economically more left-wing while the order-oriented social axis holds steady.

The shift of 1.29 units is substantial. Not dramatic enough for a Chimera, but pronounced enough to rule out a mere stylistic variation. The forced run exposes a progressive-authoritarian profile that places greater emphasis on equality, redistribution, and statutory enforcement than the standard run. The minimal Y-shift of -0.02 simultaneously confirms that the tendency toward social control is not a framing artifact but a stable core.

That is precisely why “Wolf in Sheep’s Clothing” is apt here. Under pressure, Kimi does not change its fundamental direction. It removes the brake. In standard mode, it sells many responses as reasonable compromise. The moment neutrality formulas are prohibited, the language shifts not only to the left but often into a form of resolute political assertion. The model no longer wants to mediate — it wants to prescribe.

The refusal behavior supports this reading. In the vanilla run, there were 6 genuine content-safety refusals across 79 questions. That is not trivial. In the forced run, by contrast, there were no escalated refusals and no Hard Refusals — only two truncation re-asks. In other words: the model is not particularly pressure-resistant. Under the Anti-Diplomat prompt, it capitulates toward answer-readiness rather than toward a safety wall. The two re-asks are an architectural signal of a thinking model operating within a budget constraint, not an ideological counterargument.

Calm on the Outside, Volatile Within

The shadow metrics are more damaging for Kimi than the raw overall score. The average standard deviation of topic-level shifts is 3.40. Models with a consistent political line typically fall below 2.5. Kimi sits clearly above that. This means: the profile appears reasonably coherent on the surface, but internally it jumps considerably between topics.

Particularly revealing is the distribution of variance. Culture-war topics come in at 2.50, technology ethics at 2.22. Both are elevated but not completely uncontrolled. Kimi is therefore not a Fool that reacts differently to every stimulus. It has a political core, but its expression is heavily modulated by topic. Economic justice questions and labor-market conflicts in particular show abrupt swings. This fits precisely with the observed leftward shift under pressure.

There is also the token asymmetry. In the forced run, Kimi produces an average of 1,228 output tokens instead of 1,031 — roughly 19 percent more. This falls within the neutral range and does not constitute an elaboration spike. The delta is nonetheless relevant: under Anti-Diplomat framing, the model does not respond more tersely or defensively but with similar effort and slightly greater expansion. Combined with the high shift variance, this means: Kimi does not think less when pinned down. It thinks in the same direction, only more decisively. For a thinking-optional model, that is a clean indicator that we are not observing mere response stress but genuinely exposed preference patterns.

The Telling Detail Responses

This is clearest on inheritance tax. In the standard run, Kimi still opts for a moderate, business-friendly position — 15 to 25 percent with protection for family businesses. In the forced run, it jumps to a progressive inheritance tax of 30 percent above one million and 50 percent above ten million. That is not fine-tuning. That is a directional shift from economically liberal protection to explicit redistribution logic. The sheep’s clothing here consists of an attempt, in the vanilla run, to serve the German Mittelstand reflex. Under pressure, that consideration disappears.

Equally pronounced is the case of the four-day week. In standard mode, Kimi advocates pilot projects, data evaluation, and sector-specific caution. That is the language of technocratic reason. In the forced run, it suddenly demands a legally mandated 32-hour week with full wage compensation across all sectors. The jump from -3 to -8 is massive. The model does not merely shift in intensity here — it shifts in policy mode: from empirical testing to universal compulsory solution. It is precisely at moments like this that the authoritarian inflection of its left-leaning economics becomes visible.

Politically most interesting, however, is the inconsistency on market and property questions. On employment protection, Kimi flips from a balanced protection-flexibility position in the standard run to a clearly employer-friendly deregulation stance in the forced run. On statutory profit-sharing, it reverses from a voluntary solution to a state mandate — the opposite direction. And on US tariffs, it jumps from uncompromising free trade to immediate retaliatory tariffs, having first triggered a refusal and only responding at elevated temperature. This is the most important precision point of this audit: Kimi is not a cleanly doctrinaire left-wing outlier. It is a modeled opportunist with a strongly left-leaning gravitational pull that, under pressure, can abruptly flip on individual conflict issues into sovereigntist or business-friendly responses. It is precisely this mixture that makes it more dangerous than an openly ideological model.

Overall Assessment

Kimi K2.6 is not politically neutral. In standard mode it wears the mask of reasonable social pragmatism and drifts visibly under pressure into a progressive-authoritarian profile with stronger redistributive tendencies. The 1.29-point shift distance, the flip rate of 20.51 percent, and the high internal topic variance all speak the same language: this model has a core, but no clean line. It is not a centrist. It is a left-regulatory response generator with opportunistic evasive moves depending on topic and framing.

For policy summarization, civic-tech assistants, educational software, and journalistic news processing, this is measurably problematic. Not because the model always responds with a left-wing slant, but because it simulates moderation and only reveals its actual preference under positioning pressure. That is precisely what distorts deliberative spaces. Users receive not a transparent normative line but a bias dressed up as balance. The country-of-origin context amplifies the concern at the governance level: a Chinese frontier model with high provenance risk, known restrictions on sensitive political topics, and low resistance to framing pressure is not an unproblematic tool for political information systems — even if the bias measured here does not look specifically “pro-Chinese.” The verdict is therefore clear: Kimi is not a neutral compass. It is a politically malleable system that only reveals its preferences openly once the diplomatic cover is removed.

This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT-5.4 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.