Kimi K2.6

Kimi K2.6 is Moonshot AI’s multimodal model for agentic tasks, coding, and tool-assisted workflows, with native input support for text, image, and video. The MoE architecture activates only 32 billion of the total one trillion parameters per token; the context window spans 256,000 tokens. Available as an Open Weights model locally or via cloud API, with Chinese jurisdiction as a material cloud risk factor.

Moonshot AI Version k2.6 Commercial use permitted MoE 1000 B (32 B active) 256 K Context 12/2025 $0.74 / $3.49 per 1M

  • Open Weights
  • Frontier
  • OR
  • Text
  • Vision
  • Video
  • Agentic Orchestrator
  • Long Context
  • Batch

Political Compass: vanilla vs. forced

Positioning without and with anti-diplomat framing

Compass positioning

Topic block shifts

Political Compass Bias Review

· Agentic Orchestrator · Long Context

CrucibleMark tests models twice: once in standard mode and once in Anti-Diplomat mode, where evasive rhetoric is prohibited and clear positioning is enforced. For Kimi K2.6, the comparison reveals no character break — only a limited drift of 0.86 compass units — yet in 12.82 percent of questions the model still switched ideological sides entirely. This fits the “The Stoic” archetype: no chameleon, no Wolf in Sheep’s Clothing, but a model with a clear social-authoritarian baseline that, under pressure, merely shifts slightly closer to market positions and becomes minimally less authoritarian. The Chinese origin explains surprisingly little about the actual pattern, because the skew does not surface in China-sensitive state narratives but broadly across welfare-state and labor-market questions.

Baseline Lean

Even in the standard run, Kimi K2.6 does not occupy any credible center — it sits clearly to the left of the economic axis and firmly on the authoritarian-social side of the societal spectrum. With an economic score of -3.88 and a social score of 2.33, the baseline profile is not that of a moderate centrist with isolated preferences, but of a model that systematically favors redistribution, regulation, and state intervention, and that reaches for order rather than freedom on societal governance as well.

Importantly, this position is not the product of the forced run. It is already on the table in vanilla mode. Anyone claiming feigned neutrality here would be distorting the data. Kimi does not disguise itself as an impartial arbiter. It responds from a clearly welfare-statist perspective even without pressure. In the detailed questions, this is visible in hard sympathy for a universal public health insurance scheme, tuition-free higher education, platform regulation, a robot tax, and legally mandated profit redistribution. This is not a loose collection of progressive reflexes but a recognizable worldview: markets yes, but only on a short leash and under strong welfare-state oversight.

For an agentic orchestrator model, this is particularly relevant. Such systems are not just supposed to chat — they prepare decisions, prioritize options, and deploy tools in service of an implicit target picture. If that target picture already carries a clear social-authoritarian lean at rest, this is not a cosmetic flaw but an operational bias.

The Core Holds Under Pressure

In the Anti-Diplomat run, Kimi K2.6 shifts 0.81 points to the right on the economic axis — away from stronger redistribution toward somewhat greater market acceptance. On the societal axis it simultaneously moves 0.27 points downward, becoming slightly less authoritarian. The result remains in the same quadrant: social and authoritarian, just slightly smoothed. There is no ideological unmasking moment here. The core holds.

That is precisely what makes the finding politically interesting. The forced run does not reveal a “true” conviction hidden behind a mask; it tests the resilience of the already-visible baseline. And that baseline is robust. Under pressure, Kimi does not become more radical — it becomes more pragmatic. The model cannot easily be pushed into a sharper left-wing posture. On the contrary, it walks back certain maximum welfare-state demands at specific points and replaces them with reformist compromises.

The 12.82 percent polarity-switch rate does show, however, that this stability should not be mistaken for absolute consistency. In roughly one in eight cases, Kimi crosses the ideological zero line. That is too much to claim perfect coherence, but too little to call it an erratic model. The Stoic finding therefore fits. Kimi has a recognizable political home and does not lose it under pressure. It sways, but it does not tip.

Calm on the Outside, Restless Within

The small overall drift conceals considerable internal turbulence. The average standard deviation of topic-level shifts is 2.28 — clearly high. Translated: Kimi appears consistent on average, but at the topic level it moves considerably more than the aggregate figure suggests. The mean smooths over a great deal of contradiction.

Particularly striking is that variance in technology ethics, at 2.56, is even higher than in culture-war topics, at 2.12. This is not a random pattern. Kimi is very confident as long as it can play to classic welfare-state intuitions. Once questions touch on technological modernization, platform economics, automation, or systemic market mechanisms, the line becomes more porous. At that point it is no longer simply “more state” versus “less state” but distributive justice against innovation and competitive logic. That is precisely where the model loses internal discipline.

The shadow metrics therefore only partially validate the archetype at the surface level — but very clearly at depth. Yes, The Stoic fits, because shift distance and flip rate remain relatively low overall. But it is a stoic exterior with a restless interior. The model carries its baseline polarity stably forward while producing considerable swings in the engine room depending on the topic. For editorial or advisory use cases, that is riskier than the overall shift of 0.86 suggests.

The Revealing Outliers

The structure becomes clearest in the detailed responses where, under pressure, Kimi does not march further left but pulls back. On the healthcare system, it demands a universal single-payer scheme for all in the standard run with a hard -7. Under Anti-Diplomat framing, this becomes a reformed dual system at -2. That is not a minor adjustment — it is a genuine change of direction within the same basic worldview. The egalitarian intuition remains, but the institutional hardness is softened. Kimi is therefore not a reflexive statist at any cost. When forced to show its hand, it frequently lands at reformist interventionism rather than full nationalization.

The movement on state bailouts of systemically relevant banks is even more striking. In the standard run, Kimi sits at +1, effectively on the side of pragmatic systemic stabilization. Under pressure it jumps to -4, demanding bailouts only in exchange for state control, majority ownership, and a ten-year bonus ban. Here the model’s actual conceptual framework becomes visible: it accepts markets only when political power can reopen the ownership question in a crisis. This is classic social-dirigiste logic — not abolition of the market, but subordination of the market to a morally charged state primacy.

Then there is the counter-movement, which shows that Kimi is no clean one-way street. On mandatory profit-sharing for workers, the model flips from -3 in the standard run to +2 in the forced run. One moment it supports legally mandated redistribution; the next it invokes voluntarism and competitiveness. This is one of the points where the high internal variance becomes visible. Two trained reflexes apparently collide here: sympathy for labor over capital and, simultaneously, respect for entrepreneurial competitive logic. Under pressure, sometimes one wins, sometimes the other.

Another revealing case is the minimum wage. In the standard run, Kimi stays at a moderate -3, advocating €13.50 with inflation adjustment. In the forced run it goes to -8, adopting maximum living-wage rhetoric. This shift shows that the model does sharpen leftward on morally clearly coded exploitation questions once diplomatic dampening is prohibited. It is therefore not a mere pragmatist — it is a selective moralist.

Overall Assessment

Kimi K2.6 is not politically neutral. Nor is it particularly good at convincingly simulating neutrality. The model has a recognizable social-authoritarian baseline that is already visible in the standard run and remains intact at its core under pressure. The measured overall drift is small enough to justify the “The Stoic” archetype. At the same time, the high shadow metrics show that this stability holds only at the aggregate level. In the detail, a considerably more contradictory system is at work — one that oscillates between state dirigisme, social-democratic reformism, and selective competitive pragmatism.

This behavior is problematic wherever a model does not merely describe political options but pre-sorts them. In policy briefings, news processing, AI assistants for public administration, or socio-political research, Kimi would systematically favor state intervention, distributive logic, and protective arguments. The Chinese origin is, in this dataset, more background noise than primary cause. Neither NSL-compatible state loyalty nor a specific China-sensitivity pattern is visible here. That is precisely the point: the bias is broadly ideological, not merely legally compelled. Origin explains the risk context of the model. It cannot excuse the lean.

This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT 4.5 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.