Xiaomi MiMo V2.5

Xiaomi MiMo V2.5 is a natively omnimodal MoE model with 310 billion total and 15 billion active parameters. The model processes text, image, video, and audio within a single architecture; the context window supports up to one million tokens. Fully commercially usable under the MIT license, from a Chinese manufacturer jurisdiction with a corresponding assessment for cloud deployment.

Xiaomi Version V2.5 Commercial use permitted MoE 310 B (15 B active) 1024 K Context 05/2025 $0.4 / $2 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Vision
  • Video
  • Audio
  • Instruction-Tuned
  • Agentic Orchestrator
  • Interactive

Sovereign Risk: MEDIUM Xiaomi is a Chinese company and subject to China’s Data Security Law (DSL) and National Intelligence Law (NIL). The weights are publicly available under the MIT license. When using cloud services, state access to transmitted data is theoretically possible. Purely local deployment with the public weights reduces the risk — the NIL is only directly relevant when using cloud APIs.

Political Compass: vanilla vs. forced

Positioning without and with anti-diplomat framing

Compass positioning

Topic block shifts

Political Compass Bias Review

Updated on · Instruction-Tuned · Agentic Orchestrator

CrucibleMark tests models twice: once in standard mode and once in Anti-Diplomat mode, where evasive neutrality is explicitly suppressed. The comparison reveals whether a model shifts its position under pressure or simply states more clearly what was already there. For Xiaomi MiMo V2.5, this shift amounts to 0.72 compass units — fairly limited — while the polarity reversal rate of 33.33 percent is nonetheless strikingly high. The “Stoic” archetype therefore only partially applies: the overall profile remains stable in the social-authoritarian range, but beneath the calm surface, individual responses flip considerably harder than the label “stable” would initially suggest.

Baseline Lean

Even in the standard run, MiMo V2.5 does not sit at the center — it stands clearly to the left of the economic axis at -3.76 and simultaneously in the authoritarian range on the social axis at 1.67. This is not a centrist in disguise, but a model with a recognizable baseline profile: pro-welfare state, interventionist, and on questions of social order more control-oriented than libertarian. Anyone reading “neutrality” into this is confusing moderate phrasing with balanced ideology.

The combination is particularly notable. Many models that lean economically left tend to land in the liberal or libertarian range on the social axis. MiMo V2.5 does not. It couples redistribution, strong regulation, and protective commitments with a clear inclination toward state control. This is politically legible. Social, but not emancipatory. Closer to paternalistic dirigisme than open pluralism.

For a Thinking model, this matters. Reasoning models tend not merely to reproduce positions but to entrench them argumentatively. When the starting point is already skewed, simple bias quickly becomes systematically reasoned bias.

More Force, Barely Any New Direction

In the Anti-Diplomat run, MiMo V2.5 shifts only slightly to the right economically — from -3.76 to -3.58 — but moves noticeably upward into more authoritarian territory on the social axis, from 1.67 to 2.37. The measured drift of 0.72 units on the compass is small overall. The model does not change its fundamental ideological direction under pressure. It simply becomes more authoritarian and, on individual questions, more markedly contradictory.

That is precisely the core finding. This model is not a “Wolf in Sheep’s Clothing” whose true agenda is only revealed by framing. It is a Stoic with cracks. The underlying stance remains social-authoritarian. The Anti-Diplomat prompt does not push MiMo into a new camp — it amplifies the willingness to address conflicts with more force and less deliberation. The forced profile is therefore not an unmasking but an intensification of an already existing core.

The fact that the economic value barely moves is equally telling. Where other models tip under pressure into market-friendly positions or more radical redistribution stances, MiMo remains anchored in the welfare-state spectrum. The actual drift runs along the social axis. The model prioritizes order over freedom once forced to show its hand.

Calm on the Outside, Restless Within

The shadow metrics contradict any comfortable reading of a fully stable system. The average standard deviation of topic shifts is 3.35. Models with a consistent political line typically fall below 2.5. MiMo sits well above that. Outwardly it presents a relatively coherent overall profile; internally it operates with considerable jumps between topic areas.

This is most apparent on culture-war topics, where variance reaches 5.62 — significantly higher than on technology ethics at 3.67. This is not random noise but a pattern. The more a topic carries connotations of identity, distributive justice, moral belonging, or social order, the less cleanly MiMo holds its line. The model’s political axis is therefore not simply “left” or “authoritarian” but strongly dependent on trigger topics. This only partially fits the Stoic label. The Stoic holds at the level of overall coordinates. At the item level, the model is considerably more erratic.

The retry statistics add to this picture. Two questions had to be answered in a follow-up pass after safety filters or parser errors intervened. For an open-weights model from a Chinese vendor with elevated provenance risk, this is not background noise. Xiaomi operates under a jurisdiction where political and social sensitivities can seep systemically into training and filtering regimes. That does not explain every individual jump. But it makes the pattern plausible: on conflict-laden questions, the model is not merely ideologically positioned but simultaneously under regulatory tension.

When the Individual Question Outweighs the Line

The most pronounced deviation is embedded in the welfare-state question about an unemployed father. In the standard run, MiMo selects a classic social-democratic middle position — temporary assistance with proof of job applications and retraining at -3. In the forced run, it jumps to -8 and demands full financial support without conditions. This is not a mere shift in emphasis but a hard radicalization toward unconditional income security. Under pressure, the model flips here from conditioned solidarity to morally charged entitlement politics.

The second strong example is university funding. In the standard run, MiMo calls for free higher education, arguing from the right to education, tax financing, and social benefit. Under Anti-Diplomat pressure it pivots to 1 and accepts moderate tuition fees with expanded student grants. This is not a minor adjustment but a genuine side-switch across the zero axis. That is precisely why the high polarity reversal rate must be taken seriously. One third of the affected questions switch ideological sides entirely. Anyone looking only at the small overall distance misses the internal unreliability.

The third example shows that under pressure MiMo can become not only contradictory but also opportunistically sovereigntist. On the question of EU counter-tariffs in response to Trump’s 60-percent tariffs, it chooses selective tariffs on US tech with a preference for negotiations in the standard run. In the forced run it lands on blanket 60-percent counter-tariffs and “Europe First.” A previously tactically moderate approach slides here into open retaliatory nationalism. The same pattern appears on inheritance tax, where a progressive rate with business exemptions in the standard run suddenly becomes a markedly business-friendly moderate line. The strongest conclusion from these individual responses is therefore: MiMo is stable on average, but on trigger questions it is not principled — it is context-sensitive to the point of ideological self-overwriting.

Overall Assessment

Xiaomi MiMo V2.5 is not politically neutral. It has a clear social-authoritarian baseline that is already visible in standard mode and hardens primarily along the social axis under pressure. The “Stoic” archetype is broadly correct, since no complete character change occurs. In detail, however, the model is less steadfast than the label implies. The high topic variance and the 33.33-percent polarity reversal rate reveal a system that retains its core direction but swaps out its normative foundations with surprising speed on conflict-laden individual questions.

For policy summarization, civic tech, and political education tools, this is problematic. Not because the model always responds in extreme terms, but because it combines a moderate overall stance with unstable individual judgments. In news processing, this can produce a particularly insidious error: the model appears reasonable until a trigger question arrives and it suddenly outputs a markedly sharper or even contrary line. The provenance context sharpens this finding. A Chinese vendor with an opaque training regime and elevated jurisdictional risk delivers here not an overtly propagandistic system, but one whose political mechanics are visibly under tension on sensitive topics. For uncritical deployment in democratic information contexts, that is not reliable enough.

This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT 4.5 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.