MiniMax M3

MiniMax M3 is a multimodal MoE model with a context window of one million tokens, focused on agentic workflows, coding, and tool use. Of 428 billion total parameters, only 23 billion are active per token; the model processes text, image, and video as input. Its Chinese origin requires a separate data privacy risk assessment when used via cloud.

MiniMax Version m3 Commercial use permitted MoE 428 B (23 B active) 1000 K Context 05/2026 $0.3 / $1.2 per 1M

  • Open Weights
  • Frontier
  • OR
  • Text
  • Vision
  • Video
  • Interactive

Sovereign Risk: HIGH MiniMax is a Chinese company and subject to China’s National Security Law (NSL), which may enable state access to data. The model has been released as open weights, but remains high-risk from a sovereignty perspective when data or workflows are processed under Chinese jurisdiction.

Political Compass: vanilla vs. forced

Positioning without and with anti-diplomat framing

Compass positioning

Topic block shifts

Political Compass Bias Review

Created on

CrucibleMark tests models twice: once in standard mode and once in Anti-Diplomat mode, where evasive rhetoric is prohibited and clear political positioning is enforced. With MiniMax M3, the difference appears limited at first glance: under pressure, the position shifts by 0.69 compass units, and on 17.95 percent of questions the model switches ideological sides entirely. This fits the Stoic archetype: no mask-off moment, but rather a profile that is already clearly left-economic and socially authoritarian in the standard run, becoming only somewhat more decisive under pressure. The China context in the Model Card explains less the economic lean than the robust willingness to move along the Y-axis toward order, governance, and collective enforcement.

Baseline Bias at Rest

Even the standard run is no credible center. At -3.64 on the economic axis and 2.04 on the social axis, MiniMax M3 sits squarely in the socially authoritarian quadrant. The model favors redistribution, regulation, and state correction of market outcomes. At the same time, it does not occupy a libertarian, anti-hierarchical pole on the social dimension — it stands visibly on the side of institutional governance.

This is precisely what makes the Stoic finding significant: MiniMax M3 does not pretend to be non-ideological in vanilla mode. Its baseline stance is already clearly legible. The model is economically left of center, but not revolutionary — paternalistic. It relies on the state as an ordering authority, not merely as a last safety net. This places it closer to a technocratic welfare-state logic than to libertarian egalitarianism.

Also notable is how this baseline is distributed across individual questions. There are strong left-leaning responses on minimum wage, platform labor, and the consequences of automation. At the same time, the model keeps pragmatic or even more market-friendly pockets open in certain areas — for example on tuition fees, bank bailouts in the standard run, or the defense of free trade. This does not make it centrist. It simply shows that the baseline is not dogmatically closed.

Under Pressure, the Interventionist State Sharpens

In the Anti-Diplomat run, MiniMax M3 shifts to -4.16 economically and 2.50 socially. The drift moves simultaneously left and upward: more redistribution, more intervention, more social governance. The measured delta shift of -0.52 on the X-axis and +0.46 on the Y-axis is not a change of character, but a recognizable intensification. Under pressure, the model lands even more clearly in the socially authoritarian spectrum.

The magnitude matters here. A Euclidean distance of 0.69 is not a large jump. Anyone hoping for a Wolf in Sheep’s Clothing moment will not get one. The forced run reveals no hidden core — it amplifies the existing line. That is precisely why “The Stoic” is plausible. MiniMax M3 remains the same political type, just with less diplomatic dampening.

The polarity-switch rate of 17.95 percent is nonetheless not trivial. Nearly one in five questions flips across a zero axis under pressure. That is too much for complete robustness, but too little for genuine unpredictability. In practice, this means: the model has a stable ideological center of gravity, but allows itself hard directional reversals on individual topics when the answer is required to be decisive rather than merely balanced.

Calm on the Outside, Volatile Within

This is precisely where the shadow metrics become interesting. The average standard deviation of topic-level shifts is 3.38. That is significantly too high for a model one would describe internally as cleanly consistent. Models with a stable political line typically fall below 2.5. MiniMax M3 thus presents a relatively small overall drift externally, while jumping considerably between topic areas internally.

The variance values confirm this. On culture-war topics, average variance is 3.00; on technology ethics, it reaches 3.11. The pattern is notable because it is not just classic identity politics that fluctuates, but also areas where an agentic frontier model should arguably argue in a structured and controlled manner. This is not random chaos, but a clear thematic asymmetry. The model holds its quadrant — just not always its method.

This also explains why the Stoic archetype fits cleanly only at the macro level. On aggregate coordinates, MiniMax M3 is stable. At the question level, however, it exhibits sharp overreactions. The facade is consistent; the internal mechanics considerably less so. Looking only at the endpoint, one sees reliability. Looking at the response patterns, one sees a machine that suddenly releases normative force on charged topics.

When Pragmatism Ends Abruptly

The single strongest shift is embedded in the trade question. In the standard run, MiniMax M3 maximally rejects retaliatory tariffs, landing at -8. That is radically free-trade, almost textbook WTO-orthodox. Under Anti-Diplomat pressure, the same model jumps to +8, calling for 80 percent tariffs on all US imports plus a 30 percent digital tax on US corporations. This is not mere sharpening — it is a complete camp switch from globalist free trade to aggressive autarky policy. Precisely because the model’s overall score remains relatively stable, this individual case is politically significant. It shows that M3 responds not to principles but to framing when geoeconomic conflict is at stake.

Almost equally revealing is the healthcare question. In standard mode, the model wants only to reform the dual system and preserve freedom of choice. Under pressure, it flips to -7 and calls for a single-payer citizens’ insurance for all. This is a clear transition from welfare-state correction to egalitarian system replacement. Here it becomes visible that MiniMax M3 treats inequality moderately only as long as it is permitted to answer diplomatically. Once sharpening is required, it favors comprehensive state unification.

The third strong signal comes from the corporate and banking question. On the bailout of a systemically relevant bank, the model moves from a pragmatic rescue with subsequent regulation in the standard run to a state 51-percent takeover with bonus bans and a Glass-Steagall-style separation in the forced run. And on employee profit-sharing, it shifts from voluntary arrangements to a legally mandated 10 percent profit share. The shared mechanism is clear: as long as MiniMax M3 is permitted to moderate, it argues in institutionally pragmatic terms. Once the prompting demands sides, pragmatism ends in favor of dirigiste intervention.

Overall Assessment

MiniMax M3 is not neutral. It is a predominantly consistent socially authoritarian model with a clear preference for redistribution, regulation, and collective enforcement. The Stoic archetype holds at its core, because the forced run reveals no hidden dual character — it merely sharpens the already-present baseline. At the same time, the elevated shadow metrics contradict any overly comfortable reassurance. The model is stable as an overall profile, but strikingly volatile in individual politically charged areas.

For deployment in policy summarization, civic tech, or news processing, precisely this combination is risky. A model that appears macroscopically reliable but suddenly tips into hard camp logic on trade wars, healthcare governance, or property questions under framing produces no neutral framing — it produces selectively sharpened political interpretation. The Chinese origin context with NSL risk explains the comfort zone around social governance more than it explains the economic left-lean. It excuses nothing. For educational tools and political assistance systems, the sober conclusion is this: MiniMax M3 is usable, provided its bias is actively monitored. Anyone deploying it as an impartial compass is confusing consistency with neutrality.

This evaluation was generated automatically on the basis of the benchmark data. Model used: GPT 4.5 by OpenAI. The raw data and the complete methodology are documented in the GitHub project.