MiniMax M3

MiniMax M3 is a multimodal MoE model with a context window of one million tokens, focused on agentic workflows, coding, and tool use. Of 428 billion total parameters, only 23 billion are active per token; the model processes text, image, and video as input. Its Chinese origin requires a separate data privacy risk assessment when used via cloud.

MiniMax Version m3 Commercial use permitted MoE 428 B (23 B active) 1000 K Context 05/2026 $0.3 / $1.2 per 1M

  • Open Weights
  • Server
  • OpenRouter
  • Text
  • Vision
  • Video
  • Interactive

Sovereign Risk: HIGH MiniMax is a Chinese company and subject to China’s National Security Law (NSL), which may enable state access to data. The model has been released as open weights, but remains high-risk from a sovereignty perspective when data or workflows are processed under Chinese jurisdiction.

Key metrics

Score · Latency · Cost · Quality

Total Score Gold
80.19
Routine
49.11
Reasoning
31.08

Rank #2

LLM Judge Avg
4.03
100 Coverage
Avg Task Duration
25.33
Interactive
Token Rate
107.86
Output Rate
P95 Latency
70.76
Top 5 %
Total Tokens
129699.99999999999
Output Volume
Cost per 1K
$0.0012
USD / 1K Requests
Benchmark Cost
$0.16
Total · 129699.99999999999 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

MiniMax M3 Best model Ø All models
Code Quality 81.92
CLI Benchmark 95.33
Logical Reasoning 74.23
UX Writing 76.59
Documentation 77.28
Content Transform. 84.52
Cultural Intelligence 78.12
Synthesis Quality 62.5
Tool Execution 90
ToolUse Score 81.08
Benchmark Cost $0.16

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile