MiniMax M2.7
MiniMax M2.7 is a Chinese Frontier generalist model with a context window of 205,000 tokens for large-scale documents and multilingual applications. The MoE architecture delivers high performance for general language and reasoning tasks; the model is available as a cloud variant and designed for productive applications. When used via cloud, Chinese jurisdiction applies with the corresponding data privacy implications.
- Restricted Weights
- Frontier
- OpenRouter
- Text
- Interactive
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.03
- Routine
- 44.67
- Reasoning
- 28.37
- LLM Judge Avg
- 3.71 / 5
- 100 Coverage
- Avg Task Duration
- 34.17s
- Interactive
- Token Rate
- 49.03tok/s
- Output Rate
- P95 Latency
- 89.19s
- Top 5 %
- Total Tokens
- 109900
- Output Volume
- Cost per 1K
- $0.0012
- USD / 1K Requests
- Benchmark Cost
- $0.13
- Total · 109900 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median