Kimi K2.5
Kimi K2.5 is Moonshot AI’s flagship model featuring active chain-of-thought reasoning, multimodal input for text and images, and a focus on reasoning and agentic tasks. The MoE architecture activates 32 billion of the total one trillion parameters per token; the context window spans 128,000 tokens. Available as an Open Weights variant locally or via cloud, with Chinese jurisdiction as a cloud risk factor.
- Open Weights
- Frontier
- OpenRouter
- Text
- Vision
- Agentic Orchestrator
- Batch
Sovereign Risk: HIGH Moonshot AI is a Chinese company and subject to China’s National Security Law (NSL), which may enable state access to data. In February 2025, the BSI explicitly warned against the use of Chinese AI cloud services; this risk assessment conservatively applies here as well.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 76.62
- Routine
- 46.21
- Reasoning
- 30.4
- LLM Judge Avg
- 3.86 / 5
- 100 Coverage
- Avg Task Duration
- 99.6s
- Batch
- Token Rate
- 37.68tok/s
- Output Rate
- P95 Latency
- 246.92s
- Top 5 %
- Total Tokens
- 196100
- Output Volume
- Cost per 1K
- $0.002
- USD / 1K Requests
- Benchmark Cost
- $0.39
- Total · 196100 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median