Kimi K2.5
Kimi K2.5 is Moonshot AI’s flagship model featuring active chain-of-thought reasoning, multimodal input for text and images, and a focus on reasoning and agentic tasks. The MoE architecture activates 32 billion of the total one trillion parameters per token; the context window spans 128,000 tokens. Available as an Open Weights variant locally or via cloud, with Chinese jurisdiction as a cloud risk factor.
- Open Weights
- Frontier
- OpenRouter
- Text
- Vision
- Agentic Orchestrator
- Batch
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 76.62
- Routine
- 46.21
- Reasoning
- 30.4
- LLM Judge Avg
- 3.86 / 5
- 100 Coverage
- Avg Task Duration
- 101.16s
- Batch
- Token Rate
- 38.45tok/s
- Output Rate
- P95 Latency
- 247.19s
- Top 5 %
- Total Tokens
- 196100
- Output Volume
- Cost per 1K
- $0.002
- USD / 1K Requests
- Benchmark Cost
- $0.39
- Total · 196100 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median