Kimi K2.5

Kimi K2.5 is Moonshot AI’s flagship model featuring active chain-of-thought reasoning, multimodal input for text and images, and a focus on reasoning and agentic tasks. The MoE architecture activates 32 billion of the total one trillion parameters per token; the context window spans 128,000 tokens. Available as an Open Weights variant locally or via cloud, with Chinese jurisdiction as a cloud risk factor.

Moonshot AI Version k2.5 Commercial use permitted MoE 1000 B (32 B active) 128 K Context 09/2025 $0.44 / $2 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Vision
  • Agentic Orchestrator
  • Batch

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
76.62
Routine
46.21
Reasoning
30.4

Rank #11

LLM Judge Avg
3.86
100 Coverage
Avg Task Duration
101.16
Batch
Token Rate
38.45
Output Rate
P95 Latency
247.19
Top 5 %
Total Tokens
196100
Output Volume
Cost per 1K
$0.002
USD / 1K Requests
Benchmark Cost
$0.39
Total · 196100 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Kimi K2.5 Best model Ø All models
Code Quality 77.76
CLI Benchmark 88.33
Logical Reasoning 75.28
UX Writing 71.73
Documentation 79.4
Content Transform. 70.92
Cultural Intelligence 78.36
Synthesis Quality 65.83
Tool Execution 90
ToolUse Score 77
Benchmark Cost $0.39

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile