Kimi K2.5

Kimi K2.5 ist Moonshot AIs Flaggschiff mit aktivem Chain-of-Thought-Reasoning, multimodalem Eingang für Text und Bild sowie Fokus auf Reasoning und agentische Aufgaben. Die MoE-Architektur aktiviert pro Token nur 32 Milliarden der insgesamt eine Billion Gesamtparameter, das Kontextfenster umfasst 128.000 Tokens. Als Open-Weights-Variante lokal oder über die Cloud verfügbar, mit chinesischer Jurisdiktion als Cloud-Risikofaktor.

Moonshot AI Version k2.5 Kommerzielle Nutzung erlaubt MoE 1000 B (32 B aktiv) 128 K Context 09/2025 $0.44 / $2 per 1M

  • Open Weights
  • Frontier
  • OR
  • Text
  • Vision
  • Agentic Orchestrator
  • Batch

Schlüsselmetriken

Score · Latenz · Kosten · Qualität

Total Score Silver
73.32
Routine
44.2
Reasoning
29.13

Rank #39

LLM Judge Avg
3.67
100 Coverage
Avg Task Duration
65.12
Batch
Token Rate
21.07
Output Rate
P95 Latency
185.38
Top 5 %
Total Tokens
165000
Output Volume
Cost per 1K
$0.002
USD / 1K Requests
Benchmark Cost
$0.33
Total · 165000 tok

Benchmark-Module

10 Module · gewichtet · vs. Modellmedian & Spitzenreiter

Kimi K2.5 Bestes Modell Ø Alle Modelle
Code Quality 70.6
CLI Benchmark 86.67
Logical Reasoning 74.28
UX Writing 68.69
Documentation 73.09
Content Transform. 78.99
Cultural Intelligence 64.64
Synthesis Quality 65.83
Tool Execution 90
ToolUse Score 77.42
Benchmark Cost $0.33

Token-Effizienz & Latenz

Verbrauch pro Modul vs. Modellmedian

Token-Verbrauch pro Modul

Performance-Profil