Codestral 25.08

Trained for code tasks, not for general use: Codestral 25.08 is Mistral AI’s specialized developer model from August 2025, optimized for fill-in-the-middle and a broad range of programming languages. With 22 billion parameters, it supports 128,000 tokens of context, runs either locally or via the Mistral API, but is subject to a restrictive license with limited commercial use.

Mistral AI Version 25.08 Commercial use restricted Dense 22 B 128 K Context 07/2025 $0.2 / $0.6 per 1M

  • Restricted Weights
  • Frontier
  • Mistral AI
  • Text
  • Real-Time

Sovereign Risk: LOW Mistral AI is a French company headquartered in the EU and is not subject to any government access obligations regarding model weights such as the US CLOUD Act or China’s NSL.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
65.48
Routine
39.92
Reasoning
25.56

Rank #83

LLM Judge Avg
3.29
100 Coverage
Avg Task Duration
3.8
Real-Time
Token Rate
192.29
Output Rate
P95 Latency
11.64
Top 5 %
Total Tokens
52600
Output Volume
Cost per 1K
$0.0006
USD / 1K Requests
Benchmark Cost
$0.03
Total · 52600 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Codestral 25.08 Best model Ø All models
Code Quality 66.28
CLI Benchmark 84.34
Logical Reasoning 64.11
UX Writing 60.77
Documentation 62.72
Content Transform. 64.2
Cultural Intelligence 62
Synthesis Quality 56.67
Tool Execution 83.33
ToolUse Score 68.83
Benchmark Cost $0.03

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile