Codestral 25.08
Trained for code tasks, not for general use: Codestral 25.08 is Mistral AI’s specialized developer model from August 2025, optimized for fill-in-the-middle and a broad range of programming languages. With 22 billion parameters, it supports 128,000 tokens of context, runs either locally or via the Mistral API, but is subject to a restrictive license with limited commercial use.
- Restricted Weights
- Frontier
- Mistral AI
- Text
- Real-Time
Sovereign Risk: LOW Mistral AI is a French company headquartered in the EU and is not subject to any government access obligations regarding model weights such as the US CLOUD Act or China’s NSL.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 65.48
- Routine
- 39.92
- Reasoning
- 25.56
- LLM Judge Avg
- 3.29 / 5
- 100 Coverage
- Avg Task Duration
- 3.8s
- Real-Time
- Token Rate
- 192.29tok/s
- Output Rate
- P95 Latency
- 11.64s
- Top 5 %
- Total Tokens
- 52600
- Output Volume
- Cost per 1K
- $0.0006
- USD / 1K Requests
- Benchmark Cost
- $0.03
- Total · 52600 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median