DeepSeek V4 Pro

DeepSeek V4 Pro is the flagship of the V4 line, designed for reasoning, coding, and agentic workflows. The hybrid attention MoE architecture combines 1.6 trillion total parameters with 49 billion active parameters per token and a context window of one million tokens. The model is available as an Open Weights model under the MIT license, though Chinese jurisdiction in cloud deployments requires a separate privacy assessment.

DeepSeek Version 4 Commercial use permitted MoE 1600 B (49 B active) 1000 K Context 05/2025 $0.87 / $1.74 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Agentic Orchestrator
  • Long Context
  • Batch

Sovereign Risk: HIGH DeepSeek is a Chinese company subject to China’s National Security Law (NSL), which may allow state access to data and models. On 04.02.2025, Germany’s BSI explicitly warned against using the DeepSeek cloud service: user data is stored on Chinese servers; use for official or sensitive data is not recommended. This warning applies without restriction to cloud API deployments.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
72.56
Routine
43.86
Reasoning
28.7

Rank #57

LLM Judge Avg
3.55
100 Coverage
Avg Task Duration
52.79
Batch
Token Rate
53.32
Output Rate
P95 Latency
135.13
Top 5 %
Total Tokens
170600
Output Volume
Cost per 1K
$0.0017
USD / 1K Requests
Benchmark Cost
$0.3
Total · 170600 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

DeepSeek V4 Pro Best model Ø All models
Code Quality 74.76
CLI Benchmark 83.67
Logical Reasoning 65.71
UX Writing 71.27
Documentation 63.84
Content Transform. 76.08
Cultural Intelligence 73.84
Synthesis Quality 63.33
Tool Execution 90
ToolUse Score 76.83
Benchmark Cost $0.3

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile