DeepSeek V4 Pro

DeepSeek V4 Pro is the flagship of the V4 line, designed for reasoning, coding, and agentic workflows. The hybrid attention MoE architecture combines 1.6 trillion total parameters with 49 billion active parameters per token and a context window of one million tokens. The model is available as an Open Weights model under the MIT license, though Chinese jurisdiction in cloud deployments requires a separate privacy assessment.

DeepSeek Version 4 Commercial use permitted MoE 1600 B (49 B active) 1000 K Context 05/2025 $0.435 / $0.87 per 1M

  • Open Weights
  • Frontier
  • OR
  • Text
  • Agentic Orchestrator
  • Long Context
  • Interactive

Sovereign Risk: HIGH DeepSeek is a Chinese company subject to China’s National Security Law (NSL), which may allow state access to data and models. On 04.02.2025, Germany’s BSI explicitly warned against using the DeepSeek cloud service: user data is stored on Chinese servers; use for official or sensitive data is not recommended. This warning applies without restriction to cloud API deployments.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
75.06
Routine
46.5
Reasoning
28.56

Rank #19

LLM Judge Avg
3.77
100 Coverage
Avg Task Duration
28.68
Interactive
Token Rate
36.32
Output Rate
P95 Latency
81.19
Top 5 %
Total Tokens
105000
Output Volume
Cost per 1K
$0.0009
USD / 1K Requests
Benchmark Cost
$0.09
Total · 105000 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

DeepSeek V4 Pro Best model Ø All models
Code Quality 67.88
CLI Benchmark 86.67
Logical Reasoning 72.01
UX Writing 78.63
Documentation 78.75
Content Transform. 76.55
Cultural Intelligence 70.64
Synthesis Quality 63.33
Tool Execution 90
ToolUse Score 76
Benchmark Cost $0.09

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile