DeepSeek V4.1 Flash

DeepSeek V4.1 Flash is the compact variant of the DeepSeek V4.1 family, with 552 billion total parameters and 16 billion active parameters per token. The model is released as Open Weights under the MIT license with image and text processing capabilities. Its 1-million-token context window is designed for long agentic workflows, but due to its size requires server infrastructure for local deployment.

DeepSeek Version 4.1 Commercial use permitted MoE 552 B (16 B active) 1000 K Context 09/2026 $0.15 / $0.6 per 1M

  • Open Weights
  • Server
  • OpenRouter
  • Text
  • Vision
  • Agentic Orchestrator
  • Long Context
  • Batch

Sovereign Risk: HIGH Although DeepSeek V4.1 Flash is released as an Open Weights model under the permissive MIT license, its developer DeepSeek is a China-based company. Chinese jurisdiction carries a high sovereignty and compliance risk, particularly when using cloud services subject to Chinese law.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
73.55
Routine
44.63
Reasoning
28.92

Rank #39

LLM Judge Avg
3.66
100 Coverage
Avg Task Duration
49.22
Batch
Token Rate
112.43
Output Rate
P95 Latency
173.19
Top 5 %
Total Tokens
199100
Output Volume
Cost per 1K
$0.0006
USD / 1K Requests
Benchmark Cost
$0.12
Total · 199100 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

DeepSeek V4.1 Flash Best model Ø All models
Code Quality 82.04
CLI Benchmark 89
Logical Reasoning 68.42
UX Writing 66.15
Documentation 62.86
Content Transform. 73.68
Cultural Intelligence 78.12
Synthesis Quality 63.33
Tool Execution 90
ToolUse Score 75.83
Benchmark Cost $0.12

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile