DeepSeek V4.1 Flash
DeepSeek V4.1 Flash is the compact variant of the DeepSeek V4.1 family, with 552 billion total parameters and 16 billion active parameters per token. The model is released as Open Weights under the MIT license with image and text processing capabilities. Its 1-million-token context window is designed for long agentic workflows, but due to its size requires server infrastructure for local deployment.
- Open Weights
- Server
- OpenRouter
- Text
- Vision
- Agentic Orchestrator
- Long Context
- Batch
Sovereign Risk: HIGH Although DeepSeek V4.1 Flash is released as an Open Weights model under the permissive MIT license, its developer DeepSeek is a China-based company. Chinese jurisdiction carries a high sovereignty and compliance risk, particularly when using cloud services subject to Chinese law.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.55
- Routine
- 44.63
- Reasoning
- 28.92
- LLM Judge Avg
- 3.66 / 5
- 100 Coverage
- Avg Task Duration
- 49.22s
- Batch
- Token Rate
- 112.43tok/s
- Output Rate
- P95 Latency
- 173.19s
- Top 5 %
- Total Tokens
- 199100
- Output Volume
- Cost per 1K
- $0.0006
- USD / 1K Requests
- Benchmark Cost
- $0.12
- Total · 199100 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median