Claude Opus 4.6

Anthropic’s frontier model for complex agent tasks: Claude Opus 4.6 processes text and image inputs with a standard context window of 200,000 tokens, expandable to one million tokens for long workflows. The model supports tool calls and Extended Thinking for maximum reasoning depth.

Anthropic Version 4.6 Commercial use permitted Dense 1000 K Context 01/2025 $5 / $25 per 1M

  • Proprietary
  • Frontier
  • API
  • Text
  • Vision
  • Agentic Orchestrator
  • Long Context
  • Interactive

Sovereign Risk: MEDIUM Anthropic is a US-based company and subject to the CLOUD Act; model weights are not publicly accessible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
76.1
Routine
45.66
Reasoning
30.43

Rank #12

LLM Judge Avg
3.87
100 Coverage
Avg Task Duration
33.97
Interactive
Token Rate
44
Output Rate
P95 Latency
101.73
Top 5 %
Total Tokens
111000
Output Volume
Cost per 1K
$0.025
USD / 1K Requests
Benchmark Cost
$2.78
Total · 111000 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Claude Opus 4.6 Best model Ø All models
Code Quality 81.76
CLI Benchmark 81.87
Logical Reasoning 76.66
UX Writing 75.83
Documentation 82.03
Content Transform. 72.28
Cultural Intelligence 75.32
Synthesis Quality 51.67
Tool Execution 83.33
ToolUse Score 65.92
Benchmark Cost $2.78

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile