Grok 4.6

Grok 4.6 is xAI’s Frontier model from August 12, 2026, designed for coding, long agent sessions, and knowledge work — proprietary, cloud-only, under US jurisdiction (CLOUD Act). The model processes text and images with a context of 500,000 tokens and offers four reasoning levels (low/medium/high/xhigh). An optional Priority Processing Service Tier doubles API costs in exchange for lower latency.

xAI Version 4.6 Commercial use restricted Dense 500 K Context 02/2026 $2 / $6 per 1M

  • Proprietary
  • Frontier
  • xAI
  • Text
  • Vision
  • Batch

Sovereign Risk: MEDIUM The model is developed and hosted by a US-based company. Due to US jurisdiction, it is potentially subject to the CLOUD Act, which represents a moderate risk of data access by US authorities. Since the weights are proprietary and not distributed, there is no additional risk from distribution of the weights themselves.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
74.72
Routine
45.85
Reasoning
28.87

Rank #24

LLM Judge Avg
3.72
100 Coverage
Avg Task Duration
63.69
Batch
Token Rate
13.87
Output Rate
P95 Latency
166.45
Top 5 %
Total Tokens
73900
Output Volume
Cost per 1K
$0.006
USD / 1K Requests
Benchmark Cost
$0.44
Total · 73900 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Grok 4.6 Best model Ø All models
Code Quality 81.12
CLI Benchmark 88.33
Logical Reasoning 64.42
UX Writing 73.11
Documentation 74.14
Content Transform. 69.66
Cultural Intelligence 80.64
Synthesis Quality 54
Tool Execution 90
ToolUse Score 73.17
Benchmark Cost $0.44

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile