Grok 4.5

Grok 4.5 has been xAI’s current flagship since July 2026, with native real-time access to web and X data, a 500,000-token context window, and multimodal text and image input. Reasoning runs server-side at all times and is controlled via configurable reasoning effort levels, with no visible chain-of-thought tags in the response text. Function calling, structured outputs, and prompt caching round out the profile for coding and agentic workflows.

xAI Version 4.5 Commercial use restricted Dense 500 K Context 02/2026 $2 / $6 per 1M

  • Proprietary
  • Frontier
  • xAI
  • Text
  • Vision
  • Interactive

Sovereign Risk: MEDIUM xAI is a US company subject to the CLOUD Act. The operational risk lies not in the distribution of weights (proprietary, not public) but in the use of the cloud API: data is processed server-side, with a default retention period of 30 days unless a zero-data-retention option is activated for enterprise accounts.[web:590]

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
73.85
Routine
45.27
Reasoning
28.58

Rank #34

LLM Judge Avg
3.66
100 Coverage
Avg Task Duration
21.9
Interactive
Token Rate
25.65
Output Rate
P95 Latency
53.37
Top 5 %
Total Tokens
86600
Output Volume
Cost per 1K
$0.006
USD / 1K Requests
Benchmark Cost
$0.52
Total · 86600 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Grok 4.5 Best model Ø All models
Code Quality 77.24
CLI Benchmark 90.67
Logical Reasoning 67.65
UX Writing 64.43
Documentation 79.15
Content Transform. 72.43
Cultural Intelligence 70.64
Synthesis Quality 60
Tool Execution 90
ToolUse Score 77
Benchmark Cost $0.52

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile