GPT 5.6 Luna

GPT-5.6 Luna is the cheapest and fastest tier of OpenAI’s three-tier GPT-5.6 series (Sol, Terra, Luna) for high-volume, latency-sensitive tasks — available since July 30, 2026 at $0.20 / $1.20 per million tokens, roughly 80 percent below Sol. The 1-million-token context variant with 128,000 output tokens delivers frontier-adjacent agentic performance according to OpenAI, but falls off noticeably against its larger siblings on context recall beyond 512,000 tokens.

OpenAI Version 5.6 Commercial use permitted Dense 1000 K Context 02/2026 $0.2 / $1.2 per 1M

  • Proprietary
  • Frontier
  • OpenAI
  • Text
  • Vision
  • Real-Time

Sovereign Risk: MEDIUM The model is developed and hosted by a US-based company. Due to US jurisdiction, it is potentially subject to the CLOUD Act, which represents a moderate risk of data access by US authorities. Since the weights are proprietary and not distributed, there is no additional risk from disclosure of the weights themselves.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
73.64
Routine
45.12
Reasoning
28.51

Rank #52

LLM Judge Avg
3.63
100 Coverage
Avg Task Duration
10.72
Real-Time
Token Rate
87.57
Output Rate
P95 Latency
30.11
Top 5 %
Total Tokens
80500
Output Volume
Cost per 1K
$0.0012
USD / 1K Requests
Benchmark Cost
$0.1
Total · 80500 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GPT 5.6 Luna Best model Ø All models
Code Quality 80.52
CLI Benchmark 89
Logical Reasoning 59.79
UX Writing 67.19
Documentation 67.85
Content Transform. 77.17
Cultural Intelligence 75.32
Synthesis Quality 75
Tool Execution 87.5
ToolUse Score 79.96
Benchmark Cost $0.1

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile