GPT 5.6 Luna
GPT-5.6 Luna is the cheapest and fastest tier of OpenAI’s three-tier GPT-5.6 series (Sol, Terra, Luna) for high-volume, latency-sensitive tasks — available since July 30, 2026 at $0.20 / $1.20 per million tokens, roughly 80 percent below Sol. The 1-million-token context variant with 128,000 output tokens delivers frontier-adjacent agentic performance according to OpenAI, but falls off noticeably against its larger siblings on context recall beyond 512,000 tokens.
- Proprietary
- Frontier
- OpenAI
- Text
- Vision
- Real-Time
Sovereign Risk: MEDIUM The model is developed and hosted by a US-based company. Due to US jurisdiction, it is potentially subject to the CLOUD Act, which represents a moderate risk of data access by US authorities. Since the weights are proprietary and not distributed, there is no additional risk from disclosure of the weights themselves.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.64
- Routine
- 45.12
- Reasoning
- 28.51
- LLM Judge Avg
- 3.63 / 5
- 100 Coverage
- Avg Task Duration
- 10.72s
- Real-Time
- Token Rate
- 87.57tok/s
- Output Rate
- P95 Latency
- 30.11s
- Top 5 %
- Total Tokens
- 80500
- Output Volume
- Cost per 1K
- $0.0012
- USD / 1K Requests
- Benchmark Cost
- $0.1
- Total · 80500 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median