GPT-5.4
GPT-5.4 is the larger variant from OpenAI’s 5.4 generation for demanding workloads, delivering higher response quality than the Mini versions. The model operates with a context window of 272,000 tokens, processes text and image inputs, and is available exclusively via the OpenAI API. Proprietary and designed for productive, complex tasks.
- Proprietary
- Frontier
- OpenAI
- Text
- Vision
- Instruction-Tuned
- Real-Time
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 72.93
- Routine
- 44.57
- Reasoning
- 28.36
- LLM Judge Avg
- 3.76 / 5
- 100 Coverage
- Avg Task Duration
- 9.52s
- Real-Time
- Token Rate
- 63.63tok/s
- Output Rate
- P95 Latency
- 30.27s
- Top 5 %
- Total Tokens
- 62900
- Output Volume
- Cost per 1K
- $0.015
- USD / 1K Requests
- Benchmark Cost
- $0.94
- Total · 62900 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median