GPT-5.4 Mini
GPT-5.4 Mini is the compact GPT-5.4 variant for fast and cost-efficient everyday tasks. With a context window of 272,000 tokens and multimodal input for text and image, the model targets applications requiring low latency with solid output quality. Available exclusively via the OpenAI API.
- Proprietary
- Frontier
- OpenAI
- Text
- Vision
- Instruction-Tuned
- Real-Time
Sovereign Risk: MEDIUM OpenAI is a US-based company and subject to the CLOUD Act. When using the API, input data leaves the local network — government access to processed data is legally possible.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 70.17
- Routine
- 41.93
- Reasoning
- 28.24
- LLM Judge Avg
- 3.53 / 5
- 100 Coverage
- Avg Task Duration
- 5.3s
- Real-Time
- Token Rate
- 119.7tok/s
- Output Rate
- P95 Latency
- 12.69s
- Top 5 %
- Total Tokens
- 55000
- Output Volume
- Cost per 1K
- $0.0045
- USD / 1K Requests
- Benchmark Cost
- $0.25
- Total · 55000 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median