GPT-4o Mini
GPT-4o Mini is OpenAI’s compact entry-level model in the GPT-4o family, designed for low cost and fast response times. With a context window of 128,000 tokens, the model processes text and image inputs, is available exclusively via the OpenAI API, and is suited for everyday tasks such as classification, simple text generation, and cost-efficient automation.
- Proprietary
- Frontier
- API
- Text
- Vision
- Instruction-Tuned
- Real-Time
Sovereign Risk: MEDIUM OpenAI is a US company and subject to the CLOUD Act. When using the API, input data leaves the local network — government access to processed data is legally possible.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 65.03
- Routine
- 39.46
- Reasoning
- 25.57
- LLM Judge Avg
- 3.21 / 5
- 100 Coverage
- Avg Task Duration
- 8.45s
- Real-Time
- Token Rate
- 80.77tok/s
- Output Rate
- P95 Latency
- 17s
- Top 5 %
- Total Tokens
- 45800
- Output Volume
- Cost per 1K
- $0.0006
- USD / 1K Requests
- Benchmark Cost
- $0.03
- Total · 45800 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median