GPT-4o Mini

GPT-4o Mini is OpenAI’s compact entry-level model in the GPT-4o family, designed for low cost and fast response times. With a context window of 128,000 tokens, the model processes text and image inputs, is available exclusively via the OpenAI API, and is suited for everyday tasks such as classification, simple text generation, and cost-efficient automation.

OpenAI Version 2024-07-18 Commercial use permitted Dense 128 K Context 10/2023 $0.15 / $0.6 per 1M

  • Proprietary
  • Frontier
  • API
  • Text
  • Vision
  • Instruction-Tuned
  • Real-Time

Sovereign Risk: MEDIUM OpenAI is a US company and subject to the CLOUD Act. When using the API, input data leaves the local network — government access to processed data is legally possible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
65.03
Routine
39.46
Reasoning
25.57

Rank #88

LLM Judge Avg
3.21
100 Coverage
Avg Task Duration
8.45
Real-Time
Token Rate
80.77
Output Rate
P95 Latency
17
Top 5 %
Total Tokens
45800
Output Volume
Cost per 1K
$0.0006
USD / 1K Requests
Benchmark Cost
$0.03
Total · 45800 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GPT-4o Mini Best model Ø All models
Code Quality 55.6
CLI Benchmark 83.89
Logical Reasoning 62.68
UX Writing 63.95
Documentation 53.48
Content Transform. 72.82
Cultural Intelligence 71.04
Synthesis Quality 48.33
Tool Execution 83.33
ToolUse Score 66.21
Benchmark Cost $0.03

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile