GPT-5.4 Mini

GPT-5.4 Mini is the compact GPT-5.4 variant for fast and cost-efficient everyday tasks. With a context window of 272,000 tokens and multimodal input for text and image, the model targets applications requiring low latency with solid output quality. Available exclusively via the OpenAI API.

OpenAI Version 5.4 Commercial use permitted Dense 272 K Context 09/2025 $0.75 / $4.5 per 1M

  • Proprietary
  • Frontier
  • OpenAI
  • Text
  • Vision
  • Instruction-Tuned
  • Real-Time

Sovereign Risk: MEDIUM OpenAI is a US-based company and subject to the CLOUD Act. When using the API, input data leaves the local network — government access to processed data is legally possible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
70.17
Routine
41.93
Reasoning
28.24

Rank #64

LLM Judge Avg
3.53
100 Coverage
Avg Task Duration
5.3
Real-Time
Token Rate
119.7
Output Rate
P95 Latency
12.69
Top 5 %
Total Tokens
55000
Output Volume
Cost per 1K
$0.0045
USD / 1K Requests
Benchmark Cost
$0.25
Total · 55000 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GPT-5.4 Mini Best model Ø All models
Code Quality 74.4
CLI Benchmark 81
Logical Reasoning 65.74
UX Writing 63.53
Documentation 61.46
Content Transform. 77.7
Cultural Intelligence 75.32
Synthesis Quality 55.83
Tool Execution 83.33
ToolUse Score 67.62
Benchmark Cost $0.25

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile