o4-mini

o4-mini is OpenAI’s compact reasoning model with native vision input for images, diagrams, and screenshots. The model processes text and image, operates with a context window of 200,000 tokens, and offers three adjustable reasoning levels for balancing response depth and latency. Full tool use including parallel tool calling for lightweight agentic workflows.

OpenAI Version 4-mini Commercial use permitted Dense 200 K Context 06/2024 $1.1 / $4.4 per 1M

  • Proprietary
  • Frontier
  • API
  • Text
  • Vision
  • Instruction-Tuned
  • Agentic Orchestrator
  • Real-Time

Sovereign Risk: MEDIUM OpenAI is a US-based company and subject to the CLOUD Act. Data transmitted via the API may be made accessible to US authorities. Local deployment is not possible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
69.47
Routine
42.98
Reasoning
26.49

Rank #70

LLM Judge Avg
3.49
100 Coverage
Avg Task Duration
12.31
Real-Time
Token Rate
69.73
Output Rate
P95 Latency
23.84
Top 5 %
Total Tokens
90100
Output Volume
Cost per 1K
$0.0044
USD / 1K Requests
Benchmark Cost
$0.4
Total · 90100 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

o4-mini Best model Ø All models
Code Quality 75.3
CLI Benchmark 90.56
Logical Reasoning 56.05
UX Writing 64.35
Documentation 66.19
Content Transform. 72.98
Cultural Intelligence 78.3
Synthesis Quality 40.83
Tool Execution 85
ToolUse Score 62.54
Benchmark Cost $0.4

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile