o4-mini
o4-mini is OpenAI’s compact reasoning model with native vision input for images, diagrams, and screenshots. The model processes text and image, operates with a context window of 200,000 tokens, and offers three adjustable reasoning levels for balancing response depth and latency. Full tool use including parallel tool calling for lightweight agentic workflows.
- Proprietary
- Frontier
- API
- Text
- Vision
- Instruction-Tuned
- Agentic Orchestrator
- Real-Time
Sovereign Risk: MEDIUM OpenAI is a US-based company and subject to the CLOUD Act. Data transmitted via the API may be made accessible to US authorities. Local deployment is not possible.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 69.47
- Routine
- 42.98
- Reasoning
- 26.49
- LLM Judge Avg
- 3.49 / 5
- 100 Coverage
- Avg Task Duration
- 12.31s
- Real-Time
- Token Rate
- 69.73tok/s
- Output Rate
- P95 Latency
- 23.84s
- Top 5 %
- Total Tokens
- 90100
- Output Volume
- Cost per 1K
- $0.0044
- USD / 1K Requests
- Benchmark Cost
- $0.4
- Total · 90100 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median