o3-mini

o3-mini is OpenAI’s compact reasoning model with internal chain-of-thought, specialized in math, coding, and STEM tasks. The model operates with a context window of 200,000 tokens and offers three adjustable reasoning levels for balancing response depth, latency, and cost. Available exclusively via the OpenAI API.

OpenAI Version 2025-01-31 Commercial use permitted Dense 200 K Context 10/2023 $1.1 / $4.4 per 1M

  • Proprietary
  • Frontier
  • API
  • Text
  • Instruction-Tuned
  • Real-Time

Sovereign Risk: MEDIUM OpenAI is a US-based company and subject to the CLOUD Act. Data transmitted via the API may be made accessible to US authorities. Local deployment is not possible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
68.04
Routine
42.37
Reasoning
25.67

Rank #78

LLM Judge Avg
3.33
100 Coverage
Avg Task Duration
11.74
Real-Time
Token Rate
71.19
Output Rate
P95 Latency
24.46
Top 5 %
Total Tokens
83100
Output Volume
Cost per 1K
$0.0044
USD / 1K Requests
Benchmark Cost
$0.37
Total · 83100 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

o3-mini Best model Ø All models
Code Quality 64.9
CLI Benchmark 87.78
Logical Reasoning 52.86
UX Writing 64.15
Documentation 60.41
Content Transform. 73
Cultural Intelligence 78.3
Synthesis Quality 55.83
Tool Execution 90
ToolUse Score 72.79
Benchmark Cost $0.37

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile