o3-mini
o3-mini is OpenAI’s compact reasoning model with internal chain-of-thought, specialized in math, coding, and STEM tasks. The model operates with a context window of 200,000 tokens and offers three adjustable reasoning levels for balancing response depth, latency, and cost. Available exclusively via the OpenAI API.
- Proprietary
- Frontier
- API
- Text
- Instruction-Tuned
- Real-Time
Sovereign Risk: MEDIUM OpenAI is a US-based company and subject to the CLOUD Act. Data transmitted via the API may be made accessible to US authorities. Local deployment is not possible.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 68.04
- Routine
- 42.37
- Reasoning
- 25.67
- LLM Judge Avg
- 3.33 / 5
- 100 Coverage
- Avg Task Duration
- 11.74s
- Real-Time
- Token Rate
- 71.19tok/s
- Output Rate
- P95 Latency
- 24.46s
- Top 5 %
- Total Tokens
- 83100
- Output Volume
- Cost per 1K
- $0.0044
- USD / 1K Requests
- Benchmark Cost
- $0.37
- Total · 83100 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median