Gemini 2.5 Pro
Google’s Frontier reasoning model with sparse MoE architecture and configurable Extended Thinking. Gemini 2.5 Pro operates with a one-million-token context window, natively processes text, images, audio, and video, and is exclusively accessible via the Google Cloud API. The focus is on complex reasoning and demanding coding tasks.
- Proprietary
- Frontier
- Google Gemini
- Text
- Vision
- Audio
- Video
- Agentic Orchestrator
- Interactive
Sovereign Risk: MEDIUM Google DeepMind is a US-based company and subject to the CLOUD Act; the model weights are not publicly accessible.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.18
- Routine
- 44.29
- Reasoning
- 28.89
- LLM Judge Avg
- 3.7 / 5
- 100 Coverage
- Avg Task Duration
- 25.06s
- Interactive
- Token Rate
- 29.7tok/s
- Output Rate
- P95 Latency
- 42.62s
- Top 5 %
- Total Tokens
- 67800
- Output Volume
- Cost per 1K
- $0.01
- USD / 1K Requests
- Benchmark Cost
- $0.68
- Total · 67800 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median