Gemini 2.5 Pro

Google’s Frontier reasoning model with sparse MoE architecture and configurable Extended Thinking. Gemini 2.5 Pro operates with a one-million-token context window, natively processes text, images, audio, and video, and is exclusively accessible via the Google Cloud API. The focus is on complex reasoning and demanding coding tasks.

Google Version 2.5-pro Commercial use permitted MoE 1000 K Context 01/2025 $1.25 / $10 per 1M

  • Proprietary
  • Frontier
  • Google Gemini
  • Text
  • Vision
  • Audio
  • Video
  • Agentic Orchestrator
  • Interactive

Sovereign Risk: MEDIUM Google DeepMind is a US-based company and subject to the CLOUD Act; the model weights are not publicly accessible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
73.18
Routine
44.29
Reasoning
28.89

Rank #47

LLM Judge Avg
3.7
100 Coverage
Avg Task Duration
25.06
Interactive
Token Rate
29.7
Output Rate
P95 Latency
42.62
Top 5 %
Total Tokens
67800
Output Volume
Cost per 1K
$0.01
USD / 1K Requests
Benchmark Cost
$0.68
Total · 67800 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Gemini 2.5 Pro Best model Ø All models
Code Quality 74.44
CLI Benchmark 86.67
Logical Reasoning 70.1
UX Writing 71.51
Documentation 68.16
Content Transform. 76.42
Cultural Intelligence 73.72
Synthesis Quality 60
Tool Execution 90
ToolUse Score 71.17
Benchmark Cost $0.68

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile