Gemini 2.5 Pro

Google’s Frontier reasoning model with sparse MoE architecture and configurable Extended Thinking. Gemini 2.5 Pro operates with a one-million-token context window, natively processes text, images, audio, and video, and is exclusively accessible via the Google Cloud API. The focus is on complex reasoning and demanding coding tasks.

Google Version 2.5-pro Commercial use permitted MoE 1000 K Context 01/2025 $1.25 / $10 per 1M

  • Proprietary
  • Frontier
  • API
  • Text
  • Vision
  • Audio
  • Video
  • Agentic Orchestrator
  • Interactive

Sovereign Risk: MEDIUM Google DeepMind is a US company and subject to the CLOUD Act; model weights are not publicly available.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
75.29
Routine
45.11
Reasoning
30.18

Rank #16

LLM Judge Avg
3.81
100 Coverage
Avg Task Duration
26.15
Interactive
Token Rate
30.93
Output Rate
P95 Latency
44
Top 5 %
Total Tokens
68700
Output Volume
Cost per 1K
$0.01
USD / 1K Requests
Benchmark Cost
$0.69
Total · 68700 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Gemini 2.5 Pro Best model Ø All models
Code Quality 75.44
CLI Benchmark 86.67
Logical Reasoning 75.77
UX Writing 69.43
Documentation 73.06
Content Transform. 78.32
Cultural Intelligence 75.32
Synthesis Quality 60
Tool Execution 90
ToolUse Score 74
Benchmark Cost $0.69

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile