Gemini 3.5 Flash

Gemini 3.5 Flash is Google’s fastest Frontier-class model, delivering near Pro-level performance at Flash pricing. With Dynamic Thinking across four configurable levels, a one-million-token context window, and full multimodality for text, images, audio, video, and PDF, the model is well-suited for agentic workflows, coding, and high-throughput workloads.

Google Version 3.5-flash Commercial use permitted MoE 1000 K Context 01/2025 $1.5 / $9 per 1M

  • Proprietary
  • Frontier
  • Google Gemini
  • Text
  • Vision
  • Audio
  • Video
  • Agentic Orchestrator
  • Real-Time

Sovereign Risk: MEDIUM Google DeepMind is a US-based company and subject to the CLOUD Act; the model weights are not publicly accessible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
77.17
Routine
46.95
Reasoning
30.22

Rank #12

LLM Judge Avg
3.88
100 Coverage
Avg Task Duration
9.69
Real-Time
Token Rate
58.67
Output Rate
P95 Latency
21.37
Top 5 %
Total Tokens
54300
Output Volume
Cost per 1K
$0.009
USD / 1K Requests
Benchmark Cost
$0.49
Total · 54300 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Gemini 3.5 Flash Best model Ø All models
Code Quality 74.8
CLI Benchmark 93
Logical Reasoning 73.86
UX Writing 76.87
Documentation 74.3
Content Transform. 73.33
Cultural Intelligence 82.76
Synthesis Quality 63.33
Tool Execution 83.33
ToolUse Score 76.33
Benchmark Cost $0.49

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile