Gemini 3.5 Flash
Gemini 3.5 Flash is Google’s fastest Frontier-class model, delivering near Pro-level performance at Flash pricing. With Dynamic Thinking across four configurable levels, a one-million-token context window, and full multimodality for text, images, audio, video, and PDF, the model is well-suited for agentic workflows, coding, and high-throughput workloads.
- Proprietary
- Frontier
- Google Gemini
- Text
- Vision
- Audio
- Video
- Agentic Orchestrator
- Real-Time
Sovereign Risk: MEDIUM Google DeepMind is a US-based company and subject to the CLOUD Act; the model weights are not publicly accessible.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 77.17
- Routine
- 46.95
- Reasoning
- 30.22
- LLM Judge Avg
- 3.88 / 5
- 100 Coverage
- Avg Task Duration
- 9.69s
- Real-Time
- Token Rate
- 58.67tok/s
- Output Rate
- P95 Latency
- 21.37s
- Top 5 %
- Total Tokens
- 54300
- Output Volume
- Cost per 1K
- $0.009
- USD / 1K Requests
- Benchmark Cost
- $0.49
- Total · 54300 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median