Google
Gemini 3.5 Flash
Gemini 3.5 Flash by Google — compare inference API pricing across providers.
Input from
$0.750 / 1M tokens
across 3 providers
API Pricing
Cheapest on Google Cloud — 33% below avg| Provider | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|
| $0.750 | $4.50 | $0.075 | |
| $0.750 | $4.50 | $0.075 | |
| $1.50 | $9.00 | - | |
| $1.50 | $9.00 | $0.150 |
Prices updated daily. Last check: Aug 7, 2026
Gemini 3.5 Flash pricing by provider
Input, output, and batch rates, plus alternatives, for one provider at a time.
Performance & Benchmarks
Source: Artificial Analysis →Intelligence
52.0 / 100
Coding
70.1 / 100
Output Speed
233 t/s
Latency (TTFT)
11.1s
Reasoning & Knowledge
- GPQA Diamond92.2%
- Humanity's Last Exam42.7%
Coding
- SciCode53.1%
Agentic & Tool Use
- Terminal-Bench Hard40.9%
- Terminal-Bench v2.178.7%
- τ²-bench95.3%
- τ-bench Banking32.2%
Instruction & Long Context
- IFBench76.3%
- Long-Context Reasoning81.0%
Benchmarks measured Aug 2026. Scores are independent evaluations, not vendor-reported.
Model Details
General
- Creator
- Modalities
- Text
Capabilities
- Tool Calling
- No
- Open Source
- No