Google
Gemini 3.8 Flash
Gemini 3.8 Flash by Google — compare inference API pricing across providers.
Input from
$0.375 / 1M tokens
across 1 provider
API Pricing
Cheapest on OpenRouter — 33% below avg| Provider | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|
| $0.375 | $1.88 | $0.037 | |
| $0.750 | $3.75 | $0.075 |
Prices updated daily. Last check: Sep 3, 2026
Gemini 3.8 Flash pricing by provider
Input, output, and batch rates, plus alternatives, for one provider at a time.
Performance & Benchmarks
Source: Artificial Analysis →Intelligence
58.7 / 100
Coding
76.3 / 100
Output Speed
297 t/s
Latency (TTFT)
10.0s
Reasoning & Knowledge
- GPQA Diamond95.3%
- Humanity's Last Exam47.8%
Coding
- SciCode53.6%
Agentic & Tool Use
- Terminal-Bench v2.187.6%
- τ-bench Banking44.9%
Instruction & Long Context
- Long-Context Reasoning81.0%
Benchmarks measured Sep 2026. Scores are independent evaluations, not vendor-reported.
Model Details
General
- Creator
- Modalities
- Text
Capabilities
- Tool Calling
- No
- Open Source
- No