Skip to main content
LightweightGoogle

Gemini 3.6 Flash

Gemini 3.6 Flash by Google — compare inference API pricing across providers.

Context 1.0M
Tier Lightweight
Tools Supported
Modalities text, image, video, audio
Input from
$0.750 / 1M tokens
across 2 providers

API Pricing

Cheapest on Google Cloud 25% below avg
ProviderInput / 1MOutput / 1MCached / 1M
$0.750$3.75$0.075
$0.750$3.75$0.075
$1.50$7.50$0.150

Prices updated daily. Last check: Aug 9, 2026

Gemini 3.6 Flash pricing by provider

Input, output, and batch rates, plus alternatives, for one provider at a time.

Performance & Benchmarks

Source: Artificial Analysis →
Intelligence
51.6 / 100
Coding
69.2 / 100
Output Speed
184 t/s
Latency (TTFT)
13.0s

Reasoning & Knowledge

  • GPQA Diamond92.8%
  • Humanity's Last Exam40.8%

Coding

  • SciCode52.7%

Agentic & Tool Use

  • Terminal-Bench v2.177.5%
  • τ-bench Banking29.9%

Instruction & Long Context

  • Long-Context Reasoning79.0%

Benchmarks measured Aug 2026. Scores are independent evaluations, not vendor-reported.

Model Details

General

Creator
Google
Family
Gemini
Tier
Lightweight
Context Window
1.0M
Modalities
Text, Image, Video, Audio

Capabilities

Tool Calling
Yes
Open Source
No
Aliases
gemini-3.6-flash, Gemini 3.6 Flash, models/gemini-3.6-flash