LightweightGoogle
Gemini 3.5 Flash-Lite
Gemini 3.5 Flash-Lite by Google — compare inference API pricing across providers.
Context 1.0M
Tier Lightweight
Tools Supported
Modalities text, image, video, audio
Input from
$0.150 / 1M tokens
across 2 providers
API Pricing
Cheapest on Google Cloud — 40% below avg| Provider | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|
| $0.150 | $1.25 | $0.015 | |
| $0.300 | $2.50 | $0.030 | |
| $0.300 | $2.50 | $0.030 |
Prices updated daily. Last check: Aug 10, 2026
Gemini 3.5 Flash-Lite pricing by provider
Input, output, and batch rates, plus alternatives, for one provider at a time.
Performance & Benchmarks
Source: Artificial Analysis →Intelligence
37.4 / 100
Coding
49.3 / 100
Output Speed
337 t/s
Latency (TTFT)
7.8s
Reasoning & Knowledge
- GPQA Diamond83.8%
- Humanity's Last Exam18.8%
Coding
- SciCode40.9%
Agentic & Tool Use
- Terminal-Bench v2.153.6%
- τ-bench Banking17.5%
Instruction & Long Context
- Long-Context Reasoning74.7%
Benchmarks measured Aug 2026. Scores are independent evaluations, not vendor-reported.
Model Details
General
- Creator
- Family
- Gemini
- Tier
- Lightweight
- Context Window
- 1.0M
- Modalities
- Text, Image, Video, Audio
Capabilities
- Tool Calling
- Yes
- Open Source
- No
- Aliases
- gemini-3.5-flash-lite, Gemini 3.5 Flash-Lite, models/gemini-3.5-flash-lite