Skip to main content
LightweightGoogle

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite by Google — compare inference API pricing across providers.

Context 1.0M
Tier Lightweight
Tools Supported
Modalities text, image, video, audio
Input from
$0.150 / 1M tokens
across 2 providers

API Pricing

Cheapest on Google Cloud 40% below avg
ProviderInput / 1MOutput / 1MCached / 1M
$0.150$1.25$0.015
$0.300$2.50$0.030
$0.300$2.50$0.030

Prices updated daily. Last check: Aug 10, 2026

Gemini 3.5 Flash-Lite pricing by provider

Input, output, and batch rates, plus alternatives, for one provider at a time.

Performance & Benchmarks

Source: Artificial Analysis →
Intelligence
37.4 / 100
Coding
49.3 / 100
Output Speed
337 t/s
Latency (TTFT)
7.8s

Reasoning & Knowledge

  • GPQA Diamond83.8%
  • Humanity's Last Exam18.8%

Coding

  • SciCode40.9%

Agentic & Tool Use

  • Terminal-Bench v2.153.6%
  • τ-bench Banking17.5%

Instruction & Long Context

  • Long-Context Reasoning74.7%

Benchmarks measured Aug 2026. Scores are independent evaluations, not vendor-reported.

Model Details

General

Creator
Google
Family
Gemini
Tier
Lightweight
Context Window
1.0M
Modalities
Text, Image, Video, Audio

Capabilities

Tool Calling
Yes
Open Source
No
Aliases
gemini-3.5-flash-lite, Gemini 3.5 Flash-Lite, models/gemini-3.5-flash-lite