Gemini 3.1 Flash Lite API pricing on Deep Infra
Every Deep Infra Gemini 3.1 Flash Lite rate we track, compared against 3 providers serving the same model.
Deep Infra currently has the cheapest Gemini 3.1 Flash Lite input price among the 3 providers we track.
Deep Infra Gemini 3.1 Flash Lite rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.250/1M | $1.50/1M |
About Gemini 3.1 Flash Lite
- Creator
- Context
- 1.0M
- Modalities
- text, image, audio, video
Gemini 3.1 Flash Lite on other providers
About Deep Infra
Deep Infra is an AI inference cloud that provides serverless APIs for 100+ models and dedicated GPU rentals. Features OpenAI-compatible endpoints, custom model deployments, and infrastructure optimized for scale.
Frequently Asked Questions
How much does Gemini 3.1 Flash Lite cost on Deep Infra?
Deep Infra serves Gemini 3.1 Flash Lite from $0.250 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Deep Infra the cheapest way to run Gemini 3.1 Flash Lite?
Of the 3 providers we track serving Gemini 3.1 Flash Lite, Deep Infra currently has the lowest input price. Rates change often, and throughput and latency differ between providers — check the comparison above before committing.
Does Deep Infra offer batch or cached pricing for Gemini 3.1 Flash Lite?
We track 1 Gemini 3.1 Flash Lite offering from Deep Infra. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.