DeepSeek V4 Flash API pricing on Baseten
Every Baseten DeepSeek V4 Flash rate we track, compared against 14 providers serving the same model.
The cheapest DeepSeek V4 Flash input price we currently track is $0.060/1M on Deep Infra.
Baseten DeepSeek V4 Flash rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.130/1M | $0.260/1M |
About DeepSeek V4 Flash
- Creator
- DeepSeek
- Context
- 1.0M
- Modalities
- text
DeepSeek V4 Flash on other providers
About Baseten
Baseten is an inference platform that serves popular open models through per-token Model APIs and runs custom or fine-tuned models on dedicated GPU deployments billed per minute, with managed training jobs on the same hardware.
Frequently Asked Questions
How much does DeepSeek V4 Flash cost on Baseten?
Baseten serves DeepSeek V4 Flash from $0.130 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Baseten the cheapest way to run DeepSeek V4 Flash?
Baseten ranks #4 of 14 providers we track serving DeepSeek V4 Flash. The cheapest input price right now is on Deep Infra. Price is only one factor — throughput, latency, and context limits differ between providers.
Does Baseten offer batch or cached pricing for DeepSeek V4 Flash?
We track 1 DeepSeek V4 Flash offering from Baseten. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.