DeepSeek V4 Flash Vision API pricing on Novita AI
Every Novita AI DeepSeek V4 Flash Vision rate we track, compared against 2 providers serving the same model.
The cheapest DeepSeek V4 Flash Vision input price we currently track is $0.220/1M on OpenRouter.
Novita AI DeepSeek V4 Flash Vision rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.440/1M | $1.32/1M |
About DeepSeek V4 Flash Vision
- Creator
- DeepSeek
- Modalities
- text
DeepSeek V4 Flash Vision on other providers
About Novita AI
Novita AI is an inference platform serving 200+ models through a single OpenAI-compatible API, alongside serverless GPU jobs, on-demand GPU instances, and bare-metal clusters. Its catalogue is weighted toward open-weight families — Qwen, DeepSeek, GLM, Kimi, MiniMax, Llama and Gemma — with per-token billing and prompt caching on most chat models.
Frequently Asked Questions
How much does DeepSeek V4 Flash Vision cost on Novita AI?
Novita AI serves DeepSeek V4 Flash Vision from $0.440 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is Novita AI the cheapest way to run DeepSeek V4 Flash Vision?
Novita AI ranks #2 of 2 providers we track serving DeepSeek V4 Flash Vision. The cheapest input price right now is on OpenRouter. Price is only one factor — throughput, latency, and context limits differ between providers.
Does Novita AI offer batch or cached pricing for DeepSeek V4 Flash Vision?
We track 1 DeepSeek V4 Flash Vision offering from Novita AI. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.