Kimi K2 Thinking API pricing on GMI Cloud
Every GMI Cloud Kimi K2 Thinking rate we track, compared against 5 providers serving the same model.
The cheapest Kimi K2 Thinking input price we currently track is $0.300/1M on Amazon AWS.
GMI Cloud Kimi K2 Thinking rates
Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.
| Mode | Input | Output |
|---|---|---|
| Standard | $0.800/1M | $1.20/1M |
About Kimi K2 Thinking
- Creator
- Kimi
- Modalities
- text
Kimi K2 Thinking on other providers
About GMI Cloud
GMI Cloud is a neocloud combining dedicated NVIDIA GPU compute (H100, H200, B200, GB200, with GB300 on pre-order) and a serverless, OpenAI-compatible inference API for text, image, video, and audio models.
Frequently Asked Questions
How much does Kimi K2 Thinking cost on GMI Cloud?
GMI Cloud serves Kimi K2 Thinking from $0.800 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.
Is GMI Cloud the cheapest way to run Kimi K2 Thinking?
GMI Cloud ranks #5 of 5 providers we track serving Kimi K2 Thinking. The cheapest input price right now is on Amazon AWS. Price is only one factor — throughput, latency, and context limits differ between providers.
Does GMI Cloud offer batch or cached pricing for Kimi K2 Thinking?
We track 1 Kimi K2 Thinking offering from GMI Cloud. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.