Skip to main content
GMI Cloud logo

Kimi K2 Thinking API pricing on GMI Cloud

Every GMI Cloud Kimi K2 Thinking rate we track, compared against 5 providers serving the same model.

Input from /1M
$0.800
Output /1M
$1.20
Price rank
#5 of 5
Last updated
September 30, 2026

The cheapest Kimi K2 Thinking input price we currently track is $0.300/1M on Amazon AWS.

GMI Cloud Kimi K2 Thinking rates

Prices per 1M tokens. Batch and cached-input rates are shown where the provider publishes them.

GMI Cloud Kimi K2 Thinking pricing by mode
ModeInputOutput
Standard$0.800/1M$1.20/1M

About Kimi K2 Thinking

Creator
Kimi
Modalities
text
Full Kimi K2 Thinking details and every provider →

Kimi K2 Thinking on other providers

About GMI Cloud

GMI Cloud is a neocloud combining dedicated NVIDIA GPU compute (H100, H200, B200, GB200, with GB300 on pre-order) and a serverless, OpenAI-compatible inference API for text, image, video, and audio models.

Frequently Asked Questions

How much does Kimi K2 Thinking cost on GMI Cloud?

GMI Cloud serves Kimi K2 Thinking from $0.800 per 1M input tokens, across 1 tracked pricing mode. Prices are collected daily — see the table above for current input, output, and cached-input rates.

Is GMI Cloud the cheapest way to run Kimi K2 Thinking?

GMI Cloud ranks #5 of 5 providers we track serving Kimi K2 Thinking. The cheapest input price right now is on Amazon AWS. Price is only one factor — throughput, latency, and context limits differ between providers.

Does GMI Cloud offer batch or cached pricing for Kimi K2 Thinking?

We track 1 Kimi K2 Thinking offering from GMI Cloud. Batch and cached-input rates appear in the table above when the provider publishes them; blank cells mean we have no current figure for that mode.