Kimi K2 is Moonshot's flagship text model with 128K token context window, featuring tool calling capabilities and optimized for chat and code generation tasks.
| Provider | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|
| $0.300 | $1.25 | - | |
| $0.570 | $2.30 | $0.285 | |
| $0.570 | $2.30 | - | |
| $0.570 | $2.30 | - | |
| $0.600 | $2.50 | - |
Prices updated daily. Last check: Sep 6, 2026
Input, output, and batch rates, plus alternatives, for one provider at a time.
Benchmarks measured Sep 2026. Scores are independent evaluations, not vendor-reported.
Kimi K2 is designed for flagship-tier applications requiring sophisticated text processing and reasoning capabilities. Its 128K context window makes it suitable for long document analysis, extended coding sessions, and multi-turn conversations that require maintaining context over thousands of tokens. The tool calling functionality enables agentic workflows, API integrations, and structured data processing tasks. Code generation capabilities position it for software development assistance, while the thinking variant suggests optimization for reasoning-intensive applications like mathematical problem solving or complex analysis. The model targets enterprise developers and organizations needing reliable performance for production text processing workloads.
Kimi K2 pricing varies by provider and may differ between input and output tokens. Check the pricing table above for current rates across all available providers offering this model.
Kimi K2 excels at long-form text processing with its 128K context window, code generation tasks, and applications requiring tool calling capabilities. It's particularly suited for extended conversations, document analysis, software development assistance, and agentic workflows that need structured API interactions.
While both are Kimi K2 variants, the specific differences aren't detailed in available specifications. The naming suggests kimi-k2-instruct may be optimized for following instructions and general chat, while kimi-k2-thinking could be specialized for reasoning and analytical tasks, but you should test both variants for your specific use case.