Kimi Token API Pricing, Top-up, Deals & Official Entry
Kimi API pricing comparison should start from Kimi K3 flagship, Kimi K2.7 Code programming, and tiered cache billing: Kimi K3 is ¥20/1M tokens input (¥2 cache hit) and ¥100 output with 1M context and always-on inference; Kimi K2.7 Code standard is ¥1.30 / ¥6.50 / ¥27 cache hit / uncached input / output, HighSpeed at ¥2.60 / ¥13 / ¥54; Kimi K2.6 multimodal is ¥1.10 / ¥6.50 / ¥27. This page connects the Moonshot platform, API key setup, automatic context caching, ToolCalls, JSON Mode, `$web_search` at ¥0.03/call, free file extraction, Claude Code / Cline / Roo Code integration, and Tier0–Tier5 rate limits, with deprecation notes that Kimi K2.5 and Moonshot V1 will go offline by 2026-08-31.
Below is the complete pricing comparison for Kimi Token API, covering 5 plan tiers, core model capabilities, quotas, and official entry. All data is sourced from the official website to help you decide whether it’s worth it.
Last updated:
Kimi Token API Price Comparison: Core Models
Kimi's strongest flagship `kimi-k3`: 2.8T parameters, native vision, 1M-token context, always-on inference, ToolCalls (with `tool_choice` and dynamic tool loading), structured output, context caching. ¥2.00 cache, ¥20.00 uncached input, ¥100.00 output per 1M tokens.
Current coding workhorse `kimi-k2.7-code` and `kimi-k2.7-code-highspeed`: text, image, and video input in reasoning-only mode with 256k context. Standard ¥1.30 / ¥6.50 / ¥27.00; HighSpeed ~180 tokens/s at ¥2.60 / ¥13.00 / ¥54.00. Built for complex engineering, multi-step agents, and long-horizon refactors.
General multimodal model `kimi-k2.6`: vision/text/video input, reasoning/non-reasoning modes, chat and agent tasks, 256k context. ¥1.10 cache, ¥6.50 uncached input, ¥27.00 output per 1M tokens—suited to comprehensive workflows needing vision and general agent capabilities.
`kimi-k2.5` stopped for new users, going offline by 2026-08-31. Current pricing ¥0.70 / ¥4.00 / ¥21.00 per 1M tokens, 256k context. Existing users should migrate to Kimi K3, Kimi K2.7 Code, or Kimi K2.6 immediately.
Moonshot V1 series stopped for new users, going offline by 2026-08-31. Includes 8K (¥2/¥10), 32K (¥5/¥20), 128K (¥10/¥30) text and Vision Preview. Migration: short text/vision → Kimi K2.6, long text → Kimi K3 (1M context), lightweight high-frequency → Kimi K2.7 Code standard.
Kimi Token API Price Comparison: Plans
kimi-k3 is Kimi's most capable model to date—2.8 trillion parameters, native visual understanding, 1M-token context window, always-on inference, and adjustable reasoning_effort.tool_choice constraints and dynamic tool loading), JSON Mode, Structured Output (response_format / JSON Schema), Partial Mode, automatic context caching, and web search.kimi-k2.7-code is Kimi's current coding model, supporting text, image, and video input in reasoning-only mode with 256k context; more reliable instruction following in long contexts at higher coding success rates.kimi-k2.7-code-highspeed shares the same model with ~180 tokens/s output (up to 260 tokens/s in short context) at 2x pricing: ¥2.60 cache, ¥13 uncached input, ¥54 output per 1M tokens.kimi-k2.6 is a general multimodal model with text, image, and video input, reasoning and non-reasoning modes, chat and agent tasks, and 256k context; stable instruction following and self-correction.kimi-k2.5 has stopped accepting new users and will go offline across the platform by 2026-08-31. Teams still using Kimi K2.5 should migrate to Kimi K3 or Kimi K2.7 Code as soon as possible.Kimi Token API Price Comparison: Notes
- Kimi K3 flagship: ¥2.00 cache hit, ¥20.00 uncached input, ¥100.00 output per 1M tokens, 1,048,576-token context, always-on inference with configurable `reasoning_effort` (low / high / max, default max).
- Kimi K2.7 Code standard: ¥1.30 / ¥6.50 / ¥27.00 (cache/input/output). HighSpeed same model faster (~180 tokens/s): ¥2.60 / ¥13.00 / ¥54.00. Kimi K2.6 multimodal: ¥1.10 / ¥6.50 / ¥27.00.
- Kimi K2.5 (¥0.70 / ¥4.00 / ¥21.00) and Moonshot V1 have stopped accepting new users and will go offline by 2026-08-31. Migrate to Kimi K3 or Kimi K2.7 Code promptly.
- Each successful `$web_search` trigger costs an extra ¥0.03; file extraction and storage APIs are temporarily free, but extracted document content billed as model input tokens.
Kimi Token API Price Comparison: Tools & Integration
Kimi Token API Pricing, Top-up, Deals, Quota, Usage, Setup & Updates
Page published: