Kimi / Moonshot AI model and pricing guide: K2.7 Code, K2.6, K2.5, Moonshot V1, and Kimi Code
Editorial standards and source policy: Editorial standards, Team. Content links to primary sources; see Methodology.
As of July 2, 2026, choose Kimi / Moonshot AI from the current model and pricing table. Coding and repo-agent work starts with kimi-k2.7-code; high-speed IDE-like interaction compares kimi-k2.7-code-highspeed; general multimodal and agent work starts with kimi-k2.6; cheaper Kimi trials can use kimi-k2.5; older V1 generation tiers include moonshot-v1-128k.
| Model | Context | Input cache hit | Input cache miss | Output | Best fit |
|---|---|---|---|---|---|
kimi-k2.7-code |
262,144 tokens | $0.19 / 1M tokens | $0.95 / 1M tokens | $4.00 / 1M tokens | coding model and repo-agent tasks |
kimi-k2.7-code-highspeed |
262,144 tokens | $0.38 / 1M tokens | $1.90 / 1M tokens | $8.00 / 1M tokens | high-speed coding sessions, about 180 tokens/s and up to about 260 tokens/s in short context |
kimi-k2.6 |
262,144 tokens | $0.16 / 1M tokens | $0.95 / 1M tokens | $4.00 / 1M tokens | general multimodal, tool use, JSON mode, and agent work |
kimi-k2.5 |
262,144 tokens | $0.10 / 1M tokens | $0.60 / 1M tokens | $3.00 / 1M tokens | lower-cost Kimi K2.x experiments |
moonshot-v1-128k |
131,072 tokens | n/a | $2.00 / 1M input tokens | $5.00 / 1M output tokens | V1 long-context generation reference |
deepseek-v4-flash |
1M | $0.0028 / 1M tokens | $0.14 / 1M tokens | $0.28 / 1M tokens | low-cost API baseline |
Kimi API pricing is token usage, cache hit/cache miss behavior, recharge balance, limits, and promotions, not a fixed monthly subscription plan. The K2.7 Code top-up rebate page listed a promotion from June 11, 2026 09:00 PDT to July 2, 2026 08:59 PDT: no voucher below $100, 20% voucher for $100-$299, 25% for $300-$999, and 30% for $1,000 or more, with a $4,000 voucher cap and 90-day voucher validity.
| 50k input + 10k output sample | Cache miss cost | Cache hit cost | Read this as |
|---|---|---|---|
deepseek-v4-flash |
$0.0098 | $0.00294 | lowest API cost baseline here |
deepseek-v4-pro |
$0.03045 | $0.00888 | stronger DeepSeek tier, still cheaper than Kimi K2.x in this sample |
kimi-k2.5 |
$0.0600 | $0.0350 | cheapest Kimi option in this set |
kimi-k2.6 |
$0.0875 | $0.0480 | general Kimi multimodal/agent option |
kimi-k2.7-code |
$0.0875 | $0.0495 | Kimi coding model baseline |
kimi-k2.7-code-highspeed |
$0.1750 | $0.0990 | higher price for interactive speed |
Kimi Code is a workflow entry. The official README describes it as a terminal AI coding agent that can read and edit code, run shell commands, search files, fetch web pages, and integrate through MCP, plugins, subagents, hooks, and ACP/IDE. Use the API price table for the token bill and a repo pilot for workflow value.
Official sources: Kimi model list, Kimi K2.7 Code pricing, Kimi K2.6 pricing, Kimi K2.5 pricing, Moonshot V1 pricing, Kimi Code GitHub, DeepSeek models pricing, DeepSeek API docs.
Related reading
- Top China-Built AI Models to Watch in 2026: DeepSeek, Qwen, Kimi & More
- Kimi Pricing and API Access for Builders
- Kimi Code CLI and IDE Evaluation for Builders
- How to Use the DeepSeek API for Builder Workflows
RadarAI helps builders track AI updates, compare source-backed signals, and decide which changes are worth acting on.