Specialized coding model
Kimi K2.7 Code is positioned as a specialized coding model in the official Kimi API documentation and model list. [K3, K5]
Kimi K2.7 Code guide
Kimi K2.7 Code is Moonshot AI's specialized coding model. It is served through the Kimi API under the model IDs kimi-k2.7-code and kimi-k2.7-code-highspeed, with a 262,144-token (256K) context window and always-on thinking. As of July 24, 2026 it appeared in the official Kimi API model list alongside Kimi K3, with lower official CNY per-token pricing than K3 across all billing categories. [K3, K5, K7]
Overview
Kimi K2.7 Code is positioned as a specialized coding model, in contrast to Kimi K3, which Moonshot AI positions as its flagship model for long-horizon coding, knowledge work, and deep reasoning. K2.7 Code remains available through the Kimi API after the K3 release. [K3, K1]
The official sources reviewed for this page did not state that Kimi K3 replaces Kimi K2.7 Code. Both models appeared in the official Kimi API model list as of July 24, 2026. [K5]
Key features
All capability descriptions are based on official Kimi API documentation, each with a dated access check.
Kimi K2.7 Code is positioned as a specialized coding model in the official Kimi API documentation and model list. [K3, K5]
Kimi K2.7 Code supports a context window of 262,144 tokens (256K). [K3, K4, K7]
Thinking is always on. Kimi K2.7 Code uses thinking: {"type": "enabled"}; no reasoning-intensity control is documented. [K3, K4]
A kimi-k2.7-code-highspeed variant is listed in the official model list. Moonshot AI states it outputs approximately 180 tokens per second, reaching up to 260 tokens per second in short-context scenarios. Independent speed verification has not been performed. [K5]
Kimi K2.7 Code supports tool calling with tool_choice values of auto and none. The required value is not supported and returns an error. [K3, K4]
Context caching is automatic, stated in the pricing page feature summary. A cache hit rate was not stated in the official sources accessed. [K7]
API reference
These fields come from the official Kimi API documentation and may change over time.
| Field | Value | Sources |
|---|---|---|
| API model ID | kimi-k2.7-code; high-speed variant: kimi-k2.7-code-highspeed | K3, K5 |
| Context window | 262,144 tokens (256K) | K3, K4, K7 |
| Thinking / reasoning control | Thinking is always on via thinking: {"type": "enabled"}; no intensity control is documented. | K3, K4 |
| Default max output tokens | 32,768. The maximum configurable value was not confirmed in the accessed official sources as of July 24, 2026. | K3 |
| tool_choice | Supports auto and none only; required is not supported and returns an error. | K3, K4 |
| Dynamic tool loading | Not documented in the K2.7 Code quickstart as of July 24, 2026; absence from that guide does not confirm unsupported status. | K3 |
| Structured output | The K2.7 Code pricing page lists JSON Mode; detailed interface documentation was not found in the quickstart as of July 24, 2026. | K7 |
| Partial Mode | Listed in the K2.7 Code pricing page feature summary; detailed usage documentation was not found in the quickstart as of July 24, 2026. | K7 |
| Context caching | Automatic, stated in the pricing page feature summary. A cache hit rate was not stated in the official sources accessed. | K7 |
API pricing
Pricing below comes from the official Kimi K2.7 Code CNY pricing page, accessed July 24, 2026. USD pricing was not stated in the official sources accessed. Prices may change; check the current official pricing page for the latest rates.
| Billing Category | CNY per 1M tokens |
|---|---|
| Input — cache hit | ¥1.30 standard; ¥2.60 Highspeed |
| Input — cache miss | ¥6.50 standard; ¥13.00 Highspeed |
| Output | ¥27.00 standard; ¥54.00 Highspeed |
CNY pricing is documented on the Kimi K2.7 Code pricing page. USD pricing was not stated in the official sources accessed as of July 24, 2026. [K7]
K2.7 Code and K3
Both kimi-k3 and kimi-k2.7-code, plus kimi-k2.7-code-highspeed, appear in the official Kimi API model list as of July 24, 2026. [K5]
The official sources reviewed for this page did not state that Kimi K3 replaces Kimi K2.7 Code. [K5]
For a side-by-side breakdown, see the Kimi K3 vs K2.7 Code comparison and the Kimi K3 guide.
Known limitations
| Limitation | Detail |
|---|---|
| Architecture and parameter count not stated | The official sources accessed did not publish Kimi K2.7 Code's architecture, total parameters, or active parameters. |
| USD pricing not stated | Kimi K2.7 Code USD pricing was not stated in the official sources accessed as of July 24, 2026. |
| Open weights status not covered | This page covers the hosted Kimi K2.7 Code API. An official open-weight release for K2.7 Code was not confirmed in the sources accessed. |
| Vendor-reported speed not independently verified | The Highspeed throughput figures (~180 tokens/sec, up to ~260 tokens/sec) are vendor-stated; OpenK3 has not independently verified them. |
| No independent benchmark testing | OpenK3 has not executed original coding benchmarks comparing K2.7 Code to other models. |
FAQ
Kimi K2.7 Code is Moonshot AI's specialized coding model, served through the Kimi API under the model IDs kimi-k2.7-code and kimi-k2.7-code-highspeed, with a 262,144-token (256K) context window and always-on thinking. [K3, K5]
The standard model ID is kimi-k2.7-code, and the high-speed variant is kimi-k2.7-code-highspeed, as listed in the official Kimi API model list. [K3, K5]
Kimi K2.7 Code supports a context window of 262,144 tokens (256K), based on official documentation accessed July 24, 2026. [K3, K4, K7]
As of July 24, 2026, official CNY pricing per 1M tokens is ¥1.30 cache-hit input (¥2.60 Highspeed), ¥6.50 cache-miss input (¥13.00 Highspeed), and ¥27.00 output (¥54.00 Highspeed). USD pricing was not stated in the official sources accessed. Prices may change. [K7]
No official source accessed as of July 24, 2026 says Kimi K3 replaces Kimi K2.7 Code. Both kimi-k3 and kimi-k2.7-code appeared in the official Kimi API model list on the access date. [K5]
Both are listed in the official model list. Moonshot AI states the Highspeed variant outputs approximately 180 tokens per second, reaching up to 260 tokens per second in short-context scenarios, and it has higher official CNY per-token pricing than the standard variant. Independent speed verification has not been performed. [K5, K7]
Related
Official sources
Each source below includes its own access date.
https://www.kimi.com/blog/kimi-k3https://platform.kimi.com/docs/guide/kimi-k2-7-code-quickstarthttps://platform.kimi.com/docs/api/models-overviewhttps://platform.kimi.com/docs/modelshttps://platform.kimi.com/docs/pricing/chat-k27-code