Kimi K2.7 Code by Kimi: thinking-only agentic coding MoE, 1T params (32B active), 256K context, native multimodality, 30% fewer thinking tokens than K2.6.
Capabilities, design details, and architectural traits
Kimi K2.7 Code is Kimi's dedicated coding model, built on K2.6 with a focus on long-horizon coding task completion and instruction compliance in extended contexts. Its defining characteristic is that it operates exclusively in thinking mode - non-thinking mode is unsupported and disabling it causes an API error.
The model reduces overthinking tendencies by 30% on average compared to K2.6, using fewer thinking tokens while improving task success rates. It also improves agentic capabilities by 10% over K2.6.
| Trait | Detail |
|---|---|
| Thinking-only mode | No non-thinking mode available; disabling thinking causes an API error |
| Reduced overthinking | ~30% fewer thinking tokens than K2.6 on average |
| Architecture | 1T-parameter MoE with 32B active, 384 experts, 8 selected per token, 1 shared expert |
| Native INT4 quantization | MoE weights stored in INT4; BF16 for non-MoE layers |
| Multimodal input | MoonViT vision encoder supports text, image, and video input |
| Fixed sampling | temperature locked at 1.0, top_p at 0.95, n at 1, penalties at 0.0 |
| HighSpeed variant | Same model at ~180 tokens/s, up to 260 tokens/s in short contexts |
| Context window | 256K tokens |
Kimi K2.7 Code supports multi-step tool invocation and reasoning, combining visual understanding with function calling. The model targets complex logical reasoning, mathematical problems, and code writing within its 256K context window.
Independent evaluations · Artificial Analysis
Evaluate specifications, pricing, and independent benchmark indices
| Model Details | |||
|---|---|---|---|
| General Info | |||
| Provider | Kimi | Anthropic | Anthropic |
| Release Date | June 12, 2026 | July 24, 2026 | June 9, 2026 |
| Knowledge Cutoff | - | May 2026 | - |
| Context & Limits | |||
| Context Window | 262K | 1M Best Context Window | 1M Best Context Window |
| Pricing (per 1M tokens) | |||
| Input Pricing | $0.95 Best Input Pricing | $5 | $10 |
| Output Pricing | $4 Best Output Pricing | $25 | $50 |
| Modalities | |||
| Inputs | textimagevideo | textimage | textimagefile |
| Outputs | text | text | text |
| Benchmarks (0-100) | |||
| Intelligence Index | 43.0 | 63.1 Best Intelligence Index | 62.1 |
| Coding Index | 60.8 | 78.0 Best Coding Index | 76.5 |
| Agentic Index | 30.3 | 59.2 Best Agentic Index | 56.6 |
Graduate-level reasoning and expert Q&A evaluation.
Extremely difficult logical reasoning and knowledge.
Logical reasoning over long context windows.
Independent evaluation data provided by Artificial Analysis. To view the latest benchmarks and full details, visit their official site.