Model Configuration
This page covers the models Kimi Code provides and how to switch between them in each client.
Model Overview
Kimi Code currently offers two models—Kimi K3 and Kimi K2.7 Code—across four model IDs, selectable by model ID in clients or third-party tools. Model specs:
Recommended model launch k3-256k is now available. Within 256k context, it delivers the same results. k3 (1M) consumes about twice as much quota as k3-256k . Ideal for everyday Q&A, code completion, routine feature development, and single-file or small-file edits — video input is not supported. 💓 Reminder Switching from K3 (1M) to K3-256k: When switching from k3 (1M) to k3-256k , if the current session's context already exceeds 256k, some coding tools such as Kimi Code CLI and Claude Code will perform a compact on the tool side. Switching recommendations:
(1) Because different agent tools handle this differently, manually run compact once before switching to compress the context to within 256k. This preserves the key points of the task, keeps the session intact, and lets you benefit from more durable quota after switching.
(2) If the conversation history includes video files, switching directly will fail because K3-256k does not support video input. Please compact first, then switch. Switching from K3-256k to K3 (1M): When switching from k3-256k to k3 (1M), if k3-256k is close to the 256k limit and you don't want compact to lose information, you can switch directly to 1M. The current version switching from 256k to 1M does not affect the cache.
Model ID k3 k3-256k kimi-for-coding kimi-for-coding-highspeed Model version Kimi K3 Kimi K3 Kimi K2.7 Code K2.7 Code HighSpeed Description Kimi's most capable flagship coding model: 2.8T parameters, 1M context window The 256K context version of Kimi K3, effectively reducing consumption Good at code completion and routine development tasks The high-speed version of K2.7 Code, with the same coding ability and ~5–6× faster output Speed Regular Regular Regular HighSpeed (6× speed, 3× quota usage) Context window Up to 1M (for higher-tier members) 256k only 256k 256k Reasoning reasoning_effort:low / high / max
(default high ) reasoning_effort:low / high / max (default high ) Thinking:ON Thinking:ON Availability Available to Moderato and above; 1M context for Allegretto and above Available to all Moderato members and above All members Allegretto plan or above Multimodal input Image, video Image only Image, video Image, video
Need a higher membership plan? Different membership plans unlock different models, context windows, and speeds. Upgrade your plan →
... continue reading