Kimi Code has introduced a new model, Kimi K3-256k, which is now available for users. This model provides a 256k context window and is positioned alongside the existing Kimi K3 and K2.7 Code models. It is accessible via model ID in various clients and third-party tools.
The K3-256k model is designed for common tasks such as everyday question-and-answer interactions, code completion, routine feature development, and single-file or small-file edits. Within its 256k context limit, it delivers comparable results to the K3 (1M) model. However, it does not support video input.
The K3-256k model consumes approximately half the quota of the K3 (1M) model. When switching from K3 (1M) to K3-256k, users are advised to manually compact their session context to under 256k to preserve task key points and maintain session integrity. Direct switching with video files in conversation history will fail, requiring prior compaction. Switching from K3-256k to K3 (1M) does not affect the cache if the 256k limit is approached.
Access to different models, context windows, and speeds is determined by membership plans. For instance, K3 and K3-256k models require a Moderato plan or higher. The full 1M context for K3 is available on Allegretto and higher tiers, while K3-256k has a fixed 256K context limit. Users may encounter 401 errors if their requested capabilities exceed their plan's entitlements.
✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →
One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.
One email a day. Unsubscribe in one click, any time.
Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.
▶ Play today's briefNew every morning, and the back catalogue is archived by date.
Kimi Code has released K3-256k, a new model with a 256k context window, designed for everyday Q&A, code completion, and routine development tasks. This model offers the same results as the K3 (1M) model within its context limit but consumes less quota.