Back to articles
Coding AI

Kimi K2.8 Preview Launches with 1M Context for All Tiers

3 min read

Introduction

Moonshot AI has expanded the Kimi model lineup with K2.8 Preview. The model went live across Kimi Code and Kimi Work on September 11. According to the company, its overall performance is close to K3 while its product positioning is more practical for everyday development. The most notable change is not only the model upgrade, but the broader availability of a 1M-token context window.

Key points

  • A practical coding model. K2.8 Preview is designed to improve coding and agent workflows. It supports low, high, and max reasoning effort, matching the reasoning-level structure used by K3.
  • Multimodal input. The model supports image and video inputs in addition to text.
  • 1M context for every tier. Kimi Code members across all listed plans can access the million-token context through K2.8 Preview, while K3 access and its long-context capability previously required higher-tier plans.
  • A largely seamless migration. The existing kimi-for-coding endpoint is upgraded to K2.8 Preview without changing its Model ID. Some requests that would otherwise use K3 can also be routed to a non-thinking version of K2.8 Preview.
  • Limited public evidence so far. Moonshot has not released benchmark results for this preview, so the exact gap between K2.8 and K3 remains unclear.

Why a K2.8 model matters

K3 is positioned as Moonshot’s high-end model, with strong coding performance and a very long context window. Such a flagship can attract attention, but it is not automatically the best option for every workflow. Higher capability can bring higher inference costs, slower responses, and greater sensitivity to how an agent framework preserves prior reasoning context.

The source material notes that K3 may become less stable when historical thinking content is not fully passed back or when a session switches between models. It can also act too proactively on ambiguous tasks. For code completion, routine edits, and clearly bounded development work, users may prefer a model that is cheaper, more predictable, and easier to access.

K2.8 Preview therefore looks less like a smaller K3 and more like a scalable workhorse. Its role is to handle a broader volume of ordinary coding and office tasks while reserving K3 for cases that genuinely require the highest capability ceiling.

Business implications

The launch arrives as Moonshot accelerates commercialization. The material reports rapid ARR growth after K3’s release and says the company has begun the Hong Kong IPO process. If usage continues to expand, relying on a costly flagship for every request would be difficult to sustain.

K2.8 can help create a broader access layer for Kimi: lower permission barriers, wider long-context adoption, and more predictable daily usage. Still, the claim that it is “close to K3” has not yet been supported by public benchmarks or independent testing. Its real value will depend on cost, latency, reliability, and performance across actual coding workloads.

QbitAI

Comments

Checking sign-in status...

Loading comments...

Related articles