Kimi K3-256k(kimi.com)
455 points by monneyboi 16 hours ago | 130 comments
tl;dr: Kimi has launched K3-256k, a 256K-context variant of its flagship K3 coding model (2.8T params, 1M context) that uses roughly half the quota of full K3 while delivering equivalent results within its context limit. The model supports image input only (no video), requires Thinking mode enabled (otherwise requests fall back to K2.6), and is available to Moderato-tier subscribers and above via the Kimi Code CLI, VS Code extension, or OpenAI/Anthropic-compatible APIs. Switching models invalidates context caches, so starting a fresh session is recommended to avoid re-prefill overhead.
HN Discussion:
  • 256k context is sufficient for practical use, making the cheaper variant appealing
  • This mirrors industry pricing patterns for context-length tiers and reflects real compute costs
  • Questioning whether this is truly a new model or just an API/pricing change
  • ~Open source appeal is undermined by prohibitive VRAM requirements for local use
  • Excitement about the effective price cut for users under 256k context