Skip to content

Kimi K3 Is Live — Moonshot's Flagship Multimodal Reasoner

July 17, 2026

Kimi K3 is in the model picker: Moonshot's 2.8T-parameter open-weight flagship, built for complex coding, knowledge work, and long-horizon agentic runs. It thinks before every answer, reads text and images, and holds a full 1M-token context — the strongest open-weight model Alfrada OS offers today. It bills at the Premium band (6×).

What you can do

  • Pick Kimi K3 from the selector — it sits at #2 in the recommended picks, right after GLM 5.2. Routed via OpenRouter (Global).
  • Keep whole projects in one pass with a 1M-token context window (1,048,576 tokens) and up to 50K output tokens per response.
  • Get always-on deep reasoning — K3 runs with thinking mode permanently enabled at maximum effort; you don't toggle anything.
  • Send images, not just text — K3 is multimodal, so screenshots, diagrams, and photos can go straight into the same reasoning pass.
  • Use it everywhere agents run — full tool calling, Swarm eligibility, and Auto routing.
  • Stay open-weight at the frontier — 2.8T parameters, the largest open-weight model in the catalog, for teams that prefer open models without giving up frontier quality.

Where this shows up

  • You have a gnarly multi-file refactor plus the design doc plus the logs. You used to split that across sessions or prune context. Now the whole thing fits in one 1M-token pass with reasoning on.
  • You've been using Kimi K2.7 Code as your Moonshot pick. K3 is the newer flagship lane — same family, bigger model, image input, and four times the context.
  • You want an open-weight model for a long agentic run. K3 is swarm-eligible and Auto-routable, so it can lead a run or be pinned explicitly.

Try it

  • Pick Kimi K3 in the picker, then: "Read this whole repository plus the attached architecture diagram, and propose the refactor plan with the riskiest changes first."
  • With Kimi K3 selected: "Work through this contract, the spreadsheet, and these screenshots together, and flag every inconsistency between them."
  • With Kimi K3 selected: "Run this multi-step research task end to end and keep the full source material in context instead of summarizing early."

Heads up

  • K3 bills at the Premium band (6×) against your plan quota — same token rate as GLM 5.2, GPT-5.6 Terra, and Kimi K2.7 Code. It is not a 12× model.
  • Thinking is always on. There is no fast/no-reasoning variant of K3; for quick lightweight turns a Workhorse-band model is the cheaper choice.
  • The recommended row grew to four picks — GLM 5.2, Kimi K3, GPT-5.6 Terra, and Claude Opus 4.8. Kimi K2.7 Code leaves the recommended row but stays fully selectable in the picker.

Built for Alfrada OS.