KIMI K3

Kimi K3: the largest open-weight model ever shipped

Released July 16, 2026. 2.8 trillion parameters — open weights. 1M-token context, multimodal reasoning, strong at large-repo navigation and long-horizon agentic work. $3/$15 per million tokens.

Try Kimi K3 on Council See it in a council

Kimi K3 is Moonshot AI's flagship, released July 16, 2026, and a milestone for the open-weight world: at 2.8 trillion parameters it is the largest open-weight model ever shipped — roughly triple its predecessor K2.6's 1T. It pairs a 1M-token context window with multimodal reasoning (text + image input) and a training focus Moonshot has doubled down on since K2: large-repo navigation and long-horizon agentic work, where the model must hold a goal across hundreds of files and many tool calls. Pricing lands at $3 per million input tokens and $15 per million output — a big step up from K2.6's $0.60/$2.80 and squarely in premium territory alongside Claude Sonnet 5's standard sticker. K3 is the first open-weight model priced and positioned as a direct flagship peer rather than a budget alternative. The case for Kimi K3: agentic coding across huge codebases, open-weight requirements at frontier quality, and Chinese-language strength. The case against: no native web search, and one-shot reasoning where GPT-5.6 Sol and Claude Fable 5 keep the edge.

Specs

ProviderMoonshot AI (via OpenRouter on Council)
ReleasedJuly 16, 2026
Parameters2.8 trillion — largest open-weight model ever
Context window1,000,000 tokens
MultimodalText + image input
Web searchNo native search
Price (in / out per 1M)$3.00 / $15.00

Best at

Where it loses

Kimi K3 vs Kimi K2.6

K2.6 (April 2026) is a 1T-parameter MoE with a 256K context window at $0.60/$2.80 — an aggressive-value agentic model. K3 nearly triples the parameter count to 2.8T, quadruples the context window to 1M, adds a deeper multimodal reasoning stack, and quintuples the price to $3/$15. That's a repositioning, not just a scale-up: K2.6 competed on value, K3 competes on capability. Both stay in Council's lineup — seat K2.6 when the loop is long but the task is contained, K3 when the repository is huge or the plan horizon is measured in hours.

How Council AI runs it

Kimi K3 is available on every paid Council AI plan (Plus, Pro, Ultra), routed through OpenRouter rather than a dedicated Moonshot integration — functionally identical from the user's side. In a council it runs in parallel with models from up to 8 other labs, and the moderator synthesizes a consensus answer with an agreement score. K3 earns its seat as the open-weight counterweight: when the closed Western flagships all agree, K3's independently-trained perspective is a genuinely uncorrelated check — and when it dissents on a coding question, it's often because it navigated the repo differently.

Frequently asked questions

How much does Kimi K3 cost via API?

$3.00 per million input tokens and $15.00 per million output tokens — a premium-tier price, five times its predecessor K2.6. Council AI bundles it into monthly plan budgets.

What is Kimi K3's context window?

1,000,000 tokens — four times K2.6's 256K, sized for whole-repository work.

Is Kimi K3 really open-weight?

Yes — at 2.8 trillion parameters it's the largest open-weight model ever released. In practice most users run it through hosted providers; on Council AI it's routed via OpenRouter.

How is K3 different from Kimi K2.6?

Nearly 3× the parameters (2.8T vs 1T), 4× the context (1M vs 256K), deeper multimodal reasoning, and 5× the price ($3/$15 vs $0.60/$2.80). K2.6 stays in the lineup as the value agentic pick.

Does Kimi K3 have web search?

No native search. In a Council AI run, pair it with search-capable models like Grok 4.5 or Gemini 3.5 Flash so the council covers real-time questions.

Is Kimi K3 available in Council AI?

Yes — on all paid plans (Plus, Pro, Ultra), running in parallel councils with models from 8 other labs and consensus scoring on every answer.