Status tracker · shipped

Kimi K3 Is Officially Out
Confirmed vs rumor

Moonshot AI officially released Kimi K3 on 2026-07-16. This page keeps tracking what is officially confirmed versus what remains community claims.

Current status (updated 2026-08-22)

Kimi K3 officially launched on 2026-07-16, and its weights were open-sourced on 07-27 (2.8T Stable LatentMoE — the largest open-weight model ever). Official API pricing is public ($3/$15 per million tokens). QCode /models already lists kimi-k3, with real calls over the last 30 days.

Confirmed vs community rumor

Officially confirmed

Release date 2026-07-16; weights open-sourced 07-27 (2.8T — the largest open weights ever); 1M-token context; official API price $3/$15; coding is one of the headline use cases.

Still not fully confirmed

Clarified: the parameter count (2.8T), the open-sourcing of weights on 07-27 and official API pricing ($3/$15) have all been published. The next-generation K4 has started training (multiple media reports on 07-29) — no date, no specs. Note: QCode catalog pricing follows the live /models list, which is a different basis from the vendor's official list price.

Why track Kimi K3

For Chinese developers, additional strong long-context options matter for agent workflows. Diversifying providers helps when any single quota window tightens.

How Kimi K3 fits in the 2026 landscape

Vs GLM-5.2

Different training, closed vs open weights trade-offs. Track both.

Vs GPT long context

Chinese ecosystem strength and cost profile may differ.

Provider hedge

Another option when Claude or Codex windows tighten.

Access via QCode today

QCode already surfaces current Moonshot/Kimi models when available. Watch /models for new K3 ids. One key gives access to multiple providers including this family when live.

Release date 2026-07-16; weights open-sourced 07-27 (2.8T — the largest open weights ever); 1M-token context; official API price $3/$15; coding is one of the headline use cases.

Clarified: the parameter count (2.8T), the open-sourcing of weights on 07-27 and official API pricing ($3/$15) have all been published. The next-generation K4 has started training (multiple media reports on 07-29) — no date, no specs. Note: QCode catalog pricing follows the live /models list, which is a different basis from the vendor's official list price.

For Chinese developers, additional strong long-context options matter for agent workflows. Diversifying providers helps when any single quota window tightens.

Kimi K3 officially launched on 2026-07-16, and its weights were open-sourced on 07-27 (2.8T Stable LatentMoE — the largest open-weight model ever). Official API pricing is public ($3/$15 per million tokens). QCode /models already lists kimi-k3, with real calls over the last 30 days.

Release date 2026-07-16; weights open-sourced 07-27 (2.8T — the largest open weights ever); 1M-token context; official API price $3/$15; coding is one of the headline use cases.

Clarified: the parameter count (2.8T), the open-sourcing of weights on 07-27 and official API pricing ($3/$15) have all been published. The next-generation K4 has started training (multiple media reports on 07-29) — no date, no specs. Note: QCode catalog pricing follows the live /models list, which is a different basis from the vendor's official list price.

For Chinese developers, additional strong long-context options matter for agent workflows. Diversifying providers helps when any single quota window tightens.

Kimi K3 officially launched on 2026-07-16, and its weights were open-sourced on 07-27 (2.8T Stable LatentMoE — the largest open-weight model ever). Official API pricing is public ($3/$15 per million tokens). QCode /models already lists kimi-k3, with real calls over the last 30 days.

Release date 2026-07-16; weights open-sourced 07-27 (2.8T — the largest open weights ever); 1M-token context; official API price $3/$15; coding is one of the headline use cases.

Clarified: the parameter count (2.8T), the open-sourcing of weights on 07-27 and official API pricing ($3/$15) have all been published. The next-generation K4 has started training (multiple media reports on 07-29) — no date, no specs. Note: QCode catalog pricing follows the live /models list, which is a different basis from the vendor's official list price.

Why track Kimi K3

For Chinese developers, additional strong long-context options matter for agent workflows. Diversifying providers helps when any single quota window tightens.

Kimi K3 officially launched on 2026-07-16, and its weights were open-sourced on 07-27 (2.8T Stable LatentMoE — the largest open-weight model ever). Official API pricing is public ($3/$15 per million tokens). QCode /models already lists kimi-k3, with real calls over the last 30 days.

Release date 2026-07-16; weights open-sourced 07-27 (2.8T — the largest open weights ever); 1M-token context; official API price $3/$15; coding is one of the headline use cases.

Clarified: the parameter count (2.8T), the open-sourcing of weights on 07-27 and official API pricing ($3/$15) have all been published. The next-generation K4 has started training (multiple media reports on 07-29) — no date, no specs. Note: QCode catalog pricing follows the live /models list, which is a different basis from the vendor's official list price.

Kimi K3 FAQ

Is K3 publicly usable now?

Yes. Since 2026-07-16, logged-in users can use K3·Max and K3 Cluster·Max. For API details, follow Moonshot's official platform docs.

How is this different from GLM-5.2 or GPT long context?

Different providers, different training mixes and context handling. Having options lets you route or fallback when one has quota pressure.

Should I wait for K3 or use current options?

K3 is out — start testing it on your own tasks. For production rotation, wait until API pricing and reliability data are complete; current Claude and GPT tiers remain the safe baseline.

Will QCode support K3 when released?

Already supported. kimi-k3 is listed on /models with real calls over the last 30 days; pricing follows the live list.

Stay ready with multi-provider access

One QCode key for current models; new ones added when live.