Moonshot AI · Released July 2026
Independent fan site — not affiliated with or endorsed by Moonshot AI

Kimi
K3

Kimi K3 is live — Moonshot AI's 2.8T-parameter flagship with 1M token context, native vision and max-effort thinking, available now via kimi.ai and the Kimi API.

Visit Kimi.aiLearn More
1Mctx tokensconfirmed
2.8Ttotal paramsconfirmed
896MoE expertsconfirmed

02 — Technical Specifications

Official Kimi K3 specs and pricing — confirmed at launch

Confirmed
Rumored
Confirmed

Architecture

MoE + KDA Hybrid Attention

Sparse MoE built on Kimi Delta Attention (KDA), a hybrid linear-attention mechanism with attention residuals

Confirmed

Parameters

2.8T Total / 896 Experts

2.8 trillion total parameters across 896 experts, with 16 experts activated per token

Confirmed

Context Length

1M Tokens

Up to 1M tokens for Allegretto tier and above; Moderato tier gets 256K

Confirmed

Modalities

Text + Vision

Native visual understanding built in from the ground up

Confirmed

Reasoning

Thinking Effort: max

Ships with max thinking effort (reasoning_effort: max); low and high tiers roll out later

Confirmed

Code Capability

Flagship Coding

Moonshot's strongest model for coding, game/3D and knowledge tasks — coding scores surpass Claude Fable 5

Confirmed

API Price (per 1M tokens)

$3 In / $15 Out

Cached input just $0.30/1M — Mooncake serving keeps coding cache rates above 90%, cutting real input cost ~4×

Confirmed

Open Source

Open Weights · Modified MIT

Full model weights land by July 27, 2026 under a Modified MIT license — the first open 3T-class model

* All specifications are sourced from Moonshot AI's official documentation and launch announcements (July 2026).

03 — Latest News

Stay up-to-date with everything Kimi K3

Moonshot AI OfficialJuly 17, 2026

Kimi K3 Officially Released: 2.8T-Parameter Flagship with 1M Context

Moonshot AI launched Kimi K3 on July 16, 2026 — a 2.8-trillion-parameter MoE model (896 experts, 16 active) built on KDA hybrid linear attention, with native vision and up to 1M token context. API pricing lands at $3/$15 per 1M tokens, and open weights under a Modified MIT license follow by July 27.

Read More
Moonshot AI OfficialMarch 28, 2026

Moonshot AI Teases Kimi K3: The Next Evolution of AI Intelligence

Moonshot AI officially teased the upcoming Kimi K3 model, promising significant improvements in reasoning, context length, and multimodal capabilities over its predecessor K2.

Read More
TechCrunchMarch 20, 2026

Kimi K3 Reportedly Achieves GPT-5 Level Performance in Benchmarks

Early benchmark leaks suggest Kimi K3 achieves competitive performance against leading frontier models, with particularly strong results in mathematical reasoning and code generation tasks.

Read More
The VergeMarch 15, 2026

Kimi K3 to Feature Revolutionary Long-Context Window of 1 Million Tokens

Sources close to Moonshot AI indicate that Kimi K3 will support context windows up to 1 million tokens, enabling processing of entire codebases, books, and complex multi-step workflows in a single session.

Read More
VentureBeatMarch 5, 2026

How Kimi K3 Could Reshape the Global AI Landscape

Analysis of Moonshot AI's strategic positioning with K3 suggests it could become a major challenger to Western AI labs, with strong performance across Asian languages and competitive English capabilities.

Read More
WiredFebruary 20, 2026

Inside Moonshot AI: The Story Behind China's Fastest Growing AI Startup

An in-depth look at Moonshot AI's journey from its founding to becoming a billion-dollar company, and how Kimi K3 represents their most ambitious technical achievement to date.

Read More

04 — Frequently Asked Questions

Everything you want to know about Kimi K3

Not affiliated with Moonshot AI.

Kimi K3 is Moonshot AI's flagship large language model, released on July 16, 2026. It's a 2.8-trillion-parameter Mixture-of-Experts model (896 experts, 16 activated per token) built on Kimi Delta Attention (KDA) hybrid linear attention, with native vision understanding and up to a 1 million token context window. Moonshot positions it as its strongest model for coding, game/3D and knowledge tasks.