Kimi K3, explained
Facts checked 2026-07-21Kimi K3 is Moonshot AI’s frontier model, announced on 2026-07-16 — a 2.8-trillion-parameter, natively multimodal model with a million-token context window. Within days of launch it became the highest-scoring open-weights-track model ever listed on Artificial Analysis’ Intelligence Index and took the #1 spot on LMArena’s WebDev arena. Here is where it came from, what is inside it, and what the numbers actually say.
The company behind it
Moonshot AI was founded in March 2023 by three Tsinghua University schoolmates — Yang Zhilin, Zhou Xinyu and Wu Yuxin — and is named after Pink Floyd’s *The Dark Side of the Moon*. It is one of China’s so-called “Six AI Tigers,” backed by Alibaba and Tencent, and was last reported at a $3.8B valuation with a Hong Kong IPO under consideration (Wikipedia).
Its consumer app, Kimi, made long context a household feature in China — launching in October 2023 with the ability to digest roughly 200,000 Chinese characters in one conversation, and going so viral by March 2024 that a two-day outage forced a public apology. The serving platform built for it, Mooncake, processes on the order of 100 billion tokens a day and won a Best Paper award at USENIX FAST.
The road to K3
K3 caps a remarkable twelve months. In July 2025, Kimi K2 became the first trillion-parameter open-weight model — the most-downloaded model on Hugging Face the day after release, and the moment *Nature* called “another DeepSeek moment.” Moonshot then shipped relentlessly: K2 Thinking (November 2025, an open reasoning model that could chain 200–300 tool calls), K2.5 (January 2026, native vision), K2.6 (April 2026, coding), and K2.7 Code (June 2026). Each major release pulled around a million Hugging Face downloads.
Moonshot’s own launch claim is that Kimi models have set the upper bound of open-model scale in nine of the twelve months from July 2025 to July 2026 (Kimi K3 announcement). K3 itself launched proprietary-first, with full weights pledged for 2026-07-27 and a technical report to follow; every K2-family release used a Modified MIT license, though K3’s license is not yet announced.
Inside the model
- 2.8 trillion total parameters — the first open model at that scale. Active-parameter count is not yet disclosed; the MoE routes 16 of 896 experts per token.
- Kimi Delta Attention (KDA) — a hybrid linear-attention design — plus “Attention Residuals,” which retrieve representations across depth instead of accumulating them.
- 1,048,576-token context window with native image input (video understanding in the Kimi apps).
- Trained with Per-Head Muon, an evolution of the Muon optimizer Moonshot scaled for K2, with quantization-aware training from the SFT stage — weights ship in MXFP4.
- Moonshot claims the combined changes deliver ~2.5× the scaling efficiency of K2. Training cost and token counts are undisclosed.
What the benchmarks say
The honest framing — which is also Moonshot’s own — is that K3 still trails Claude Fable 5 and GPT-5.6 Sol overall, and beats essentially everything else it was tested against, including Claude Opus 4.8, GPT-5.5 and GLM-5.2. Five days in, most numbers are vendor-reported; the independent signals so far are Artificial Analysis and LMArena.
| Benchmark | Kimi K3 | Best closed comparison |
|---|---|---|
| Artificial Analysis Intelligence Index v4.1 (independent) | 57 — #4 overall, #1 open-track | Claude Fable 5: 59.9 |
| LMArena WebDev arena (independent) | #1, Elo 1677 | — |
| LMArena text arena (independent) | #8, ~1487 Elo | — |
| Terminal-Bench 2.1 | 88.3 | GPT-5.6 Sol: 88.8 |
| GPQA-Diamond | 93.5 | GPT-5.6 Sol: 94.1 |
| BrowseComp (with compaction) | 91.2 | GPT-5.6 Sol: 90.4 |
| DeepSWE v1.1 (agentic SWE) | 67.5 | GPT-5.6 Sol: 73.0 |
| Humanity’s Last Exam (no tools) | 43.5 | Claude Fable 5: 53.3 |
Full table with sources in the K3 announcement; comparisons mix Moonshot’s runs with published leaderboard numbers. Classic suites like SWE-bench Verified and AIME have no published K3 numbers yet.
Two practical caveats from Artificial Analysis’ independent measurement: served from Kimi’s own API, K3 output ran at ~39.5 tokens/second, and the model is, in their words, “notably slow and very verbose” — verbosity that cost $2,709.75 to run their full index eval. Fast third-party hosting will likely follow the weights release.
Pricing: the end of the shock-cheap era
Moonshot’s list price for K3 is $3.00 per million input tokens ($0.30 on cache hits) and $15.00 per million output tokens. That is a deliberate 3–5× step up from the K2 era — K2 launched at $0.15/$2.50 in 2025, and K2.6 still lists at $0.95/$4.00 — pricing K3 at Claude-Sonnet-class rates while claiming near-flagship intelligence. Moonshot counters that its coding workloads see >90% cache-hit rates, making effective cost far lower than list. Smouter’s live rate for kimi-k3 right now: $2.16 in / $10.78 out — 28% under Moonshot’s list, same weights, same 1M context.
The week it moved markets
K3’s launch coincided with a Xi Jinping speech at the World AI Conference and knocked the Nasdaq down about 1% the next day (TechCrunch). Pre-IPO markets repriced the closed labs: IG measured a combined ~$392B drop in Anthropic’s and OpenAI’s implied valuations in the five days after launch. In Washington, reporting by Axios describes a revived push to restrict Chinese models — alongside the admission that banning downloadable weights is nearly unenforceable. Meanwhile adoption keeps compounding: Coinbase’s CEO says running GLM and Kimi models in production roughly halved the company’s AI spend.
Sources
- Moonshot AI — Kimi K3 announcement (specs, benchmarks, pricing, weights pledge)
- Kimi platform docs — K3 pricing ($0.30/$3.00/$15.00, 1,048,576 ctx)
- Artificial Analysis — Kimi K3 (Intelligence Index 57, speed, eval cost)
- LMArena leaderboard — K3 text #8, WebDev #1
- OpenRouter — kimi-k3 listing (price, launch capacity limits)
- Wikipedia — Moonshot AI (founding, funding, Kimi app history, K2 lineage)
- TechCrunch — K3 reception roundup (2026-07-18)
- IG — pre-IPO implied-valuation move after K3 (2026-07-21)
- Tom’s Hardware — US restriction push after K3; adoption notes
- CNBC — Kimi K2 launch pricing (2025-07)