TL;DR

Moonshot AI released Kimi K3 on July 16 with 2.8 trillion total parameters and API prices matching Claude Sonnet 5’s standard rates. Independent testing places it near leading Western models, but promised weights, licensing terms and a technical report remain pending.

Moonshot AI released Kimi K3 on July 16, pricing the new artificial intelligence model at $3 per million input tokens and $15 per million output tokens—the same standard rates cited for Anthropic’s Claude Sonnet 5. The launch matters because it positions a leading Chinese model as a capability competitor at Western prices, rather than primarily as a cheaper alternative.

Moonshot describes K3 as its most capable model to date, with 2.8 trillion total parameters in a sparse mixture-of-experts architecture. The company says the model routes 16 of 896 experts per token, supports text, image and video input, and offers a maximum context window of 1,048,576 tokens.

K3 is available through the Kimi app, Playground and API. Moonshot’s published API rates are $3 for one million input tokens, $15 for one million output tokens and $0.30 for cached input. Thorsten Meyer AI estimates that the input and output rates are about five times those of the K2 family. Anthropic’s temporary Sonnet 5 introductory rates of $2 and $10, reportedly running through August 31, would make K3 50% more expensive during that period.

Artificial Analysis’ Intelligence Index version 4.1 gave K3 a reported score of 57.1, 2.8 points behind Claude Fable 5 and 1.8 points behind GPT-5.6 Sol Max in the cited configuration. Its long-horizon score reportedly rose by 732 Elo points from K2.6, reaching 1,547, while K3 placed first on Design Arena. These are independent test results reported shortly after release, while Moonshot’s separate benchmark figures remain self-reported.

At a glance
analysisWhen: Released July 16, 2026; weights promise…
The developmentMoonshot AI launched Kimi K3 at Western mid-tier API prices, challenging the view that Chinese AI models compete mainly through lower costs.
AI Dispatch · Reality Check · 17 July 2026

Kimi K3: the gap closed six months early — and China stopped competing on price

Every write-up today says “China caught up.” True — and the less interesting half. The other half: K3 costs 5× its predecessor, making it the most expensive Chinese model ever, priced at exact parity with Claude Sonnet 5. A benchmark is a claim. A price is a claim the vendor has to live with.

The gap — measured by someone other than Moonshot (Artificial Analysis v4.1)
Claude Fable 5 (Opus 4.8 fallback)59.9
GPT-5.6 Sol Max58.9
Kimi K3 — open-weight*57.1
2.8 points to the frontier. #4 tested config, effectively the #3 family — and just 0.54 behind Sol xhigh. #1 on Design Arena. A 732-point Elo jump over K2.6 on AA’s long-horizon tracker, to 1547. Analysts expected this tier in early 2027.
◆ The story nobody’s writing — the discount is gone
~$0.60 / $3
K2 family (approx.)
→ 5× →
$3 / $15
Kimi K3 — priciest Chinese model ever
=
$3 / $15
Claude Sonnet 5 list

For two years the thesis was “cheap alternative.” Moonshot just abandoned it. Vendors discount when they’re compensating for something — Moonshot has stopped compensating. With Sonnet 5’s intro rate at $2/$10 through 31 Aug, K3 currently costs 50% more than the model it’s priced against. The competition just moved from cheap vs good to good vs good at the same price, with one of them open — and you can’t answer that with a discount.

⚠ Read the licence before the leaderboard — *it isn’t open yet
Weights promised by 27 July — not available today Licence unpublished — the whole ballgame Technical report unpublished Active param count undisclosed (16 of 896 experts routed) 1M context is a maximum, not an entitlement (Moderato capped at 256K) Max reasoning only at launch 2.8T = a datacentre problem, not a workstation
Everyone calling K3 “the largest open-source model ever” today is describing a press release. Inkling’s story was Apache 2.0 — real, permissive, checkable. K3’s terms are unknown.
⚑ The scale story cuts against the efficiency narrative

The story we’ve told: export controls forced Chinese labs into efficiency. But K3 is 2.8T — the largest open model ever, ~3× K2, vs DeepSeek V4-Pro’s 1.6T. That’s not more with less. That’s more with more. Caveat: sparse MoE, active params undisclosed — total ≠ FLOPs. But if the controls were binding at the frontier, this model shouldn’t exist.

⚖ The distillation asymmetry

Anthropic has accused Moonshot, Z.AI, MiniMax, Alibaba & DeepSeek of “illicit” distillation — possibly well-founded; I can’t assess it. But one day earlier, Thinking Machines said Inkling’s post-training bootstrapped on Kimi K2.5 — reported as ecosystem health. Same verb, different flag, different word. If the distinction is real, someone should articulate it.

The take

Two things changed, neither in the headlines. The discount is gone — anyone whose China strategy was “they’re cheaper” needs a new strategy. And the controls didn’t work — six months early, biggest model ever, from a lab that was supposed to be compute-starved, while Washington’s options narrow to loosening restrictions on its own labs, criminalising distillation, or subsidising American open weights. That’s not containment. It’s a menu of concessions. The gap is 2.8 points and closing. The price is Sonnet’s. The weights are ten days out. Everything that matters happens on 27 July.

Sources: Moonshot’s K3 launch materials, platform docs & pricing (2.8T params, 16-of-896 routing, Kimi Delta Attention, 1,048,576 context, text/image/video, Max-only reasoning, $3/$15/$0.30, weights by 27 July); Simon Willison; Artificial Analysis Intelligence Index v4.1 & long-horizon Elo, via AA and aggregating coverage; Sonnet 5 comparison pricing; Yutong Zhang (WEF); Thinking Machines’ Inkling (15 July) & its stated K2.5 post-training use; Anthropic’s distillation accusations and reported US policy deliberations per Fortune/Bloomberg/CNBC. Moonshot’s own benchmarks are self-reported; AA figures are independent but one day old. Licence, technical report & active params unpublished at time of writing. Not investment advice.
thorstenmeyerai.com

Western Pricing Resets Competition

The price is a direct test of whether customers will treat Chinese frontier models as peers, rather than discounted substitutes. For the past two years, Chinese developers often attracted users through lower API costs and downloadable weights. Moonshot is now asking buyers to compare K3 with Western systems on performance, reliability and deployment options.

The launch also adds pressure on Western model providers. If K3’s reported performance holds across wider testing, vendors may face a rival offering near-frontier capability with promised downloadable weights at the same API price. That could affect enterprise purchasing decisions, model-routing services and national strategies built on the assumption that Chinese laboratories remain behind at the highest performance tier.

The GPT-4 Millionaire: Future of Business Featuring Microsoft 365 Copilot: How to Leverage AI Language Models to Grow Your Company and How AI-driven Language Models Will Revolutionize the Way We Work

The GPT-4 Millionaire: Future of Business Featuring Microsoft 365 Copilot: How to Leverage AI Language Models to Grow Your Company and How AI-driven Language Models Will Revolutionize the Way We Work

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

K3 Arrives Ahead of Forecasts

Thorsten Meyer AI said models at K3’s reported performance level had been expected in early 2027, putting the release about six months ahead of that forecast. The assessment depends on early benchmark results, and no broad consensus about that timeline has yet been established.

K3 is also much larger by total parameter count than Moonshot’s roughly one-trillion-parameter K2 family and the cited 1.6-trillion-parameter DeepSeek V4-Pro. Its scale complicates claims that Chinese laboratories advanced mainly through doing more with less computing power. Because K3 uses sparse expert routing and Moonshot has not disclosed the active parameter count, however, total parameters do not reveal its computing cost.

“Our most capable model to date, with 2.8 trillion parameters.”

— Moonshot AI, in launch material cited by Thorsten Meyer AI

Licence and Compute Details Pending

K3’s promised weights were not available at launch, so descriptions of it as an open or open-source model remain premature. Moonshot says the weights will arrive by July 27, but the licence has not been published. Until its terms appear, users cannot determine whether commercial use, modification and redistribution will be permitted without material restrictions.

Moonshot has also not released a full technical report or disclosed the number of active parameters. The million-token context figure is a stated maximum, while some access tiers may have lower limits. Only the Max reasoning setting was available at launch, leaving the performance and pricing effects of other reasoning levels unknown.

July 27 Release Faces Scrutiny

Attention will shift to Moonshot’s promised July 27 weight release. Developers will examine the licence, hardware needs, model files and deployment requirements to determine whether K3 can be used outside Moonshot’s hosted services and whether its 2.8-trillion-parameter scale is practical for independent operators.

Further independent testing will show whether K3’s early scores hold across coding, reasoning, multimodal work and long-context tasks. Buyers will also watch real-world latency and reliability, while competitors may respond through price changes, stronger open-weight releases or revised product tiers.

Key Questions

When did Moonshot AI release Kimi K3?

Kimi K3 was released on July 16, 2026, through the Kimi app, Playground and API. Its downloadable weights are promised for July 27.

How much does Kimi K3 cost?

Moonshot lists K3 at $3 per million input tokens, $15 per million output tokens and $0.30 per million cached input tokens.

Is Kimi K3 open source?

Not at this stage. The weights and licence were unpublished at launch, so the extent of permitted use cannot be confirmed until Moonshot releases the licensing terms.

How does K3 compare with leading Western models?

Early Artificial Analysis results place K3 2.8 points from the cited leader and near GPT-5.6 Sol Max. Those results support a near-frontier performance claim, but wider testing is still needed.

Why is K3’s pricing drawing attention?

K3 uses the same $3 input and $15 output rates cited for Claude Sonnet 5’s standard pricing. That choice suggests Moonshot expects customers to judge K3 on capability rather than a large discount.

Source: Thorsten Meyer AI

You May Also Like

Mobilisiert, nicht ausgegeben: Was von Europas €200-Milliarden-KI-Offensive übrig bleibt

EU InvestAI is framed as EUR 200bn, but much of the headline sum depends on private capital and pending gigafactory procurement.

The 24% Rule In AI: A Wake-Up Call For Sovereignty Certification Reliability

A July 2026 analysis argues Europe’s common cloud badges certify security practice, not ownership — only SecNumCloud’s 24% cap tests who really controls a provider.

Dimon Says Rates Risk Going Much Higher After Bond Selloff

JPMorgan’s Jamie Dimon warns that interest rates could rise significantly further following recent bond market declines, raising concerns for markets and borrowers.