Skip to main content
Back to News
news/AI Models

Moonshot AI Opens Kimi K3 Weights, Challenging U.S. AI Dominance

Moonshot AI releases the full 2.8-trillion-parameter Kimi K3 model weights today, the largest open-weight AI system ever available for free download.

Stefan Trbojevic

Stefan Trbojevic

27 July 20262 min read
LinkedIn
Editorial illustration for Moonshot AI Kimi K3 open weights release

The takeaway

Open-weight frontier models are no longer hypothetical. The gap between proprietary APIs and downloadable models has functionally closed, and the question is now what builders do with that capability.

Why it matters for builders

Open-weight frontier models eliminate per-token API costs and enable self-hosting, fine-tuning, and full model inspection — changing the economics of AI deployment for technical teams.

Moonshot AI Opens Kimi K3 Weights, Challenging U.S. AI Dominance

Moonshot AI is releasing the full model weights for Kimi K3 today, July 27 — making the 2.8-trillion-parameter system the largest open-weight AI model ever made available for free download. The release, confirmed by a countdown timer on Moonshot's Hugging Face page with over 2,600 developers waiting, marks a pivotal moment in the global AI race.

What Kimi K3 brings to the table

Kimi K3 is a mixture-of-experts model that activates 16 of its 896 experts per token, with a one-million-token context window and native vision capabilities. On Arena.ai's Frontend Code Arena — a blind developer voting leaderboard — K3 scored first overall with a 76% win rate, ahead of Anthropic's Claude Fable 5. The Artificial Analysis Intelligence Index places K3 fourth globally, trailing only Claude Fable 5 and GPT-5.6 Sol.

What sets this release apart is the price tag. While Anthropic's Opus 5 costs $25 per million output tokens and GPT-5.6 Sol even more, K3's API pricing lands at roughly $15 per million output tokens. And with open weights, organizations can self-host the model entirely, eliminating per-token costs.

Kimi K3 benchmark comparison chart showing performance relative to GPT-5.6 and Claude Fable 5

Why today matters

The July 27 weight release transforms Kimi K3 from a service you can use into a tool you can own. Developers can now download, inspect, fine-tune, and deploy the model on their own infrastructure under a Modified MIT license. The download is approximately 594GB in MXFP4 format, with vLLM support shipping alongside the weights.

The timing is deliberate. Dropping at 11AM ET — 11PM in Beijing — Moonshot is clearly targeting the American market. Coming just weeks after the White House accused Moonshot of distilling Anthropic's Fable model and amid growing debate over open-weight restrictions, the release directly challenges the guarded approach of U.S. frontier labs.

What builders should watch

Independent testing by Artificial Analysis found K3's hallucination rate climbed to approximately 51%, up from 39% on its predecessor — a significant caveat for knowledge-work and RAG pipelines. The model's strengths are clearest in agentic coding and frontend development, where its benchmark results are independently verified.

For AI builders, the takeaway is straightforward: open-weight frontier models are no longer a hypothetical. The gap between what you can rent from a proprietary API and what you can download and own has functionally closed. The question is no longer whether open models will catch up — it's what you build with them now that they have.

Share𝕏

The Automation Brief

Read 5 AI stories instead of 50.

The essential moves in AI agents, models, automation and infrastructure — filtered for builders and operators, with the part that actually matters.

No noise. Unsubscribe anytime.

Editorial notes

Reported by

Stefan Trbojevic

Edited by

n8n Lab Editorial

Published

27 July 2026

Updated

27 July 2026

AI disclosure: AI assisted with research and drafting. Factual claims are reviewed by an editor.

n8n Lab is an independent service provider. We are not affiliated with, endorsed by, or sponsored by n8n GmbH. “n8n” is a trademark of n8n GmbH and is used here only to describe the platform-specific implementation and automation services we provide.