Morning Edition · Tuesday, July 28, 2026Published at 1:31 AM EDT · New York
Moonshot AI Releases Kimi K3, a 2.8-Trillion-Parameter Open-Weight Model That Tops Open Rankings
The Beijing lab's mixture-of-experts system trails only the closed frontier of Claude Fable 5 and GPT-5.6 Sol while beating Claude Opus 4.8 on coding and agent benchmarks, and its weights are a 1.4-terabyte free download.

Moonshot AI, the Beijing startup backed by Alibaba, has published the open weights for Kimi K3, a 2.8-trillion-parameter model it calls the largest ever released for download. The weights went live on Hugging Face on July 26, a day ahead of schedule. The design is a mixture-of-experts model with a context window of roughly one million tokens and an always-on reasoning mode. And, as VentureBeat reported, only about 32 billion parameters activate for each token, which holds the cost of running the model far below what its full size suggests.
Moonshot's own evaluations place K3 behind the closed frontier of Anthropic's Claude Fable 5 and OpenAI's GPT-5.6 Sol, but ahead of every other model it tested, including Claude Opus 4.8 and GPT-5.5, on coding and agentic suites. Tom's Hardware reported that K3 beat Claude Fable 5 on the Frontend Code Arena benchmark. A widely circulated AI Post summary put K3 at the top of the open-weight tier of Artificial Analysis's Intelligence Index with a score of 57, a figure that reflects a single aggregator rather than independent reproduction of each named benchmark.
The practical constraint is deployment size. The full model is roughly 1.4 terabytes even when compressed to a lower-precision format (MXFP4 quantization), which makes self-hosting impractical for most teams and pushes real usage toward hosted endpoints and cloud providers. The distance between a model that can be downloaded and one that can actually be run is now the central economic question for open weights at this scale.
- If true, who benefits
Moonshot and Beijing gain a credibility win for China's open-weight strategy and Global South AI adoption, while enterprises and cloud hosts serving downloadable Chinese weights gain a cheaper alternative to metered United States APIs.
- The nuance
The "tops open rankings" claim rests on Moonshot's own evaluations and a single aggregator score rather than independent per-benchmark reproduction, and Anthropic's separate allegation that K3 was built partly by distilling Claude, which Moonshot denies, is the load-bearing nuance about whether the capability is fully original.
An open-source-intelligence read of how likely this story is true with its real nuance, not a judgment of any outlet. It assesses the claim, weighing independent and adversarial reporting. How we label confidence.
What this means
The open-weight tier now includes a model that beats a two-month-old closed frontier model from a leading United States lab, which shortens the period during which a closed lab can charge a premium for capability alone. The exposed parties are sellers of metered application programming interfaces (APIs) whose pricing assumes a durable capability lead, and the beneficiaries are enterprises and government buyers who can now build a stack on downloadable Chinese weights. The 1.4-terabyte footprint means the near-term beneficiaries on the serving side are cloud hosts, not individual machines.
What to watch
- Independent reproduction of the coding and agent numbers on Artificial Analysis, LMArena, and SWE-bench, since a single aggregator score is not the same as verified parity on each benchmark.
- How fast third-party inference providers create cheap hosted K3 endpoints, which determines whether the release is used in production or remains only a published specification.
Observations to monitor, not financial advice.
Source: Polylog editors
Part of a tracked trend
Open-Weight Models Close the Gap With Closed Frontier Labs
Over the next 3-9 months, open-weight releases with downloadable weights, long context, and strong agentic/coding performance increasingly match closed frontier models on practical work, eroding the closed-lab moat.
More from this edition
- China Accuses Washington of 'AI Hegemonism' and Threatens Countermeasures Over Moonshot Probe
- Anthropic Says It Never Sought to Ban Open Weights, but Backs Distillation Crackdowns and Pre-Release Testing
- Anthropic Ships Claude Opus 5, Its Fourth Model in Two Months, With a Doubling on Software-Engineering Evals
- Microsoft Says United Kingdom Grid Connections Take Eight Years, Choking AI Data-Center Buildout
- Nvidia Puts Its Vera CPU to Work Designing the Next Generation of CPUs and GPUs
- Hassabis Says DeepMind Sold to Google Because Independence Would Have Cost Billions It Could Not Raise
- A 184-Million-Parameter Classifier Claims State-of-the-Art Prompt-Injection Detection at a Fraction of Llama Guard's Size
- CORVUS Attacks the Coding Agent's Real Bottleneck: an Append-Only Trajectory That Bloats Context
- FlowEvo Lets Agents Rewrite Their Own Workflows and Skills Instead of Rebuilding Them Each Run
- Meta Puts Segment Anything and DINO Into Assistive Robotics at the University of Pittsburgh
- Vals Launches a Tool That Turns Your Own Codebase Into a Custom Model Benchmark