# Moonshot's Kimi K3 Becomes the Largest Open-Weight Model, but Abandons Discount Pricing

The 2.8-trillion-parameter Chinese model scores 93.5% on GPQA Diamond and prices output at $15 per million tokens, roughly matching US flagships rather than undercutting them.

- Published: 2026-07-26T05:32:56.369Z
- Canonical: https://polylog.news/ai/2026-07-26/moonshot-s-kimi-k3-becomes-the-largest-open-weight-model-but
- Publisher: Polylog (AI desk)
- Section: tech
- Sources: [Polylog editors](https://polylog.news)

Moonshot AI released [Kimi K3](https://venturebeat.com/technology/chinas-moonshot-ai-releases-kimi-k3-the-largest-open-source-model-ever-rivaling-top-u-s-systems) on July 16, a 2.8-trillion-parameter multimodal reasoning model with a 1-million-token context window and downloadable weights expected around July 27. Independent trackers report it as the largest open-weight model shipped to date, with a GPQA Diamond score of 93.5%, the strongest reported for any open-weight system.

The pricing complicates the common view that Chinese labs always compete on lower cost. Kimi K3's API lists at [$3 per million input tokens and $15 per million output tokens](https://openrouter.ai/moonshotai/kimi-k3), well above the sub-dollar rates of DeepSeek and Qwen and described by several analysts as the most expensive model any Chinese lab has released. The "roughly half the price" framing holds on a per-task basis. Artificial Analysis measured K3 at about $0.94 per task on its intelligence index versus roughly $1.80 for Claude Opus 4.8, [an efficiency edge](https://the-decoder.com/kimis-open-model-k3-nears-gpt-5-6-sol-and-fable-5-while-signaling-the-end-of-super-cheap-chinese-ai/) that comes from architecture and throughput rather than a list-price discount. Independent testing by developer Simon Willison found K3 consumes heavy reasoning-token budgets, around 13,000 tokens for a single generation task.

## What this means

The signal is that Chinese frontier labs are moving up-market from the discount tier toward parity pricing once their capability justifies it. That changes the mechanism of the China-stack thesis. Adoption abroad will increasingly depend on downloadable weights and sovereignty rather than on a price gap that K3 no longer offers. Enterprises that standardized on cheap Chinese APIs to cut cost now face a two-track choice between low-cost older models and a flagship priced like Western incumbents.

## What to watch

- The open-weight release around July 27 and how quickly self-hosted K3 deployments appear, which would test whether the 2.8-trillion-parameter model is practical to run outside Moonshot's own infrastructure.
- Whether other Chinese labs follow Moonshot toward flagship pricing, which would mark the end of discount positioning as the region's main competitive advantage.
