Morning Edition · Sunday, August 16, 2026Published at 2:12 AM EDT · New York
The 2.4-trillion-parameter sparse model activates 95 billion parameters per token and lists at $2 and $6 per million input and output tokens, well below comparable closed-model rates.

Alibaba's Qwen team released Qwen3.8-Max on August 3, a mixture-of-experts (MoE) model with 2.4 trillion total parameters and 95 billion active per forward pass. It natively handles 262,144 tokens of context, extensible to roughly one milli…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Part of a tracked trend
Chinese Open-Weight Models Emerge as the Non-US AI Stack
As Washington restricts foreign access to US frontier models, governments and enterprises cut off from American AI increasingly standardize on downloadable Chinese open-weight models, splitting the world into competing AI supply blocs rather than a single frontier.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.