Morning Edition · Friday, July 24, 2026Published at 1:33 AM EDT · New York
One frames expert selection as a frequency-driven code over chains of thought, the other uses expert-aware contrastive decoding to reduce hallucination.

Two arXiv preprints examine the routing logic inside Mixture-of-Experts (MoE) models, the architecture behind most efficient frontier systems, where a gate sends each token to a small subset of experts. The first, asking whether MoE routing…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Part of a tracked trend
Frontier Model Efficiency Gains
Capability per unit of training and inference compute keeps improving, letting newer models match prior frontier performance far more cheaply and gradually loosening the link between raw scale and capability.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.