Morning Edition · Saturday, July 25, 2026Published at 1:43 AM EDT · New York
One arXiv study proposes expert-aware contrast decoding to cut hallucinations in mixture-of-experts models, while another argues routing follows a frequency-diversity law resembling a compression code.

Two new arXiv preprints turn attention to the internals of mixture-of-experts (MoE) models, the architecture now standard in frontier systems including the largest Chinese open-weight releases. Both treat the router, the component that deci…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Part of a tracked trend
Frontier Model Efficiency Gains
Capability per unit of training and inference compute keeps improving, letting newer models match prior frontier performance far more cheaply and gradually loosening the link between raw scale and capability.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.