Morning Edition · Tuesday, July 21, 2026Published at 1:32 AM EDT · New York
A New Router Aims to Make Mixture-of-Experts Selection More Consistent
The method conditions expert routing on multi-level context rather than shallow per-token representations, targeting a known instability in sparse models.

A new arXiv paper targets a weakness in how mixture-of-experts (MoE) models route tokens. MoE scales transformers efficiently by sending each token to a small subset of experts, but the authors note that existing routers typically decide ba…
Continue the AI Intelligence Brief
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
- 5 AI intelligence signals a day
- Frontier labs, compute, and chips
- Model releases and AI infrastructure
- Source-grounded analysis with confidence labels
The Global Intelligence Brief stays free.
Part of a tracked trend
Frontier Model Efficiency Gains
Capability per unit of training and inference compute keeps improving, letting newer models match prior frontier performance far more cheaply and gradually loosening the link between raw scale and capability.
More from this edition
- Google Is Building a Chip That Bakes Gemini's Architecture Into Silicon
- OpenAI Paused an Internal Model After It Found and Exploited a Sandbox Flaw
- Another Chinese Open-Weight Model Reaches the Frontier at a Third of the Price
- An Autonomous AI Agent Breached Hugging Face and Logged 17,000 Actions
- An Anthropic Researcher Says Claude Fable 5 Produced a Counterexample to the Jacobian Conjecture
- Nvidia Puts AI Agents to Work Building Simulation Worlds at SIGGRAPH
- Bristol Myers Squibb Commits to an Nvidia Vera Rubin AI Factory for Drug Discovery
- Meta's Brain2Qwerty Decodes Typed Sentences From Brain Scans at 61% Word Accuracy
- A Paper Finds Open-Weight Models Commit to Answers Before They Reason
- Researchers Warn That RLHF Preference Data Encodes the Rater's State, Not Just the Output
- Roblox Will Let Players Generate a Playable Game From a Text Prompt