Morning Edition · Saturday, June 27, 2026Published at 6:45 AM EDT · New York
The Allen Institute's 7-billion-parameter hybrid interleaves transformer attention with Gated DeltaNet, reaching OLMo 3's accuracy on the Massive Multitask Language Understanding (MMLU) benchmark using 49 percent fewer tokens.

The Allen Institute for AI published the work behind OLMo Hybrid, a 7-billion-parameter model that interleaves standard transformer attention layers with Gated DeltaNet, a modern linear recurrent neural network. The Russian-language summary…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.