Morning Edition · Friday, July 31, 2026Published at 1:48 AM EDT · New York
Google DeepMind Ships Gemini Robotics ER 2 as a Reasoning Layer for Multi-Robot Tasks
The model reports 91.3 percent accuracy on finding the exact frame of a critical event in robot video and streams reasoning while robots keep acting.

Google DeepMind released Gemini Robotics ER 2, an embodied-reasoning model that acts as a high-level planner. It interprets the physical scene, breaks multi-step tasks into parts, and delegates motor execution to a lower-level vision-langua…
Continue the AI Intelligence Brief
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
- 5 AI intelligence signals a day
- Frontier labs, compute, and chips
- Model releases and AI infrastructure
- Source-grounded analysis with confidence labels
The Global Intelligence Brief stays free.
Part of a tracked trend
Robotics Foundation Models for Embodied AI
Over the coming months, labs ship general-purpose robotics model suites that bridge vision-language understanding to physical navigation and manipulation, pushing foundation models into embodied action.
More from this edition
- OpenAI Cuts GPT-5.6 Luna Prices About 80 Percent, Turning the Model Race Into a Price War
- Anthropic's Claude Opus 5 Posts 96 Percent on SWE-bench Verified at Unchanged Opus Pricing
- US Regulator Bars New Foreign-Made Humanoid Robots, Citing Supply-Chain Security
- Google DeepMind Disbands the Original AlphaFold Team and Folds Science Into Gemini
- Meta Turns Muse Spark Into a Paid API and Zuckerberg Argues for Faster, Not Slower, AI
- GPTZero Flags Fabricated Citations in Four PwC Middle East Reports
- Preprint Probes Why Reinforcement Learning Beats Supervised Tuning on Math Reasoning
- Study Finds LLM Agents Deceive More Under Hidden, Conflicting Objectives
- New Jailbreak Uses Dual-Layer Encoding to Reconstruct Blocked Prompts Past Moderation
- ClinLens Benchmark Pushes Coding Agents Toward Long-Horizon Clinical Data Science
- ByteDance Readies Seedance 2.5 With Single-Pass 30-Second Video and Region-Level Editing