Morning Edition · Monday, June 22, 2026Published at 6:48 AM EDT · New York
A vision-language model that predicts where marked points will move in world coordinates aims to give robot planners and video generators a reusable, learned model of how objects move (a motion prior).
The Allen Institute for AI (Ai2) released MolmoMotion, a language-guided model that forecasts the future trajectories of points attached to objects within a 3D world frame. Given a video frame, a set of 3D points marked on an object, and a…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Part of a tracked trend
Robotics Foundation Models for Embodied AI
Over the coming months, labs ship general-purpose robotics model suites that bridge vision-language understanding to physical navigation and manipulation, pushing foundation models into embodied action.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.