Morning Edition · Monday, September 14, 2026Published at 2:22 AM EDT · New York
Researchers at the University of Pittsburgh say running DINOv3 and the Segment Anything Model (SAM) locally is what lets the system perceive in real time without depending on network connectivity.

Meta published an account of how the Human Engineering Research Laboratories at the University of Pittsburgh, a group affiliated with the United States Department of Veterans Affairs, are using its open vision models in assistive robotics.…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Part of a tracked trend
Open-Vocabulary, Promptable Vision Foundation Models
Vision foundation models shift to text-promptable, open-vocabulary detection, segmentation, and real-time tracking of arbitrary concepts, generalizing perception beyond fixed label sets across images and video and pushing open perception models toward production use.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.