Morning Edition · Monday, June 22, 2026Published at 6:48 AM EDT · New York
The model detects, segments, and tracks every instance of a concept from a phrase like "yellow school bus," roughly doubling prior systems on Meta's open-vocabulary benchmark.

Meta's Segment Anything Model 3 (SAM 3) shifts the family from pixel-level prompting to concept-level prompting. Where SAM and SAM 2 required a manual click or box per instance, SAM 3 takes an open-vocabulary text description and segments a…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Part of a tracked trend
Open-Vocabulary, Promptable Vision Foundation Models
Vision foundation models shift to text-promptable, open-vocabulary detection, segmentation, and real-time tracking of arbitrary concepts, generalizing perception beyond fixed label sets across images and video and pushing open perception models toward production use.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.