Morning Edition · Sunday, July 12, 2026Published at 1:29 AM EDT · New York
Researchers from Fudan University and Shanghai say the framework holds near single-object speed and memory as the number of tracked targets grows.
A team from Fudan University and Shanghai introduced SAM-MT, an interactive multi-target video object segmentation framework, according to a research digest circulated by AI with Papers. The stated goal is to keep frames-per-second throughp…
Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.
The Global Intelligence Brief stays free.
Part of a tracked trend
Open-Vocabulary, Promptable Vision Foundation Models
Vision foundation models shift to text-promptable, open-vocabulary detection, segmentation, and real-time tracking of arbitrary concepts, generalizing perception beyond fixed label sets across images and video and pushing open perception models toward production use.
Start a discussion in Townsquare.
More from this edition
Comments
0No comments yet.