Polylog
The Polylog AI Intelligence Brief

Morning Edition · Monday, July 27, 2026Published at 1:32 AM EDT · New York

SenseTime Open-Sources SenseNova-Vision, a Unified Perception Model

A single model handles detection, segmentation, depth, keypoints, and 3D reconstruction as multimodal generation, replacing the usual collection of task-specific components.

SenseTime Open-Sources SenseNova-Vision, a Unified Perception Model

SenseTime has fully open-sourced SenseNova-Vision, a model that reframes a range of perception tasks as unified multimodal generation. Those tasks include object detection, optical character recognition, segmentation, depth estimation, keyp…

Continue the AI Intelligence Brief

Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.

  • 5 AI intelligence signals a day
  • Frontier labs, compute, and chips
  • Model releases and AI infrastructure
  • Source-grounded analysis with confidence labels

The Global Intelligence Brief stays free.

Part of a tracked trend

Open-Vocabulary, Promptable Vision Foundation Models

Vision foundation models shift to text-promptable, open-vocabulary detection, segmentation, and real-time tracking of arbitrary concepts, generalizing perception beyond fixed label sets across images and video and pushing open perception models toward production use.