Our #CVPR2026 poster starts in 30 minutes! Come by Poster Session 2, #333 to learn about VideoCUPS — the first unsupervised video panoptic segmentation method, learning to detect, segment, and track objects in video without human supervision.
📢 [CVPR’26] Can we learn to detect, segment, and track every object in a video without human supervision? Yes, we introduce VideoCUPS, the first unsupervised video panoptic segmentation (VPS) method: 1. Get pseudo-labels from monocular videos. 2. Train a VPS model on them.