Home Knowledge Base Video instance segmentation (VIS)

Video instance segmentation (VIS) is the task of segmenting each object instance in every frame while maintaining consistent identity across time - it unifies detection, pixel-wise masking, and tracking into one temporal perception problem.

What Is Video Instance Segmentation?

Why VIS Matters

VIS Pipeline Components

Per-Frame Instance Proposal:

Temporal Association:

Mask Refinement:

How It Works

Step 1:

Step 2:

Video instance segmentation is a high-resolution temporal perception task that tracks who is where and with what shape through time - it is a cornerstone capability for advanced video scene intelligence.

video instance segmentationvideo understanding

Explore 500+ Semiconductor & AI Topics

From EUV lithography to CUDA optimization — search the full knowledge base or chat with our AI assistant.