Datasets › OVIS

OVIS (Occluded Video Instance Segmentation)

Introduced by Jiyang Qi et al. in Occluded Video Instance Segmentation: A Benchmark2 Feb 2021 archive 2025-07-28

OVIS is a new large scale benchmark dataset for video instance segmentation task. It is designed with the philosophy of perceiving object occlusions in videos, which could reveal the complexity and the diversity of real-world scenes. OVIS consists of:

  • 296k high-quality instance masks
  • 25 commonly seen semantic categories
  • 901 videos with severe object occlusions
  • 5,223 unique instances

If the description or image is from a different paper, please refer to it as follows: Source: http://songbai.site/ovis/

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Video Instance Segmentation OVIS validation DVIS-DAQ(VIT-L, Offline) mask AP 57.1 DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries zhang-tao-whu/DVIS +2 44 Compare

Papers archive 2025-07-28

29 shown of 29 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 75. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Context-Aware Video Instance Segmentation 1 1 3 Jul 2024 not harvested
DVIS-DAQ: Improving Video Segmentation via Dynamic Anchor Queries 3 1 29 Mar 2024 ran 1 of 5 samples (4 unverified; 3 pointer-only for licence)
UniVS: Unified and Universal Video Segmentation with Prompts as Queries 1 1 28 Feb 2024 ran 12 of 14 samples (2 unverified; 14 pointer-only for licence)
DVIS++: Improved Decoupled Framework for Universal Video Segmentation 1 4 20 Dec 2023 not harvested
General Object Foundation Model for Images and Videos at Scale 1 1 14 Dec 2023 ran 8 of 13 samples (5 unverified)
NOVIS: A Case for End-to-End Near-Online Video Instance Segmentation 0 2 29 Aug 2023 not harvested
CTVIS: Consistent Training for Online Video Instance Segmentation 1 2 24 Jul 2023 ran 4 of 7 samples (3 unverified)
RefineVIS: Video Instance Segmentation with Temporal Attention Refinement 0 1 7 Jun 2023 not harvested
DVIS: Decoupled Video Instance Segmentation Framework 1 2 6 Jun 2023 not harvested
GRAtt-VIS: Gated Residual Attention for Auto Rectifying Video Instance Segmentation 1 2 26 May 2023 not harvested
BoxVIS: Video Instance Segmentation with Box Annotations 1 1 26 Mar 2023 not harvested
MDQE: Mining Discriminative Query Embeddings to Segment Occluded Instances on Challenging Videos 1 1 25 Mar 2023 ran 3 of 4 samples (1 unverified; 4 pointer-only for licence)
Tube-Link: A Flexible Cross Tube Framework for Universal Video Segmentation 1 1 22 Mar 2023 not harvested
Universal Instance Perception as Object Discovery and Retrieval 1 2 12 Mar 2023 ran 3 of 4 samples (1 unverified)
TarViS: A Unified Approach for Target-based Video Segmentation 1 3 6 Jan 2023 ran 0 of 6 samples (6 unverified)
Robust Online Video Instance Segmentation with Track Queries 1 1 16 Nov 2022 not harvested
A Generalized Framework for Video Instance Segmentation 1 1 16 Nov 2022 not harvested
InstanceFormer: An Online Video Instance Segmentation Framework 1 2 22 Aug 2022 not harvested
MinVIS: A Minimal Video Instance Segmentation Framework without Video-based Training 2 1 3 Aug 2022 not harvested
DeVIS: Making Deformable Transformers Work for Video Instance Segmentation 1 2 22 Jul 2022 ran 2 of 11 samples (9 unverified; 11 pointer-only for licence)
In Defense of Online Models for Video Instance Segmentation 2 2 21 Jul 2022 ran 3 of 5 samples (2 unverified)
VITA: Video Instance Segmentation via Object Token Association 1 1 9 Jun 2022 ran 0 of 7 samples (7 unverified)
Temporally Efficient Vision Transformer for Video Instance Segmentation 3 1 18 Apr 2022 ran 0 of 2 samples (2 unverified)
STC: Spatio-Temporal Contrastive Learning for Video Instance Segmentation 0 1 8 Feb 2022 not harvested
Mask2Former for Video Instance Segmentation 6 1 20 Dec 2021 ran 2 of 7 samples (5 unverified)
D2Conv3D: Dynamic Dilated Convolutions for Object Segmentation in Videos 2 1 15 Nov 2021 not harvested
Crossover Learning for Fast Online Video Instance Segmentation 1 2 13 Apr 2021 not harvested
Spatial Feature Calibration and Temporal Fusion for Effective One-stage Video Instance Segmentation 1 1 6 Apr 2021 not harvested
Occluded Video Instance Segmentation: A Benchmark 2 2 2 Feb 2021 not harvested

Dataset loaders archive 2025-07-28

1 loader as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Creative Commons Attribution-NonCommercial-ShareAlike 4.0 License

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • OVIS
  • OVIS validation

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections