Browse State-of-the-Art › One-shot visual object segmentation
One-shot visual object segmentation
26 papers with code · 2 benchmarks · 1 dataset archive 2025-07-28
Benchmarks archive 2025-07-28
2 leaderboard tables shown for this task, 2 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| YouTube-VOS 2018 (2 rows) | OSVOS | One-Shot Video Object Segmentation | code | — | Compare |
| YouTube-VOS (1 row) | RVOS-Mask-ST+ | RVOS: End-to-End Recurrent Network for Video Object Segmentation | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
26 shown of 26 papers with code (41 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
24 Jul 2018 5 repositories listed Syntology ran 3 of 15 samples · 12 unverified · 3 pointer-only (licence)We address semi-supervised video object segmentation, the task of automatically generating accurate and consistent pixel masks for objects in a video sequence, given the first-frame ground truth annotations.
-
3 Dec 2020 4 repositories listedIn the semi-supervised setting, the first mask of each object is provided at test time.
-
3 Sep 2018 4 repositories listedEnd-to-end sequential learning to explore spatial-temporal features for video segmentation is largely limited by the scale of available video segmentation datasets, i.
-
1 Apr 2019 3 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedIn our framework, the past frames with object masks form an external memory, and the current frame as the query is segmented using the mask information in the memory.
-
4 Jun 2021 2 repositories listedThe state-of-the-art methods learn to decode features with a single positive object and thus have to match and segment each target separately under multi-object scenarios, consuming multiple times computing resources.
-
25 Mar 2020 2 repositories listedThis allows us to achieve a rich internal representation of the target in the current frame, significantly increasing the segmentation accuracy of our approach.
-
18 Mar 2020 2 repositories listedThis paper investigates the principles of embedding learning to tackle the challenging semi-supervised video object segmentation.
-
27 Feb 2020 2 repositories listedThe target appearance model consists of a light-weight module, which is learned during the inference stage using fast optimization techniques to predict a coarse but robust target segmentation.
-
8 May 2019 2 repositories listed Syntology ran 1 of 4 samples · 3 unverifiedThen the synthesized flow field is used to guide the propagation of pixels to fill up the missing regions in the video.
-
1 Jun 2021 1 repository listedRecently, Space-Time Memory Network (STM) based methods have achieved state-of-the-art performance in semi-supervised video object segmentation (VOS).
-
24 Mar 2021 1 repository listed Syntology ran 5 of 10 samples · 5 unverifiedFor the current query frame, the query regions are tracked and predicted based on the optical flow estimated from the previous frame.
-
18 Feb 2021 1 repository listedSpecifically, we first compute a pixel-wise similarity matrix by using representations of reference and target pixels and then select top-rank reference pixels for target pixel classification.
-
21 Jan 2021 1 repository listed Syntology ran 5 of 7 samples · 2 unverified · 7 pointer-only (licence)SST extracts per-pixel representations for each object in a video using sparse attention over spatiotemporal features.
-
21 Dec 2020 1 repository listedCurrent state-of-the-art approaches for Semi-supervised Video Object Segmentation (Semi-VOS) propagates information from previous frames to generate segmentation mask for the current frame.
-
23 Oct 2020 1 repository listedIn this paper, we address several inadequacies of current video object segmentation pipelines.
-
13 Oct 2020 1 repository listed Syntology ran 0 of 2 samples · 2 unverifiedThis paper investigates the principles of embedding learning to tackle the challenging semi-supervised video object segmentation.
-
10 Oct 2020 1 repository listedIn this work, we study an RNN-based architecture and address some of these issues by proposing a hybrid sequence-to-sequence architecture named HS2S, utilizing a dual mask propagation strategy that allows incorporating…
-
1 Aug 2020 1 repository listedWe propose a unified referring video object segmentation network (URVOS).
-
26 May 2020 1 repository listedWe treat this as a grouping problem by exploiting object proposals and making a joint inference about grouping over both space and time.
-
17 Feb 2020 1 repository listedWe propose a directional deep embedding and appearance learning (DDEAL) method, which is free of the online fine-tuning process, for fast VOS.
-
1 Oct 2019 1 repository listedIn this paper, we propose AGSS-VOS to segment multiple objects in one feed-forward path via instance-agnostic and instance-specific modules.
-
30 Sep 2019 1 repository listedIn this work we propose a capsule-based approach for semi-supervised video object segmentation.
-
27 Sep 2019 1 repository listedIn practice, it performs similarly to the Hungarian algorithm during inference.
-
2 Jul 2019 1 repository listedVideo object segmentation (VOS) aims at pixel-level object tracking given only the annotations in the first frame.
-
13 Mar 2019 1 repository listedMultiple object video object segmentation is a challenging task, specially for the zero-shot case, when no object mask is given at the initial frame and the model has to find the objects to be segmented along the…
-
28 Nov 2018 1 repository listed Syntology ran 2 of 3 samples · 1 unverified · 3 pointer-only (licence)One of the fundamental challenges in video object segmentation is to find an effective representation of the target and background appearance.
Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections