Papers › ALBA : Reinforcement Learning for Video Object Segmentation

ALBA : Reinforcement Learning for Video Object Segmentation

26 May 2020arXiv:2005.13039archive 2025-07-28

Shreyank N Gowda, Panagiotis Eustratiadis, Timothy Hospedales, Laura Sevilla-Lara

We consider the challenging problem of zero-shot video object segmentation (VOS). That is, segmenting and tracking multiple moving objects within a video fully automatically, without any manual initialization. We treat this as a grouping problem by exploiting object proposals and making a joint inference about grouping over both space and time. We propose a network architecture for tractably performing proposal selection and joint grouping. Crucially, we then show how to train this network with reinforcement learning so that it learns to perform the optimal non-myopic sequence of grouping decisions to segment the whole video. Unlike standard supervised techniques, this also enables us to directly optimize for the non-differentiable overlap-based metrics used to evaluate VOS. We show that the proposed method, which we call ALBA outperforms the previous stateof-the-art on three benchmarks: DAVIS 2017 [2], FBMS [20] and Youtube-VOS [27].

PaperPDFCode

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

kini5gowda/ALBA-RL-for-VOS mentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

ObjectOne-shot visual object segmentationReinforcement LearningReinforcement Learning (RL)Semantic SegmentationUnsupervised Video Object SegmentationVideo Object SegmentationVideo Semantic Segmentationreinforcement-learning

1 archive task tag without a task page not shown.

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Unsupervised Video Object Segmentation DAVIS 2017 (val) ALBA F-measure (Mean) 60.2 #7 of 10 Archive leaderboard report
Unsupervised Video Object Segmentation DAVIS 2017 (val) ALBA F-measure (Recall) 63.1 #7 of 10 Archive leaderboard report
Unsupervised Video Object Segmentation DAVIS 2017 (val) ALBA J&F 58.4 #7 of 10 Archive leaderboard report
Unsupervised Video Object Segmentation DAVIS 2017 (val) ALBA Jaccard (Mean) 56.6 #7 of 10 Archive leaderboard report
Unsupervised Video Object Segmentation DAVIS 2017 (val) ALBA Jaccard (Recall) 63.4 #7 of 10 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections