Papers › Region Aware Video Object Segmentation with Deep Motion Modeling

Region Aware Video Object Segmentation with Deep Motion Modeling

21 Jul 2022arXiv:2207.10258archive 2025-07-28

Bo Miao, Mohammed Bennamoun, Yongsheng Gao, Ajmal Mian

Current semi-supervised video object segmentation (VOS) methods usually leverage the entire features of one frame to predict object masks and update memory. This introduces significant redundant computations. To reduce redundancy, we present a Region Aware Video Object Segmentation (RAVOS) approach that predicts regions of interest (ROIs) for efficient object segmentation and memory storage. RAVOS includes a fast object motion tracker to predict their ROIs in the next frame. For efficient segmentation, object features are extracted according to the ROIs, and an object decoder is designed for object-level segmentation. For efficient memory storage, we propose motion path memory to filter out redundant context by memorizing the features within the motion path of objects between two frames. Besides RAVOS, we also propose a large-scale dataset, dubbed OVOS, to benchmark the performance of VOS models under occlusions. Evaluation on DAVIS and YouTube-VOS benchmarks and our new OVOS dataset show that our method achieves state-of-the-art performance with significantly faster inference time, e.g., 86.1 J&F at 42 FPS on DAVIS and 84.4 J&F at 23 FPS on YouTube-VOS.

PaperPDF

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

DecoderObjectSegmentationSemantic SegmentationSemi-Supervised Video Object SegmentationVideo Object SegmentationVideo Semantic Segmentation

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Semi-Supervised Video Object Segmentation DAVIS 2016 RAVOS F-measure (Mean) 92.6 #17 of 78 Archive leaderboard report
Semi-Supervised Video Object Segmentation DAVIS 2016 RAVOS J&F 91.7 #17 of 78 Archive leaderboard report
Semi-Supervised Video Object Segmentation DAVIS 2016 RAVOS Jaccard (Mean) 90.8 #17 of 78 Archive leaderboard report
Semi-Supervised Video Object Segmentation DAVIS 2016 RAVOS Speed (FPS) 58 #17 of 78 Archive leaderboard report
Semi-Supervised Video Object Segmentation DAVIS 2017 (val) RAVOS F-measure (Mean) 89.3 #18 of 81 Archive leaderboard report
Semi-Supervised Video Object Segmentation DAVIS 2017 (val) RAVOS J&F 86.1 #18 of 81 Archive leaderboard report
Semi-Supervised Video Object Segmentation DAVIS 2017 (val) RAVOS Jaccard (Mean) 82.9 #18 of 81 Archive leaderboard report
Semi-Supervised Video Object Segmentation DAVIS 2017 (val) RAVOS Speed (FPS) 42 (on 3090) #18 of 81 Archive leaderboard report
Semi-Supervised Video Object Segmentation YouTube-VOS 2018 RAVOS F-Measure (Seen) 87.8 #20 of 53 Archive leaderboard report
Semi-Supervised Video Object Segmentation YouTube-VOS 2018 RAVOS F-Measure (Unseen) 87.4 #20 of 53 Archive leaderboard report
Semi-Supervised Video Object Segmentation YouTube-VOS 2018 RAVOS Jaccard (Seen) 83.1 #20 of 53 Archive leaderboard report
Semi-Supervised Video Object Segmentation YouTube-VOS 2018 RAVOS Jaccard (Unseen) 79.1 #20 of 53 Archive leaderboard report
Semi-Supervised Video Object Segmentation YouTube-VOS 2018 RAVOS Overall 84.4 #20 of 53 Archive leaderboard report
Semi-Supervised Video Object Segmentation YouTube-VOS 2018 RAVOS Speed (FPS) 23 #20 of 53 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

AWAREVOS

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections