{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/flow-guided-recurrent-neural-encoder-for","title":"Flow Guided Recurrent Neural Encoder for Video Salient Object Detection","arxiv_id":null,"date":"2018-06-01","proceeding":"CVPR 2018 6","authors":["Guanbin Li","Yuan Xie","Tianhao Wei","Keze Wang","Liang Lin"],"abstract":"Image saliency detection has recently witnessed significant progress due to deep convolutional neural networks. However, extending state-of-the-art saliency detectors from image to video is challenging. The performance of salient object detection suffers from object or camera motion and the dramatic change of the appearance contrast in videos. In this paper, we present flow guided recurrent neural encoder(FGRNE), an accurate and end-to-end learning framework for video salient object detection. It works by enhancing the temporal coherence of the per-frame feature by exploiting both motion information in terms of optical flow and sequential feature evolution encoding in terms of LSTM networks. It can be considered as a universal framework to extend any FCN based static saliency detector to video salient object detection. Intensive experimental results verify the effectiveness of each part of FGRNE and confirm that our proposed method significantly outperforms state-of-the-art methods on the public benchmarks of DAVIS and FBMS.","url_abs":"http://openaccess.thecvf.com/content_cvpr_2018/html/Li_Flow_Guided_Recurrent_CVPR_2018_paper.html","url_pdf":"http://openaccess.thecvf.com/content_cvpr_2018/papers/Li_Flow_Guided_Recurrent_CVPR_2018_paper.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"object","task_name":"Object"},{"task_slug":"object-detection","task_name":"Object Detection"},{"task_slug":"optical-flow-estimation","task_name":"Optical Flow Estimation"},{"task_slug":"salient-object-detection","task_name":"RGB Salient Object Detection"},{"task_slug":"saliency-detection","task_name":"Saliency Detection"},{"task_slug":"salient-object-detection-1","task_name":"Salient Object Detection"},{"task_slug":"video-salient-object-detection","task_name":"Video Salient Object Detection"},{"task_slug":"object-detection-1","task_name":"object-detection"}],"methods":[{"method_slug":"convolution","method_name":"Convolution"},{"method_slug":"fcn","method_name":"FCN"},{"method_slug":"lstm","method_name":"LSTM"},{"method_slug":"max-pooling","method_name":"Max Pooling"},{"method_slug":"sigmoid-activation","method_name":"Sigmoid Activation"},{"method_slug":"tanh-activation","method_name":"Tanh Activation"}],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/video-salient-object-detection-on-davis-2016","task":"Video Salient Object Detection","dataset":"DAVIS-2016","model":"FGRN","rank_in_archive_order":7,"of":11,"metrics":{"AVERAGE MAE":"0.043","MAX E-MEASURE":"0.917","MAX F-MEASURE":"0.783","S-Measure":"0.838"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-davsod-2","task":"Video Salient Object Detection","dataset":"DAVSOD-Difficult20","model":"FGRN","rank_in_archive_order":2,"of":8,"metrics":{"Average MAE":"0.131","S-Measure":"0.608","max E-measure":"0.698"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-davsod-1","task":"Video Salient Object Detection","dataset":"DAVSOD-Normal25","model":"FGRN","rank_in_archive_order":3,"of":8,"metrics":{"Average MAE":"0.126","S-Measure":"0.638","max E-measure":"0.700"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-davsod","task":"Video Salient Object Detection","dataset":"DAVSOD-easy35","model":"FGRN","rank_in_archive_order":4,"of":9,"metrics":{"Average MAE":"0.095","S-Measure":"0.701","max E-Measure":"0.765","max F-Measure":"0.589"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-fbms-59","task":"Video Salient Object Detection","dataset":"FBMS-59","model":"FGRN","rank_in_archive_order":7,"of":16,"metrics":{"AVERAGE MAE":"0.088","MAX E-MEASURE":"0.863","MAX F-MEASURE":"0.767","S-Measure":"0.809"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-mcl","task":"Video Salient Object Detection","dataset":"MCL","model":"FGRN","rank_in_archive_order":4,"of":8,"metrics":{"AVERAGE MAE":"0.044","MAX E-MEASURE":"0.817","MAX F-MEASURE":"0.625","S-Measure":"0.709"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-uvsd","task":"Video Salient Object Detection","dataset":"UVSD","model":"FGRN","rank_in_archive_order":3,"of":8,"metrics":{"Average MAE":"0.042","S-Measure":"0.745","max E-measure":"0.887"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-vos-t","task":"Video Salient Object Detection","dataset":"VOS-T","model":"FGRN","rank_in_archive_order":5,"of":9,"metrics":{"Average MAE":"0.097","S-Measure":"0.715","max E-measure":"0.797"},"uses_additional_data":true},{"leaderboard":"/sota/video-salient-object-detection-on-visal","task":"Video Salient Object Detection","dataset":"ViSal","model":"FGRN","rank_in_archive_order":5,"of":10,"metrics":{"Average MAE":"0.045","S-Measure":"0.861","max E-measure":"0.945"},"uses_additional_data":true}],"syntology":{"syntology_url":null,"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}