Browse State-of-the-Art › Video Saliency Prediction
Video Saliency Prediction
15 papers with code · 0 benchmarks · 3 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
3 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
15 shown of 15 papers with code (29 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
3 Jul 2019 2 repositories listedThis paper investigates modifying an existing neural network architecture for static saliency prediction using two types of recurrences that integrate information from the temporal domain.
-
25 May 2019 2 repositories listedOur results suggest that (1) audio is a strong contributing cue for saliency prediction, (2) salient visible sound-source is the natural cause of the superiority of our Audio-Visual model, (3) richer feature…
-
28 Aug 2018 2 repositories listedThis work adapts a deep neural model for image saliency prediction to the temporal domain of egocentric video.
-
23 Sep 2024 1 repository listedThe goal of the participants was to develop a method for predicting accurate saliency maps for the provided set of video sequences.
-
24 Aug 2023 1 repository listedThe growing interest in omnidirectional videos (ODVs) that capture the full field-of-view (FOV) has gained 360-degree saliency prediction importance in computer vision.
-
11 Jan 2023 1 repository listedVideo saliency prediction has recently attracted attention of the research community, as it is an upstream task for several practical applications.
-
9 Jun 2022 1 repository listedWe show that gaze direction and affective representations contribute a prediction to ground-truth correspondence improvement of at least 5% compared to dynamic saliency models without social cues.
-
24 Aug 2021 1 repository listed3D convolutional neural networks have achieved promising results for video tasks in computer vision, including video saliency prediction that is explored in this paper.
-
16 Apr 2021 1 repository listedWe note that the accuracy of the maps reconstructed from the gaze data of a fixed number of observers varies with the frame, as it depends on the content of the scene.
-
11 Dec 2020 1 repository listedWe also explore a variation of ViNet architecture by augmenting audio features into the decoder.
-
2 Oct 2020 1 repository listedWhen the base hierarchical model is empowered with domain-specific modules, performance improves, outperforming state-of-the-art models on three out of five metrics on the DHF1K benchmark and reaching the second-best…
-
2 Jan 2020 1 repository listedDue to a variety of motions across different frames, it is highly challenging to learn an effective spatiotemporal representation for accurate video saliency prediction (VSP).
-
1 Sep 2018 1 repository listedHence, an object-to-motion convolutional neural network (OM-CNN) is developed to predict the intra-frame saliency for DeepVS, which is composed of the objectness and motion subnets.
-
13 Mar 2018 1 repository listedOur model starts with a rough segmentation and quantifies several intuitive observations such as the effects of visual discomfort level, depth abruptness, motion acceleration, elements of surprise, size and compactness…
-
19 Sep 2017 1 repository listedWe further find from our database that there exists a temporal correlation of human attention with a smooth saliency transition across video frames.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections