Browse State-of-the-Art › Saliency Prediction
Saliency Prediction
105 papers with code · 4 benchmarks · 9 datasets archive 2025-07-28
A saliency map is a model that predicts eye fixations on a visual scene. Saliency prediction is informed by the human visual attention mechanism and predicts the possibility of the human eyes to stay in a certain position in the scene.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
4 leaderboard tables shown for this task, 4 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| SALECI (5 rows) | SUM | SUM: Saliency Unification through Mamba for Visual Attention Modeling | code | Syntology ran 14 of 14 samples · 0 unverified | Compare |
| SALICON (5 rows) | SUM | SUM: Saliency Unification through Mamba for Visual Attention Modeling | code | Syntology ran 14 of 14 samples · 0 unverified | Compare |
| MIT300 (2 rows) | SUM | SUM: Saliency Unification through Mamba for Visual Attention Modeling | code | Syntology ran 14 of 14 samples · 0 unverified | Compare |
| CAT2000 (1 row) | SUM | SUM: Saliency Unification through Mamba for Visual Attention Modeling | code | Syntology ran 14 of 14 samples · 0 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
9 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
2 subtasks in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 105 papers with code (268 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
7 Sep 2020 4 repositories listedOur framework includes two main models: 1) a generator model, which maps the input image and latent variable to stochastic saliency prediction, and 2) an inference model, which gradually updates the latent variable by…
-
4 Feb 2020 4 repositories listedIn this paper, we introduce and tackle the simultaneous enhancement and super-resolution (SESR) problem for underwater robot vision and provide an efficient solution for near real-time applications.
-
18 Feb 2019 4 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedTo develop robust representations for this challenging task, high-level visual features at multiple spatial scales must be extracted and augmented with contextual information.
-
4 Jan 2017 4 repositories listed Syntology ran 2 of 6 samples · 4 unverifiedWe introduce SalGAN, a deep convolutional neural network for visual saliency prediction trained with adversarial examples.
-
18 Aug 2021 3 repositories listed Syntology ran 8 of 14 samples · 6 unverified · 14 pointer-only (licence)To effectively fuse cross-modal features in the shared learning network, we propose a cross-enhanced integration module (CIM) and then propagate the fused feature to the next layer for integrating cross-level…
-
2 Apr 2020 3 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)We also present a benchmark evaluation of state-of-the-art semantic segmentation approaches based on standard performance metrics.
-
1 Jun 2019 3 repositories listedIn this paper, we propose a predict-refine architecture, BASNet, and a new hybrid loss for Boundary-Aware Salient object detection.
-
23 Mar 2019 3 repositories listedIn this paper, we present a conditional generative adversarial network-based model for real-time underwater image enhancement.
-
26 Jul 2018 3 repositories listedBenefit from the quick development of deep learning techniques, salient object detection has achieved remarkable progresses recently.
-
26 May 2021 2 repositories listed Syntology ran 13 of 20 samples · 7 unverified · 16 pointer-only (licence)Since 2014 transfer learning has become the key driver for the improvement of spatial saliency prediction; however, with stagnant progress in the last 3-5 years.
-
20 Apr 2021 2 repositories listedFor the former, we apply transformer to a deterministic model, and explain that the effective structure modeling and global context modeling abilities lead to its superior performance compared with the CNN based…
-
11 Mar 2020 2 repositories listed Syntology ran 3 of 9 samples · 6 unverifiedWe evaluate our method on the video saliency datasets DHF1K, Hollywood-2 and UCF-Sports, and the image saliency datasets SALICON and MIT300.
-
3 Jul 2019 2 repositories listedThis paper investigates modifying an existing neural network architecture for static saliency prediction using two types of recurrences that integrate information from the temporal domain.
-
25 May 2019 2 repositories listedOur results suggest that (1) audio is a strong contributing cue for saliency prediction, (2) salient visible sound-source is the natural cause of the superiority of our Audio-Visual model, (3) richer feature…
-
15 Dec 2018 2 repositories listedWe propose three specific formulations of the PiCANet via embedding the pixel-wise contextual attention mechanism into the pooling and convolution operations with attending to global or local contexts.
-
15 Nov 2018 2 repositories listedLateral connections in the primary visual cortex (V1) have long been hypothesized to be responsible of several visual processing mechanisms such as brightness induction, chromatic induction, visual discomfort and…
-
28 Aug 2018 2 repositories listedThis work adapts a deep neural model for image saliency prediction to the temporal domain of egocentric video.
-
10 Apr 2018 2 repositories listedDespite the noticeable progress in perceptual tasks like detection, instance segmentation and human parsing, computers still perform unsatisfactorily on visually understanding humans in crowded scenes, such as group…
-
24 Mar 2018 2 repositories listedWe present a new computational model for gaze prediction in egocentric videos by exploring patterns in temporal shift of gaze fixations (attention transition) that are dependent on egocentric manipulation tasks.
-
17 Jan 2018 2 repositories listed Syntology ran 3 of 3 samples · 0 unverifiedPredicting human fixations from images has recently seen large improvements by leveraging deep representations which were pretrained for object recognition.
-
22 Dec 2017 2 repositories listedIn this paper, we propose a novel two-step understanding method, namely Salient Relevance (SR) map, which aims to shed light on how deep CNNs recognize images and learn features from areas, referred to as attention…
-
29 Nov 2016 2 repositories listed Syntology ran 2 of 11 samples · 9 unverifiedData-driven saliency has recently gained a lot of attention thanks to the use of Convolutional Neural Networks for predicting gaze fixations.
-
5 Sep 2016 2 repositories listed Syntology ran 0 of 5 samples · 5 unverifiedCurrent state of the art models for saliency prediction employ Fully Convolutional networks that perform a non-linear combination of features extracted from the last convolutional layer to predict saliency maps.
-
8 Jun 2025 1 repository listedIn this work, we propose a diagnostic test for class sensitivity: a method's ability to distinguish between competing class labels on the same input.
-
15 May 2025 1 repository listedBuilding upon this, we propose a dual-branch priors embedding architecture (DBPEA) that establishes differentiated feature fusion pathways, embedding these two priors at optimal network positions to achieve performance…
-
13 Apr 2025 1 repository listedTo address this, we introduce the uncertainty guidance learning approach to SOD, intended to enhance the model's perception of uncertain regions.
-
2 Apr 2025 1 repository listedMesh saliency enhances the adaptability of 3D vision by identifying and emphasizing regions that naturally attract visual attention.
-
21 Mar 2025 1 repository listedNews outlets' competition for attention in news interfaces has highlighted the need for demographically-aware saliency prediction models.
-
11 Dec 2024 1 repository listedTextured meshes significantly enhance the realism and detail of objects by mapping intricate texture details onto the geometric structure of 3D models.
-
5 Nov 2024 1 repository listedAs object detection techniques continue to evolve, understanding their relationships with complementary visual tasks becomes crucial for optimising model architectures and computational resources.
Syntology lines on 9 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections