Papers › Hierarchical Domain-Adapted Feature Learning for Video Saliency Prediction

Hierarchical Domain-Adapted Feature Learning for Video Saliency Prediction

2 Oct 2020arXiv:2010.01220archive 2025-07-28

Giovanni Bellitto, Federica Proietto Salanitri, Simone Palazzo, Francesco Rundo, Daniela Giordano, Concetto Spampinato

In this work, we propose a 3D fully convolutional architecture for video saliency prediction that employs hierarchical supervision on intermediate maps (referred to as conspicuity maps) generated using features extracted at different abstraction levels. We provide the base hierarchical learning mechanism with two techniques for domain adaptation and domain-specific learning. For the former, we encourage the model to unsupervisedly learn hierarchical general features using gradient reversal at multiple scales, to enhance generalization capabilities on datasets for which no annotations are provided during training. As for domain specialization, we employ domain-specific operations (namely, priors, smoothing and batch normalization) by specializing the learned features on individual datasets in order to maximize performance. The results of our experiments show that the proposed model yields state-of-the-art accuracy on supervised saliency prediction. When the base hierarchical model is empowered with domain-specific modules, performance improves, outperforming state-of-the-art models on three out of five metrics on the DHF1K benchmark and reaching the second-best results on the other two. When, instead, we test it in an unsupervised domain adaptation setting, by enabling hierarchical gradient reversal layers, we obtain performance comparable to supervised state-of-the-art.

PaperPDFCode

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

perceivelab/hd2s officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Domain AdaptationSaliency DetectionSaliency PredictionUnsupervised Domain AdaptationVideo Saliency DetectionVideo Saliency Prediction

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Video Saliency Detection DHF1K HD2S AUC-J 0.908 #2 of 3 Archive leaderboard report
Video Saliency Detection DHF1K HD2S CC 0.503 #2 of 3 Archive leaderboard report
Video Saliency Detection DHF1K HD2S NSS 2.812 #2 of 3 Archive leaderboard report
Video Saliency Detection DHF1K HD2S SIM 0.406 #2 of 3 Archive leaderboard report
Video Saliency Detection DHF1K HD2S s-AUC 0.70 #2 of 3 Archive leaderboard report
Video Saliency Detection MSU Video Saliency Prediction HD2S AUC-J 0.844 #2 of 14 Archive leaderboard report
Video Saliency Detection MSU Video Saliency Prediction HD2S CC 0.707 #2 of 14 Archive leaderboard report
Video Saliency Detection MSU Video Saliency Prediction HD2S FPS 24.51 #2 of 14 Archive leaderboard report
Video Saliency Detection MSU Video Saliency Prediction HD2S KLDiv 0.545 #2 of 14 Archive leaderboard report
Video Saliency Detection MSU Video Saliency Prediction HD2S NSS 1.89 #2 of 14 Archive leaderboard report
Video Saliency Detection MSU Video Saliency Prediction HD2S SIM 0.615 #2 of 14 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections