Papers › DSNet: A Flexible Detect-to-Summarize Network for Video Summarization
DSNet: A Flexible Detect-to-Summarize Network for Video Summarization
Wencheng Zhu, Jiwen Lu, Jiahao Li, and Jie Zhou
In this paper, we propose a Detect-to-Summarize network (DSNet) framework for supervised video summarization. Our DSNet contains anchor-based and anchor-free counterparts. The anchor-based method generates temporal interest proposals to determine and localize the representative contents of video sequences, while the anchor-free method eliminates the pre-defined temporal proposals and directly predicts the importance scores and segment locations. Different from existing supervised video summarization methods which formulate video summarization as a regression problem without temporal consistency and integrity constraints, our interest detection framework is the first attempt to leverage temporal consistency via the temporal interest detection formulation. Specifically, in the anchor-based approach, we first provide a dense sampling of temporal interest proposals with multi-scale intervals that accommodate interest variations in length, and then extract their long-range temporal features for interest proposal location regression and importance prediction. Notably, positive and negative segments are both assigned for the correctness and completeness information of the generated summaries. In the anchor-free approach, we alleviate drawbacks of temporal proposals by directly predicting importance scores of video frames and segment locations. Particularly, the interest detection framework can be flexibly plugged into off-the-shelf supervised video summarization methods. We evaluate the anchor-based and anchor-free approaches on the SumMe and TVSum datasets. Experimental results clearly validate the effectiveness of the anchor-based and anchor-free approaches.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Supervised Video Summarization | SumMe | DSNet | F1-score (Augmented) | 50.7 | #12 of 21 | Archive leaderboard | report |
| Supervised Video Summarization | SumMe | DSNet | F1-score (Canonical) | 50.2 | #12 of 21 | Archive leaderboard | report |
| Supervised Video Summarization | TvSum | DSNet | F1-score (Augmented) | 63.9 | #9 of 21 | Archive leaderboard | report |
| Supervised Video Summarization | TvSum | DSNet | F1-score (Canonical) | 62.1 | #9 of 21 | Archive leaderboard | report |
| Video Summarization | SumMe | DSNet | F1-score (Augmented) | 53.3 | #3 of 6 | Archive leaderboard | report |
| Video Summarization | SumMe | DSNet | F1-score (Canonical) | 53.0 | #3 of 6 | Archive leaderboard | report |
| Video Summarization | TvSum | DSNet | F1-score (Augmented) | 63.9 | #2 of 6 | Archive leaderboard | report |
| Video Summarization | TvSum | DSNet | F1-score (Canonical) | 62.1 | #2 of 6 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections