Papers › The complementarity of a diverse range of deep learning features extracted from video...
The complementarity of a diverse range of deep learning features extracted from video content for video recommendation
Adolfo Almeida, Johan Pieter de Villiers, Allan De Freitas, Mergandran Velayudan
Following the popularisation of media streaming, a number of video streaming services are continuously buying new video content to mine the potential profit from them. As such, the newly added content has to be handled well to be recommended to suitable users. In this paper, we address the new item cold-start problem by exploring the potential of various deep learning features to provide video recommendations. The deep learning features investigated include features that capture the visual-appearance, audio and motion information from video content. We also explore different fusion methods to evaluate how well these feature modalities can be combined to fully exploit the complementary information captured by them. Experiments on a real-world video dataset for movie recommendations show that deep learning features outperform hand-crafted features. In particular, recommendations generated with deep learning audio features and action-centric deep learning features are superior to MFCC and state-of-the-art iDT features. In addition, the combination of various deep learning features with hand-crafted features and textual metadata yields significant improvement in recommendations compared to combining only the former.
Code
Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.
Code Syntology ran Syntology
Not run by Syntology. Nothing on this page verifies that the listed code works.
Tasks
Results from the paper archive 2025-07-28
| Task | Dataset | Model | Metric | Value | Rank at snapshot | Leaderboard | Report |
|---|---|---|---|---|---|---|---|
| Recommendation Systems | MovieLens 10M | scaled-CER | MAP@15 | 0.1568 | #16 of 17 | Archive leaderboard | report |
| Recommendation Systems | MovieLens 10M | scaled-CER | MAP@30 | 0.1671 | #16 of 17 | Archive leaderboard | report |
| Recommendation Systems | MovieLens 10M | scaled-CER | MAP@5 | 0.1536 | #16 of 17 | Archive leaderboard | report |
| Recommendation Systems | MovieLens 10M | scaled-CER | NDCG@15 | 0.2546 | #16 of 17 | Archive leaderboard | report |
| Recommendation Systems | MovieLens 10M | scaled-CER | NDCG@30 | 0.2971 | #16 of 17 | Archive leaderboard | report |
| Recommendation Systems | MovieLens 10M | scaled-CER | NDCG@5 | 0.1846 | #16 of 17 | Archive leaderboard | report |
Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections