{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/the-youtube-8m-kaggle-competition-challenges","title":"The YouTube-8M Kaggle Competition: Challenges and Methods","arxiv_id":"1706.09274","date":"2017-06-28","proceeding":null,"authors":["Haosheng Zou","Kun Xu","Jialian Li","Jun Zhu"],"abstract":"We took part in the YouTube-8M Video Understanding Challenge hosted on\nKaggle, and achieved the 10th place within less than one month's time. In this\npaper, we present an extensive analysis and solution to the underlying\nmachine-learning problem based on frame-level data, where major challenges are\nidentified and corresponding preliminary methods are proposed. It's noteworthy\nthat, with merely the proposed strategies and uniformly-averaging multi-crop\nensemble was it sufficient for us to reach our ranking. We also report the\nmethods we believe to be promising but didn't have enough time to train to\nconvergence. We hope this paper could serve, to some extent, as a review and\nguideline of the YouTube-8M multi-label video classification benchmark,\ninspiring future attempts and research.","url_abs":"http://arxiv.org/abs/1706.09274v2","url_pdf":"http://arxiv.org/pdf/1706.09274v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"the-youtube-8m-kaggle-competition-challenges","repo_url":"https://github.com/taufikxu/youtube","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"tf","reach":{"status":"ok","spdx":"Apache-2.0"}}],"tasks":[{"task_slug":"classification","task_name":"General Classification"},{"task_slug":"video-classification","task_name":"Video Classification"},{"task_slug":"video-understanding","task_name":"Video Understanding"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":null,"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}