Datasets › Sports-1M
Sports-1M
The Sports-1M dataset consists of over a million videos from YouTube. The videos in the dataset can be obtained through the YouTube URL specified by the authors. Approximately 7% (as of 2016) of the videos have been removed by the YouTube uploaders since the dataset was compiled. However, there are still over a million videos in the dataset with 487 sports-related categories with 1,000 to 3,000 videos per category. The videos are automatically labelled with 487 sports classes using the YouTube Topics API by analyzing the text metadata associated with the videos (e.g. tags, descriptions). Approximately 5% of the videos are annotated with more than one class.
Source: Review of Action Recognition and Detection Methods
Image Source: Computer Vision for Sports
Benchmarks archive 2025-07-28
All 2 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Action Recognition | Sports-1M | ip-CSN-152 (RGB) Video hit@1 75.5 | Video Classification with Channel-Separated... | open-mmlab/mmaction2 +6 | 9 | Compare |
| Action Recognition In Videos | Sports-1M | G-Blend Video hit@1 74.8 | What Makes Training Multi-Modal Classification Networks Hard? | facebookresearch/R2Plus1D +2 | 2 | Compare |
Papers archive 2025-07-28
8 shown of 8 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 164. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.
| Date | Samples run Syntology | |||
|---|---|---|---|---|
| What Makes Training Multi-Modal Classification Networks Hard? | 3 | 1 | 29 May 2019 | ran 1 of 11 samples (10 unverified) |
| Video Classification with Channel-Separated Convolutional Networks | 7 | 2 | 4 Apr 2019 | ran 1 of 4 samples (3 unverified) |
| A Closer Look at Spatiotemporal Convolutions for Action Recognition | 24 | 3 | 30 Nov 2017 | ran 1 of 4 samples (3 unverified; 4 pointer-only for licence) |
| Learning Spatio-Temporal Representation with Pseudo-3D Residual Networks | 2 | 1 | 28 Nov 2017 | not harvested |
| YouTube-8M: A Large-Scale Video Classification Benchmark | 7 | 1 | 27 Sep 2016 | ran 1 of 9 samples (8 unverified) |
| Beyond Short Snippets: Deep Networks for Video Classification | 1 | 1 | 31 Mar 2015 | ran 0 of 4 samples (4 unverified) |
| Learning Spatiotemporal Features with 3D Convolutional Networks | 29 | 1 | 2 Dec 2014 | not harvested |
| Large-Scale Video Classification with Convolutional Neural Networks | 1 | 1 | 23 Jun 2014 | not harvested |
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
License archive 2025-07-28
Modalities archive 2025-07-28
Languages archive 2025-07-28
No language tagged.
Variants archive 2025-07-28
- Sports-1M
1 variant name, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections