Datasets › Sports-1M

Sports-1M

Introduced by Andrej Karpathy et al. in Large-Scale Video Classification with Convolutional Neural Networks1 Jan 2014 archive 2025-07-28

The Sports-1M dataset consists of over a million videos from YouTube. The videos in the dataset can be obtained through the YouTube URL specified by the authors. Approximately 7% (as of 2016) of the videos have been removed by the YouTube uploaders since the dataset was compiled. However, there are still over a million videos in the dataset with 487 sports-related categories with 1,000 to 3,000 videos per category. The videos are automatically labelled with 487 sports classes using the YouTube Topics API by analyzing the text metadata associated with the videos (e.g. tags, descriptions). Approximately 5% of the videos are annotated with more than one class.

Source: Review of Action Recognition and Detection Methods

Image Source: Computer Vision for Sports

Benchmarks archive 2025-07-28

All 2 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Action Recognition Sports-1M ip-CSN-152 (RGB) Video hit@1 75.5 Video Classification with Channel-Separated... open-mmlab/mmaction2 +6 9 Compare
Action Recognition In Videos Sports-1M G-Blend Video hit@1 74.8 What Makes Training Multi-Modal Classification Networks Hard? facebookresearch/R2Plus1D +2 2 Compare

Papers archive 2025-07-28

8 shown of 8 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 164. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
What Makes Training Multi-Modal Classification Networks Hard? 3 1 29 May 2019 ran 1 of 11 samples (10 unverified)
Video Classification with Channel-Separated Convolutional Networks 7 2 4 Apr 2019 ran 1 of 4 samples (3 unverified)
A Closer Look at Spatiotemporal Convolutions for Action Recognition 24 3 30 Nov 2017 ran 1 of 4 samples (3 unverified; 4 pointer-only for licence)
Learning Spatio-Temporal Representation with Pseudo-3D Residual Networks 2 1 28 Nov 2017 not harvested
YouTube-8M: A Large-Scale Video Classification Benchmark 7 1 27 Sep 2016 ran 1 of 9 samples (8 unverified)
Beyond Short Snippets: Deep Networks for Video Classification 1 1 31 Mar 2015 ran 0 of 4 samples (4 unverified)
Learning Spatiotemporal Features with 3D Convolutional Networks 29 1 2 Dec 2014 not harvested
Large-Scale Video Classification with Convolutional Neural Networks 1 1 23 Jun 2014 not harvested

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

CC BY 3.0

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • Sports-1M

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections