Datasets › Something-Something V2 › Papers where code ran, page 1

Something-Something V2

Papers archive 2025-07-28

papers with a benchmark row: 85 · with a code link: 66 · where Syntology ran a sample: 38 (34 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument) Syntology

Show: all papers with a benchmark rowonly where code ran (38 of 85 with a benchmark row: 34 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument)

Syntology We ran code from the paper's repository; we did not run it on this dataset or check it against this dataset's benchmarks.

Page 1 of 1: papers 1 to 38 of the 38 papers with a benchmark row here where Syntology ran at least one harvested sample (34 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.

The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset, not that list; the archive's count for this dataset is 290. The Syntology column is from Syntology's graph, stated per sample; it is not part of any archive number. A line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified” (C is a part of N, never taken away from it); the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. When the archive marks a repository official for the paper, the cell starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record.

PaperCodeResultsDateSamples run Syntology
CAST: Cross-Attention in Space and Time for Video Action Recognition 1 1 30 Nov 2023 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (17 pointer-only for licence)
ZeroI2V: Zero-Cost Adaptation of Pre-trained Transformers from Image to Video 2 1 2 Oct 2023 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified
Temporally-Adaptive Models for Efficient Video Understanding 1 2 10 Aug 2023 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified
What Can Simple Arithmetic Operations Do for Temporal Modeling? 2 1 18 Jul 2023 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified
Hiera: A Hierarchical Vision Transformer without the Bells-and-Whistles 4 1 1 Jun 2023 official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified
Implicit Temporal Modeling with Learnable Alignment for Video Recognition 1 2 20 Apr 2023 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (4 pointer-only for licence)
VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking 1 1 29 Mar 2023 official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 2 samples that ran constructed an object rather than computing a result
MAGVIT: Masked Generative Video Transformer 1 2 10 Dec 2022 official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
InternVideo: General Video Foundation Models via Generative and Discriminative Learning 2 1 6 Dec 2022 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified
Rethinking Video ViTs: Sparse Video Tubes for Joint Image and Video Learning 1 1 6 Dec 2022 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
AdaMAE: Adaptive Masking for Efficient Spatiotemporal Learning with Masked Autoencoders 2 1 16 Nov 2022 official (archive's flag): 14 ran · 18 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (6 pointer-only for licence)
Spatiotemporal Self-attention Modeling with Temporal Patch Shift for Action Recognition 1 1 27 Jul 2022 official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result
ST-Adapter: Parameter-Efficient Image-to-Video Transfer Learning 1 1 27 Jun 2022 official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence)
VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training 9 3 23 Mar 2022 official (archive's flag): 9 ran · 10 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (12 pointer-only for licence)
DirecFormer: A Directed Attention in Transformer Approach to Robust Action Recognition 1 1 19 Mar 2022 official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (8 pointer-only for licence)
Group Contextualization for Video Recognition 1 1 18 Mar 2022 official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result
Omnivore: A Single Model for Many Visual Modalities 2 1 20 Jan 2022 official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (1 pointer-only for licence)
MorphMLP: An Efficient MLP-Like Backbone for Spatial-Temporal Representation Learning 2 1 24 Nov 2021 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified
Relational Self-Attention: What's Missing in Attention for Video Understanding 1 5 2 Nov 2021 official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (1 pointer-only for licence)
Object-Region Video Transformers 1 2 13 Oct 2021 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (7 pointer-only for licence)
Video Swin Transformer 15 1 24 Jun 2021 official (archive's flag): 3 ran · 19 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 5 where Syntology's instrument failed) · 13 unverified (7 pointer-only for licence)
VIMPAC: Video Pre-Training via Masked Token Prediction and Contrastive Learning 1 1 21 Jun 2021 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Space-time Mixing Attention for Video Transformer 1 1 10 Jun 2021 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
Keeping Your Eye on the Ball: Trajectory Attention in Video Transformers 2 3 9 Jun 2021 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (2 pointer-only for licence)
CT-Net: Channel Tensorization Network for Video Classification 1 1 3 Jun 2021 official (archive's flag): 11 ran · 11 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 7 where Syntology's instrument failed) · 2 unverified
Multiscale Vision Transformers 8 3 22 Apr 2021 community repositories only · 13 ran (of which 10 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 13 unverified (5 pointer-only for licence)
ViViT: A Video Vision Transformer 10 1 29 Mar 2021 official: no sample here; runs from other or unrecorded repositories · 14 ran (of which 10 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (1 pointer-only for licence)
MoViNets: Mobile Video Networks for Efficient Video Recognition 3 4 21 Mar 2021 community repositories only · 9 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified
Learning Self-Similarity in Space and Time as Generalized Motion for Video Action Recognition 1 3 14 Feb 2021 official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (4 pointer-only for licence)
Is Space-Time Attention All You Need for Video Understanding? 16 3 9 Feb 2021 official: no sample here; runs from other or unrecorded repositories · 35 ran (of which 17 constructed an object rather than computing a result; 21 with no instrument failure: 0 honoured, 1 violated, 20 with no contract checked; 14 where Syntology's instrument failed) · 8 unverified (14 pointer-only for licence)
TDN: Temporal Difference Networks for Efficient Action Recognition 1 2 18 Dec 2020 official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified
MotionSqueeze: Neural Motion Feature Learning for Video Understanding 2 4 20 Jul 2020 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (2 pointer-only for licence)
Temporal Pyramid Network for Action Recognition 3 1 7 Apr 2020 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
More Is Less: Learning Efficient Video Representations by Big-Little Network and Depthwise Temporal Aggregation 1 1 2 Dec 2019 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified
A Multigrid Method for Efficiently Training Video Models 3 1 2 Dec 2019 community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (3 pointer-only for licence)
SlowFast Networks for Video Recognition 15 1 10 Dec 2018 community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence)
TSM: Temporal Shift Module for Efficient Video Understanding 13 1 20 Nov 2018 official: harvested, nothing ran · 10 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 6 unverified (6 pointer-only for licence)
Temporal Relational Reasoning in Videos 5 1 22 Nov 2017 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence)