| CAST: Cross-Attention in Space and Time for Video Action Recognition |
1 |
1 |
30 Nov 2023 |
11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (17 pointer-only for licence) |
| ZeroI2V: Zero-Cost Adaptation of Pre-trained Transformers from Image to Video |
2 |
1 |
2 Oct 2023 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified |
| Temporally-Adaptive Models for Efficient Video Understanding |
1 |
2 |
10 Aug 2023 |
official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified |
| What Can Simple Arithmetic Operations Do for Temporal Modeling? |
2 |
1 |
18 Jul 2023 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified |
| Hiera: A Hierarchical Vision Transformer without the Bells-and-Whistles |
4 |
1 |
1 Jun 2023 |
official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified |
| Implicit Temporal Modeling with Learnable Alignment for Video Recognition |
1 |
2 |
20 Apr 2023 |
official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (4 pointer-only for licence) |
| VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking |
1 |
1 |
29 Mar 2023 |
official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 2 samples that ran constructed an object rather than computing a result |
| MAGVIT: Masked Generative Video Transformer |
1 |
2 |
10 Dec 2022 |
official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| InternVideo: General Video Foundation Models via Generative and Discriminative Learning |
2 |
1 |
6 Dec 2022 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified |
| Rethinking Video ViTs: Sparse Video Tubes for Joint Image and Video Learning |
1 |
1 |
6 Dec 2022 |
3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| AdaMAE: Adaptive Masking for Efficient Spatiotemporal Learning with Masked Autoencoders |
2 |
1 |
16 Nov 2022 |
official (archive's flag): 14 ran · 18 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (6 pointer-only for licence) |
| Spatiotemporal Self-attention Modeling with Temporal Patch Shift for Action Recognition |
1 |
1 |
27 Jul 2022 |
official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result |
| ST-Adapter: Parameter-Efficient Image-to-Video Transfer Learning |
1 |
1 |
27 Jun 2022 |
official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence) |
| VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training |
9 |
3 |
23 Mar 2022 |
official (archive's flag): 9 ran · 10 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (12 pointer-only for licence) |
| DirecFormer: A Directed Attention in Transformer Approach to Robust Action Recognition |
1 |
1 |
19 Mar 2022 |
official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (8 pointer-only for licence) |
| Group Contextualization for Video Recognition |
1 |
1 |
18 Mar 2022 |
official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result |
| Omnivore: A Single Model for Many Visual Modalities |
2 |
1 |
20 Jan 2022 |
official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (1 pointer-only for licence) |
| MorphMLP: An Efficient MLP-Like Backbone for Spatial-Temporal Representation Learning |
2 |
1 |
24 Nov 2021 |
official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified |
| Relational Self-Attention: What's Missing in Attention for Video Understanding |
1 |
5 |
2 Nov 2021 |
official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (1 pointer-only for licence) |
| Object-Region Video Transformers |
1 |
2 |
13 Oct 2021 |
6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (7 pointer-only for licence) |
| Video Swin Transformer |
15 |
1 |
24 Jun 2021 |
official (archive's flag): 3 ran · 19 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 5 where Syntology's instrument failed) · 13 unverified (7 pointer-only for licence) |
| VIMPAC: Video Pre-Training via Masked Token Prediction and Contrastive Learning |
1 |
1 |
21 Jun 2021 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence) |
| Space-time Mixing Attention for Video Transformer |
1 |
1 |
10 Jun 2021 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified |
| Keeping Your Eye on the Ball: Trajectory Attention in Video Transformers |
2 |
3 |
9 Jun 2021 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (2 pointer-only for licence) |
| CT-Net: Channel Tensorization Network for Video Classification |
1 |
1 |
3 Jun 2021 |
official (archive's flag): 11 ran · 11 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 7 where Syntology's instrument failed) · 2 unverified |
| Multiscale Vision Transformers |
8 |
3 |
22 Apr 2021 |
community repositories only · 13 ran (of which 10 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 13 unverified (5 pointer-only for licence) |
| ViViT: A Video Vision Transformer |
10 |
1 |
29 Mar 2021 |
official: no sample here; runs from other or unrecorded repositories · 14 ran (of which 10 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (1 pointer-only for licence) |
| MoViNets: Mobile Video Networks for Efficient Video Recognition |
3 |
4 |
21 Mar 2021 |
community repositories only · 9 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified |
| Learning Self-Similarity in Space and Time as Generalized Motion for Video Action Recognition |
1 |
3 |
14 Feb 2021 |
official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (4 pointer-only for licence) |
| Is Space-Time Attention All You Need for Video Understanding? |
16 |
3 |
9 Feb 2021 |
official: no sample here; runs from other or unrecorded repositories · 35 ran (of which 17 constructed an object rather than computing a result; 21 with no instrument failure: 0 honoured, 1 violated, 20 with no contract checked; 14 where Syntology's instrument failed) · 8 unverified (14 pointer-only for licence) |
| TDN: Temporal Difference Networks for Efficient Action Recognition |
1 |
2 |
18 Dec 2020 |
official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified |
| MotionSqueeze: Neural Motion Feature Learning for Video Understanding |
2 |
4 |
20 Jul 2020 |
6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (2 pointer-only for licence) |
| Temporal Pyramid Network for Action Recognition |
3 |
1 |
7 Apr 2020 |
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified |
| More Is Less: Learning Efficient Video Representations by Big-Little Network and Depthwise Temporal Aggregation |
1 |
1 |
2 Dec 2019 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified |
| A Multigrid Method for Efficiently Training Video Models |
3 |
1 |
2 Dec 2019 |
community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (3 pointer-only for licence) |
| SlowFast Networks for Video Recognition |
15 |
1 |
10 Dec 2018 |
community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence) |
| TSM: Temporal Shift Module for Efficient Video Understanding |
13 |
1 |
20 Nov 2018 |
official: harvested, nothing ran · 10 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 6 unverified (6 pointer-only for licence) |
| Temporal Relational Reasoning in Videos |
5 |
1 |
22 Nov 2017 |
2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence) |