Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers where code ran, page 4
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 4 of 4: papers 301 to 328 of the 328 tagged papers where Syntology ran at least one harvested sample (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Container: Context Aggregation Network 2 Jun 2021 · 4 repositories · arXiv:2106.01401Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
You Only Look at One Sequence: Rethinking Transformer in Vision through Object Detection 1 Jun 2021 · 2 repositories · arXiv:2106.00666Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 5 pointer-only (licence)
-
TransMatcher: Deep Image Matching Through Transformers for Generalizable Person Re-identification 30 May 2021 · 2 repositories · arXiv:2105.14432Syntology official (archive's flag): 5 ran · 13 ran (of which 4 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 8 unverified (of 21 harvested samples) · 1 pointer-only (licence)
-
KVT: k-NN Attention for Boosting Vision Transformers 28 May 2021 · 1 repository · arXiv:2106.00515Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Nested Hierarchical Transformer: Towards Accurate, Data-Efficient and Interpretable Visual Understanding 26 May 2021 · 6 repositories · arXiv:2105.12723Syntology official (archive's flag): 9 ran · 19 ran (of which 0 constructed an object rather than computing a result; 18 with no instrument failure: 1 honoured, 1 violated, 16 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 26 harvested samples) · 6 pointer-only (licence)
-
Vision Transformers are Robust Learners 17 May 2021 · 1 repository · arXiv:2105.07581Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
A Large-Scale Benchmark for Food Image Segmentation 12 May 2021 · 2 repositories · arXiv:2105.05409Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Segmenter: Transformer for Semantic Segmentation 12 May 2021 · 8 repositories · arXiv:2105.05633Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Instances as Queries 5 May 2021 · 5 repositories · arXiv:2105.01928Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Emerging Properties in Self-Supervised Vision Transformers 29 Apr 2021 · 32 repositories · arXiv:2104.14294Syntology official: no sample here; runs from other or unrecorded repositories · 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 20 harvested samples) · 4 pointer-only (licence)
-
Twins: Revisiting the Design of Spatial Attention in Vision Transformers 28 Apr 2021 · 9 repositories · arXiv:2104.13840Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Visual Saliency Transformer 25 Apr 2021 · 2 repositories · arXiv:2104.12099Syntology 7 ran (of which 4 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
All Tokens Matter: Token Labeling for Training Better Vision Transformers 22 Apr 2021 · 7 repositories · arXiv:2104.10858Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
VATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and Text 22 Apr 2021 · 5 repositories · arXiv:2104.11178Syntology community repositories only · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Vision Transformer Pruning 17 Apr 2021 · 2 repositories · arXiv:2104.08500Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
ViT-V-Net: Vision Transformer for Unsupervised Volumetric Medical Image Registration 13 Apr 2021 · 1 repository · arXiv:2104.06468Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Putting NeRF on a Diet: Semantically Consistent Few-Shot View Synthesis 1 Apr 2021 · 2 repositories · arXiv:2104.00677Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 2 pointer-only (licence)
-
Rethinking Spatial Dimensions of Vision Transformers 30 Mar 2021 · 12 repositories · arXiv:2103.16302Syntology official: harvested, nothing ran · 10 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified (of 20 harvested samples)
-
Multi-Scale Vision Longformer: A New Vision Transformer for High-Resolution Image Encoding 29 Mar 2021 · 3 repositories · arXiv:2103.15358Syntology official (archive's flag): 3 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
CvT: Introducing Convolutions to Vision Transformers 29 Mar 2021 · 16 repositories · arXiv:2103.15808Syntology official (archive's flag): 10 ran · 39 ran (of which 19 constructed an object rather than computing a result; 36 with no instrument failure: 2 honoured, 0 violated, 34 with no contract checked; 3 where Syntology's instrument failed) · 8 unverified (of 47 harvested samples) · 8 pointer-only (licence)
-
CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification 27 Mar 2021 · 15 repositories · arXiv:2103.14899Syntology official (archive's flag): 4 ran · 17 ran (of which 10 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 1 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 26 harvested samples)
-
AutoMix: Unveiling the Power of Mixup for Stronger Classifiers 24 Mar 2021 · 3 repositories · arXiv:2103.13027Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
Vision Transformers for Dense Prediction 24 Mar 2021 · 15 repositories · arXiv:2103.13413Syntology 65 ran (of which 19 constructed an object rather than computing a result; 32 with no instrument failure: 0 honoured, 0 violated, 32 with no contract checked; 33 where Syntology's instrument failed) · 51 unverified (of 116 harvested samples) · 15 pointer-only (licence)
-
TransFG: A Transformer Architecture for Fine-grained Recognition 14 Mar 2021 · 2 repositories · arXiv:2103.07976Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
PipeTransformer: Automated Elastic Pipelining for Distributed Training of Transformers 5 Feb 2021 · 1 repository · arXiv:2102.03161Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNet 28 Jan 2021 · 13 repositories · arXiv:2101.11986Syntology official (archive's flag): 4 ran · 21 ran (of which 16 constructed an object rather than computing a result; 21 with no instrument failure: 1 honoured, 0 violated, 20 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 26 harvested samples) · 8 pointer-only (licence)
-
DAF:re: A Challenging, Crowd-Sourced, Large-Scale, Long-Tailed Dataset For Anime Character Recognition 21 Jan 2021 · 2 repositories · arXiv:2101.08674Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale 22 Oct 2020 · 158 repositories · arXiv:2010.11929Syntology official: harvested, nothing ran · 307 ran (of which 165 constructed an object rather than computing a result; 286 with no instrument failure: 8 honoured, 2 violated, 276 with no contract checked; 21 where Syntology's instrument failed) · 112 unverified (of 419 harvested samples) · 154 pointer-only (licence)