Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers, page 20
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 20 of 22: papers 1,901 to 2,000 of 2,144, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
O-ViT: Orthogonal Vision Transformer 28 Jan 2022 · 0 repositories · arXiv:2201.12133
-
ViT-HGR: Vision Transformer-based Hand Gesture Recognition from High Density Surface EMG Signals 25 Jan 2022 · 1 repository · arXiv:2201.10060
-
Improving Chest X-Ray Report Generation by Leveraging Warm Starting 24 Jan 2022 · 1 repository · arXiv:2201.09405
-
Patches Are All You Need? 24 Jan 2022 · 12 repositories · arXiv:2201.09792Syntology official (archive's flag): 6 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 5 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Fast Differentiable Matrix Square Root 21 Jan 2022 · 1 repository · arXiv:2201.08663Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
MeMViT: Memory-Augmented Multiscale Vision Transformer for Efficient Long-Term Video Recognition 20 Jan 2022 · 1 repository · arXiv:2201.08383
-
TerViT: An Efficient Ternary Vision Transformer 20 Jan 2022 · 0 repositories · arXiv:2201.08050
-
Q-ViT: Fully Differentiable Quantization for Vision Transformer 19 Jan 2022 · 1 repository · arXiv:2201.07703Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
RePre: Improving Self-Supervised Vision Transformer with Reconstructive Pre-training 18 Jan 2022 · 0 repositories · arXiv:2201.06857
-
SwinUNet3D -- A Hierarchical Architecture for Deep Traffic Prediction using Shifted Window Transformers 17 Jan 2022 · 1 repository · arXiv:2201.06390Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
ViTBIS: Vision Transformer for Biomedical Image Segmentation 15 Jan 2022 · 0 repositories · arXiv:2201.05920
-
Lawin Transformer: Improving Semantic Segmentation Transformer with Multi-Scale Representations via Large Window Attention 5 Jan 2022 · 3 repositories · arXiv:2201.01615
-
Short Range Correlation Transformer for Occluded Person Re-Identification 4 Jan 2022 · 0 repositories · arXiv:2201.01090
-
CaFT: Clustering and Filter on Tokens of Transformer for Weakly Supervised Object Localization 3 Jan 2022 · 0 repositories · arXiv:2201.00475
-
D-Former: A U-shaped Dilated Transformer for 3D Medical Image Segmentation 3 Jan 2022 · 1 repository · arXiv:2201.00462
-
Vision Transformer Slimming: Multi-Dimension Searching in Continuous Optimization Space 3 Jan 2022 · 1 repository · arXiv:2201.00814Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Splicing ViT Features for Semantic Appearance Transfer 2 Jan 2022 · 1 repository · arXiv:2201.00424Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CADTransformer: Panoptic Symbol Spotting Transformer for CAD Drawings 1 Jan 2022 · 1 repository
-
Chitransformer: Towards Reliable Stereo From Cues 1 Jan 2022 · 1 repository
-
Continual Learning With Lifelong Vision Transformer 1 Jan 2022 · 0 repositories
-
Learning Transferable Human-Object Interaction Detector With Natural Language Supervision 1 Jan 2022 · 1 repository
-
Neural Window Fully-Connected CRFs for Monocular Depth Estimation 1 Jan 2022 · 0 repositories
-
Recurring the Transformer for Video Action Recognition 1 Jan 2022 · 0 repositories
-
Training Object Detectors From Scratch: An Empirical Study in the Era of Vision Transformer 1 Jan 2022 · 0 repositories
-
Pale Transformer: A General Vision Transformer Backbone with Pale-Shaped Attention 28 Dec 2021 · 2 repositories · arXiv:2112.14000Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Learning Generative Vision Transformer with Energy-Based Latent Space for Saliency Prediction 27 Dec 2021 · 0 repositories · arXiv:2112.13528
-
Learning Robust and Lightweight Model through Separable Structured Transformations 27 Dec 2021 · 0 repositories · arXiv:2112.13551
-
SPViT: Enabling Faster Vision Transformers via Soft Token Pruning 27 Dec 2021 · 1 repository · arXiv:2112.13890Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
ViR:the Vision Reservoir 27 Dec 2021 · 0 repositories · arXiv:2112.13545
-
Vision Transformer for Small-Size Datasets 27 Dec 2021 · 5 repositories · arXiv:2112.13492Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
SimViT: Exploring a Simple Vision Transformer with sliding windows 24 Dec 2021 · 2 repositories · arXiv:2112.13085
-
Learned Queries for Efficient Local Attention 21 Dec 2021 · 1 repository · arXiv:2112.11435Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
MIA-Former: Efficient and Robust Vision Transformers via Multi-grained Input-Adaptation 21 Dec 2021 · 0 repositories · arXiv:2112.11542
-
MPViT: Multi-Path Vision Transformer for Dense Prediction 21 Dec 2021 · 3 repositories · arXiv:2112.11010
-
Lite Vision Transformer with Enhanced Self-Attention 20 Dec 2021 · 1 repository · arXiv:2112.10809
-
A Simple Single-Scale Vision Transformer for Object Localization and Instance Segmentation 17 Dec 2021 · 3 repositories · arXiv:2112.09747
-
Towards End-to-End Image Compression and Analysis with Transformers 17 Dec 2021 · 1 repository · arXiv:2112.09300
-
How to augment your ViTs? Consistency loss and StyleAug, a random style transfer augmentation 16 Dec 2021 · 0 repositories · arXiv:2112.09260
-
SeqFormer: Sequential Transformer for Video Instance Segmentation 15 Dec 2021 · 2 repositories · arXiv:2112.08275
-
Vision Transformer Based Video Hashing Retrieval for Tracing the Source of Fake Videos 15 Dec 2021 · 0 repositories · arXiv:2112.08117
-
AdaViT: Adaptive Tokens for Efficient Vision Transformer 14 Dec 2021 · 1 repository · arXiv:2112.07658Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Improving Vision Transformers for Incremental Learning 12 Dec 2021 · 0 repositories · arXiv:2112.06103
-
Deep ViT Features as Dense Visual Descriptors 10 Dec 2021 · 1 repository · arXiv:2112.05814Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Injecting Semantic Concepts into End-to-End Image Captioning 9 Dec 2021 · 1 repository · arXiv:2112.05230Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Transformaly -- Two (Feature Spaces) Are Better Than One 8 Dec 2021 · 1 repository · arXiv:2112.04185
-
PolyphonicFormer: Unified Query Learning for Depth-aware Video Panoptic Segmentation 5 Dec 2021 · 1 repository · arXiv:2112.02582
-
Pose-guided Feature Disentangling for Occluded Person Re-identification Based on Transformer 5 Dec 2021 · 1 repository · arXiv:2112.02466
-
LAVT: Language-Aware Vision Transformer for Referring Image Segmentation 4 Dec 2021 · 1 repository · arXiv:2112.02244
-
Make A Long Image Short: Adaptive Token Length for Vision Transformers 3 Dec 2021 · 0 repositories · arXiv:2112.01686
-
TBN-ViT: Temporal Bilateral Network with Vision Transformer for Video Scene Parsing 2 Dec 2021 · 0 repositories · arXiv:2112.01033
-
Container: Context Aggregation Networks 1 Dec 2021 · 2 repositories
-
Federated Split Task-Agnostic Vision Transformer for COVID-19 CXR Diagnosis 1 Dec 2021 · 0 repositories
-
Focal Attention for Long-Range Interactions in Vision Transformers 1 Dec 2021 · 1 repository
-
HRFormer: High-Resolution Vision Transformer for Dense Predict 1 Dec 2021 · 2 repositories
-
TEDGE-Caching: Transformer-based Edge Caching Towards 6G Networks 1 Dec 2021 · 0 repositories · arXiv:2112.00633
-
A Unified Pruning Framework for Vision Transformers 30 Nov 2021 · 1 repository · arXiv:2111.15127
-
Adaptive Token Sampling For Efficient Vision Transformers 30 Nov 2021 · 1 repository · arXiv:2111.15667Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Boosting Discriminative Visual Representation Learning with Scenario-Agnostic Mixup 30 Nov 2021 · 1 repository · arXiv:2111.15454
-
Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models 30 Nov 2021 · 1 repository · arXiv:2112.00029Syntology official: harvested, nothing ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 5 pointer-only (licence)
-
Pyramid Adversarial Training Improves ViT Performance 30 Nov 2021 · 1 repository · arXiv:2111.15121
-
Building extraction with vision transformer 29 Nov 2021 · 0 repositories · arXiv:2111.15637
-
Recurrent Vision Transformer for Solving Visual Reasoning Problems 29 Nov 2021 · 0 repositories · arXiv:2111.14576
-
FQ-ViT: Post-Training Quantization for Fully Quantized Vision Transformer 27 Nov 2021 · 1 repository · arXiv:2111.13824Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Scene Representation Transformer: Geometry-Free Novel View Synthesis Through Set-Latent Scene Representations 25 Nov 2021 · 1 repository · arXiv:2111.13152Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Attention-based Dual-stream Vision Transformer for Radar Gait Recognition 24 Nov 2021 · 0 repositories · arXiv:2111.12290
-
Pruning Self-attentions into Convolutional Layers in Single Path 23 Nov 2021 · 3 repositories · arXiv:2111.11802Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Self-Supervised Pre-Training for Transformer-Based Person Re-Identification 23 Nov 2021 · 3 repositories · arXiv:2111.12084Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Benchmarking Detection Transfer Learning with Vision Transformers 22 Nov 2021 · 2 repositories · arXiv:2111.11429
-
Are Vision Transformers Robust to Patch Perturbations? 20 Nov 2021 · 0 repositories · arXiv:2111.10659
-
Rethinking Query, Key, and Value Embedding in Vision Transformer under Tiny Model Constraints 19 Nov 2021 · 0 repositories · arXiv:2111.10017
-
TransMorph: Transformer for unsupervised medical image registration 19 Nov 2021 · 2 repositories · arXiv:2111.10480Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
PatchCensor: Patch Robustness Certification for Transformers via Exhaustive Testing 19 Nov 2021 · 0 repositories · arXiv:2111.10481
-
Improved Robustness of Vision Transformer via PreLayerNorm in Patch Embedding 16 Nov 2021 · 0 repositories · arXiv:2111.08413
-
FakeTransformer: Exposing Face Forgery From Spatial-Temporal Representation Modeled By Facial Pixel Variations 15 Nov 2021 · 0 repositories · arXiv:2111.07601
-
FastFlow: Unsupervised Anomaly Detection and Localization via 2D Normalizing Flows 15 Nov 2021 · 5 repositories · arXiv:2111.07677Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
The self-supervised spectral-spatial attention-based transformer network for automated, accurate prediction of crop nitrogen status from UAV imagery 12 Nov 2021 · 0 repositories · arXiv:2111.06839
-
Sliced Recursive Transformer 9 Nov 2021 · 1 repository · arXiv:2111.05297
-
Convolutional Gated MLP: Combining Convolutions & gMLP 6 Nov 2021 · 0 repositories · arXiv:2111.03940
-
Generalized Radiograph Representation Learning via Cross-supervision between Images and Free-text Radiology Reports 4 Nov 2021 · 1 repository · arXiv:2111.03452Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
TranSMS: Transformers for Super-Resolution Calibration in Magnetic Particle Imaging 3 Nov 2021 · 2 repositories · arXiv:2111.02163
-
Can Vision Transformers Perform Convolution? 2 Nov 2021 · 0 repositories · arXiv:2111.01353
-
Federated Split Vision Transformer for COVID-19 CXR Diagnosis using Task-Agnostic Training 2 Nov 2021 · 0 repositories · arXiv:2111.01338
-
Blending Anti-Aliasing into Vision Transformer 28 Oct 2021 · 0 repositories · arXiv:2110.15156
-
Colossal-AI: A Unified Deep Learning System For Large-Scale Parallel Training 28 Oct 2021 · 1 repository · arXiv:2110.14883
-
Detecting Dementia from Speech and Transcripts using Transformers 27 Oct 2021 · 0 repositories · arXiv:2110.14769
-
Vision Transformer for Classification of Breast Ultrasound Images 27 Oct 2021 · 0 repositories · arXiv:2110.14731
-
History Aware Multimodal Transformer for Vision-and-Language Navigation 25 Oct 2021 · 1 repository · arXiv:2110.13309Syntology 7 ran (of which 3 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 15 harvested samples)
-
MVT: Multi-view Vision Transformer for 3D Object Recognition 25 Oct 2021 · 2 repositories · arXiv:2110.13083Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
CvT-ASSD: Convolutional vision-Transformer Based Attentive Single Shot MultiBox Detector 24 Oct 2021 · 1 repository · arXiv:2110.12364
-
Vis-TOP: Visual Transformer Overlay Processor 21 Oct 2021 · 0 repositories · arXiv:2110.10957
-
Bilateral-ViT for Robust Fovea Localization 19 Oct 2021 · 0 repositories · arXiv:2110.09860
-
SSAST: Self-Supervised Audio Spectrogram Transformer 19 Oct 2021 · 3 repositories · arXiv:2110.09784Syntology official (archive's flag): 1 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples) · 1 pointer-only (licence)
-
HRFormer: High-Resolution Transformer for Dense Prediction 18 Oct 2021 · 1 repository · arXiv:2110.09408Syntology official (archive's flag): 7 ran · 7 ran (of which 7 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 7 samples that ran constructed an object rather than computing a result (of 11 harvested samples)
-
MEMO: Test Time Robustness via Adaptation and Augmentation 18 Oct 2021 · 2 repositories · arXiv:2110.09506Syntology official: harvested, nothing ran · 14 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 11 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples) · 15 pointer-only (licence)
-
TLDR: Twin Learning for Dimensionality Reduction 18 Oct 2021 · 1 repository · arXiv:2110.09455
-
Transform and Bitstream Domain Image Classification 13 Oct 2021 · 0 repositories · arXiv:2110.06740
-
Dynamic Inference with Neural Interpreters 12 Oct 2021 · 0 repositories · arXiv:2110.06399
-
Certified Patch Robustness via Smoothed Vision Transformers 11 Oct 2021 · 1 repository · arXiv:2110.07719Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Global Vision Transformer Pruning with Hessian-Aware Saliency 10 Oct 2021 · 1 repository · arXiv:2110.04869Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Vision Transformer based COVID-19 Detection using Chest X-rays 9 Oct 2021 · 0 repositories · arXiv:2110.04458