Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers, page 17
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 17 of 22: papers 1,601 to 1,700 of 2,144, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
S2WAT: Image Style Transfer via Hierarchical Vision Transformer using Strips Window Attention 22 Oct 2022 · 1 repository · arXiv:2210.12381
-
Face Pyramid Vision Transformer 21 Oct 2022 · 1 repository · arXiv:2210.11974
-
GPR-Net: Multi-view Layout Estimation via a Geometry-aware Panorama Registration Network 20 Oct 2022 · 0 repositories · arXiv:2210.11419
-
SimpleClick: Interactive Image Segmentation with Simple Vision Transformers 20 Oct 2022 · 2 repositories · arXiv:2210.11006Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
A Unified View of Masked Image Modeling 19 Oct 2022 · 1 repository · arXiv:2210.10615
-
Multi-view Gait Recognition based on Siamese Vision Transformer 19 Oct 2022 · 0 repositories · arXiv:2210.10421
-
Sequence and Circle: Exploring the Relationship Between Patches 18 Oct 2022 · 0 repositories · arXiv:2210.09871
-
Histopathological Image Classification based on Self-Supervised Vision Transformer and Weak Labels 17 Oct 2022 · 1 repository · arXiv:2210.09021
-
Distributionally Robust Multiclass Classification and Applications in Deep Image Classifiers 15 Oct 2022 · 0 repositories · arXiv:2210.08198
-
Transformer-based dimensionality reduction 15 Oct 2022 · 0 repositories · arXiv:2210.08288
-
MOVE: Unsupervised Movable Object Segmentation and Detection 14 Oct 2022 · 1 repository · arXiv:2210.07920Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Optimizing Vision Transformers for Medical Image Segmentation 14 Oct 2022 · 1 repository · arXiv:2210.08066
-
CAP: Correlation-Aware Pruning for Highly-Accurate Sparse Vision Models 14 Oct 2022 · 0 repositories · arXiv:2210.09223
-
Feature-Proxy Transformer for Few-Shot Segmentation 13 Oct 2022 · 2 repositories · arXiv:2210.06908
-
How to Train Vision Transformer on Small-scale Datasets? 13 Oct 2022 · 2 repositories · arXiv:2210.07240
-
Bridging the Gap Between Vision Transformers and Convolutional Neural Networks on Small Datasets 12 Oct 2022 · 1 repository · arXiv:2210.05958Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 5 pointer-only (licence)
-
ACSeg: Adaptive Conceptualization for Unsupervised Semantic Segmentation 12 Oct 2022 · 0 repositories · arXiv:2210.05944
-
S4ND: Modeling Images and Videos as Multidimensional Signals Using State Spaces 12 Oct 2022 · 1 repository · arXiv:2210.06583Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Towards Theoretically Inspired Neural Initialization Optimization 12 Oct 2022 · 1 repository · arXiv:2210.05956
-
SaiT: Sparse Vision Transformers through Adaptive Token Pruning 11 Oct 2022 · 1 repository · arXiv:2210.05832
-
Revisiting adapters with adversarial training 10 Oct 2022 · 0 repositories · arXiv:2210.04886
-
Visual Prompt Tuning for Test-time Domain Adaptation 10 Oct 2022 · 0 repositories · arXiv:2210.04831
-
Strong Gravitational Lensing Parameter Estimation with Vision Transformer 9 Oct 2022 · 1 repository · arXiv:2210.04143Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Transformer-based Flood Scene Segmentation for Developing Countries 9 Oct 2022 · 0 repositories · arXiv:2210.04218
-
Gastrointestinal Disorder Detection with a Transformer Based Approach 6 Oct 2022 · 0 repositories · arXiv:2210.03168
-
LungViT: Ensembling Cascade of Texture Sensitive Hierarchical Vision Transformers for Cross-Volume Chest CT Image-to-Image Translation 6 Oct 2022 · 0 repositories · arXiv:2210.02625
-
Real-World Robot Learning with Masked Visual Pre-training 6 Oct 2022 · 1 repository · arXiv:2210.03109
-
Structure Representation Network and Uncertainty Feedback Learning for Dense Non-Uniform Fog Removal 6 Oct 2022 · 1 repository · arXiv:2210.03061Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SynBench: Task-Agnostic Benchmarking of Pretrained Representations using Synthetic Data 6 Oct 2022 · 0 repositories · arXiv:2210.02989
-
Vision Transformer Based Model for Describing a Set of Images as a Story 6 Oct 2022 · 0 repositories · arXiv:2210.02762
-
Exploring The Role of Mean Teachers in Self-supervised Masked Auto-Encoders 5 Oct 2022 · 1 repository · arXiv:2210.02077
-
K-means for unsupervised instance segmentation using a self-supervised transformer 4 Oct 2022 · 0 repositories
-
Early or Late Fusion Matters: Efficient RGB-D Fusion in Vision Transformers for 3D Object Recognition 3 Oct 2022 · 0 repositories · arXiv:2210.00843
-
Enhancing Fine-Grained 3D Object Recognition using Hybrid Multi-Modal Vision Transformer-CNN Models 3 Oct 2022 · 1 repository · arXiv:2210.04613
-
Introducing Vision Transformer for Alzheimer's Disease classification task with 3D input 3 Oct 2022 · 0 repositories · arXiv:2210.01177
-
Deep-OCTA: Ensemble Deep Learning Approaches for Diabetic Retinopathy Analysis on OCTA Images 2 Oct 2022 · 1 repository · arXiv:2210.00515
-
Diffusion-based Image Translation using Disentangled Style and Content Representation 30 Sep 2022 · 1 repository · arXiv:2209.15264Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Dual Progressive Transformations for Weakly Supervised Semantic Segmentation 30 Sep 2022 · 1 repository · arXiv:2209.15211
-
Impact of Face Image Quality Estimation on Presentation Attack Detection 30 Sep 2022 · 0 repositories · arXiv:2209.15489
-
MobileViTv3: Mobile-Friendly Vision Transformer with Simple and Effective Fusion of Local, Global and Input Features 30 Sep 2022 · 2 repositories · arXiv:2209.15159
-
Self-Distillation for Further Pre-training of Transformers 30 Sep 2022 · 0 repositories · arXiv:2210.02871
-
Where Should I Spend My FLOPS? Efficiency Evaluations of Visual Pre-training Methods 30 Sep 2022 · 0 repositories · arXiv:2209.15589
-
Continuous PDE Dynamics Forecasting with Implicit Neural Representations 29 Sep 2022 · 1 repository · arXiv:2209.14855Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
360FusionNeRF: Panoramic Neural Radiance Fields with Joint Guidance 28 Sep 2022 · 1 repository · arXiv:2209.14265
-
Attacking Compressed Vision Transformers 28 Sep 2022 · 1 repository · arXiv:2209.13785
-
MTU-Net: Multi-level TransUNet for Space-based Infrared Tiny Ship Detection 28 Sep 2022 · 1 repository · arXiv:2209.13756
-
Collaboration of Pre-trained Models Makes Better Few-shot Learner 25 Sep 2022 · 0 repositories · arXiv:2209.12255
-
NasHD: Efficient ViT Architecture Performance Ranking using Hyperdimensional Computing 23 Sep 2022 · 0 repositories · arXiv:2209.11356
-
Wide-Area Geolocalization with a Limited Field of View Camera 23 Sep 2022 · 0 repositories · arXiv:2209.11854
-
Colonoscopy Landmark Detection using Vision Transformers 22 Sep 2022 · 0 repositories · arXiv:2209.11304
-
NamedMask: Distilling Segmenters from Complementary Foundation Models 22 Sep 2022 · 1 repository · arXiv:2209.11228Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Pretraining the Vision Transformer using self-supervised methods for vision based Deep Reinforcement Learning 22 Sep 2022 · 1 repository · arXiv:2209.10901
-
PicT: A Slim Weakly Supervised Vision Transformer for Pavement Distress Classification 21 Sep 2022 · 1 repository · arXiv:2209.10074
-
Revisiting Image Pyramid Structure for High Resolution Salient Object Detection 20 Sep 2022 · 3 repositories · arXiv:2209.09475Syntology official (archive's flag): 8 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples)
-
Attentive Symmetric Autoencoder for Brain MRI Segmentation 19 Sep 2022 · 1 repository · arXiv:2209.08887
-
HiMFR: A Hybrid Masked Face Recognition Through Face Inpainting 19 Sep 2022 · 0 repositories · arXiv:2209.08930
-
ViT-DD: Multi-Task Vision Transformer for Semi-Supervised Driver Distraction Detection 19 Sep 2022 · 1 repository · arXiv:2209.09178
-
Panoramic Vision Transformer for Saliency Detection in 360° Videos 19 Sep 2022 · 1 repository · arXiv:2209.08956Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Uncertainty Aware Multitask Pyramid Vision Transformer For UAV-Based Object Re-Identification 19 Sep 2022 · 0 repositories · arXiv:2209.08686
-
EEG-Based Epileptic Seizure Prediction Using Temporal Multi-Channel Transformers 18 Sep 2022 · 0 repositories · arXiv:2209.11172
-
A Mosquito is Worth 16x16 Larvae: Evaluation of Deep Learning Architectures for Mosquito Larvae Classification 16 Sep 2022 · 1 repository · arXiv:2209.07718
-
Hybrid Window Attention Based Transformer Architecture for Brain Tumor Segmentation 16 Sep 2022 · 1 repository · arXiv:2209.07704
-
PPT: token-Pruned Pose Transformer for monocular and multi-view human pose estimation 16 Sep 2022 · 2 repositories · arXiv:2209.08194
-
Self-Supervised Learning of Phenotypic Representations from Cell Images with Weak Labels 16 Sep 2022 · 1 repository · arXiv:2209.07819
-
PriorLane: A Prior Knowledge Enhanced Lane Detection Approach Based on Transformer 15 Sep 2022 · 1 repository · arXiv:2209.06994
-
A lightweight Transformer-based model for fish landmark detection 13 Sep 2022 · 0 repositories · arXiv:2209.05777
-
DMTNet: Dynamic Multi-scale Network for Dual-pixel Images Defocus Deblurring with Transformer 13 Sep 2022 · 0 repositories · arXiv:2209.06040
-
Vision Transformers for Action Recognition: A Survey 13 Sep 2022 · 0 repositories · arXiv:2209.05700
-
OpenMixup: Open Mixup Toolbox and Benchmark for Visual Representation Learning 11 Sep 2022 · 1 repository · arXiv:2209.04851
-
Multi-Granularity Prediction for Scene Text Recognition 8 Sep 2022 · 3 repositories · arXiv:2209.03592
-
Video Vision Transformers for Violence Detection 8 Sep 2022 · 0 repositories · arXiv:2209.03561
-
Prior Knowledge-Guided Attention in Self-Supervised Vision Transformers 7 Sep 2022 · 0 repositories · arXiv:2209.03745
-
Transfer Learning and Vision Transformer based State-of-Health prediction of Lithium-Ion Batteries 7 Sep 2022 · 0 repositories · arXiv:2209.05253
-
Fusion of Satellite Images and Weather Data with Transformer Networks for Downy Mildew Disease Detection 6 Sep 2022 · 0 repositories · arXiv:2209.02797
-
Transformer-CNN Cohort: Semi-supervised Semantic Segmentation by the Best of Both Students 6 Sep 2022 · 0 repositories · arXiv:2209.02178
-
ViTKD: Practical Guidelines for ViT feature knowledge distillation 6 Sep 2022 · 1 repository · arXiv:2209.02432
-
Time-distance vision transformers in lung cancer diagnosis from longitudinal computed tomography 4 Sep 2022 · 1 repository · arXiv:2209.01676
-
EViT: Privacy-Preserving Image Retrieval via Encrypted Vision Transformer in Cloud Computing 31 Aug 2022 · 1 repository · arXiv:2208.14657
-
SIM-Trans: Structure Information Modeling Transformer for Fine-grained Visual Categorization 31 Aug 2022 · 1 repository · arXiv:2208.14607
-
Light-YOLOv5: A Lightweight Algorithm for Improved YOLOv5 in Complex Fire Scenarios 29 Aug 2022 · 0 repositories · arXiv:2208.13422
-
Open-Set Semi-Supervised Object Detection 29 Aug 2022 · 0 repositories · arXiv:2208.13722
-
An Access Control Method with Secret Key for Semantic Segmentation Models 28 Aug 2022 · 0 repositories · arXiv:2208.13135
-
VMFormer: End-to-End Video Matting with Transformer 26 Aug 2022 · 1 repository · arXiv:2208.12801
-
A Deep Learning Approach Using Masked Image Modeling for Reconstruction of Undersampled K-spaces 24 Aug 2022 · 1 repository · arXiv:2208.11472
-
Federated Self-Supervised Contrastive Learning and Masked Autoencoder for Dermatological Disease Diagnosis 24 Aug 2022 · 0 repositories · arXiv:2208.11278
-
Predicting microsatellite instability and key biomarkers in colorectal cancer from H&E-stained images: Achieving SOTA predictive performance with fewer data using Swin Transformer 22 Aug 2022 · 0 repositories · arXiv:2208.10495
-
ProtoPFormer: Concentrating on Prototypical Parts in Vision Transformers for Interpretable Image Recognition 22 Aug 2022 · 1 repository · arXiv:2208.10431Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Accelerating Vision Transformer Training via a Patch Sampling Schedule 19 Aug 2022 · 1 repository · arXiv:2208.09520
-
The 8-Point Algorithm as an Inductive Bias for Relative Pose Prediction by ViTs 18 Aug 2022 · 0 repositories · arXiv:2208.08988
-
Conviformers: Convolutionally guided Vision Transformer 17 Aug 2022 · 1 repository · arXiv:2208.08900
-
Video-TransUNet: Temporally Blended Vision Transformer for CT VFSS Instance Segmentation 17 Aug 2022 · 2 repositories · arXiv:2208.08315
-
ViT-ReT: Vision and Recurrent Transformer Neural Networks for Human Activity Recognition in Videos 16 Aug 2022 · 0 repositories · arXiv:2208.07929
-
Your ViT is Secretly a Hybrid Discriminative-Generative Diffusion Model 16 Aug 2022 · 2 repositories · arXiv:2208.07791Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
A Vision Transformer-Based Approach to Bearing Fault Classification via Vibration Signals 15 Aug 2022 · 0 repositories · arXiv:2208.07070
-
Self-Supervised Vision Transformers for Malware Detection 15 Aug 2022 · 1 repository · arXiv:2208.07049Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 4 pointer-only (licence)
-
Shuffle Instances-based Vision Transformer for Pancreatic Cancer ROSE Image Classification 14 Aug 2022 · 1 repository · arXiv:2208.06833
-
BEiT v2: Masked Image Modeling with Vector-Quantized Visual Tokenizers 12 Aug 2022 · 3 repositories · arXiv:2208.06366
-
Shifted Windows Transformers for Medical Image Quality Assessment 11 Aug 2022 · 0 repositories · arXiv:2208.06034
-
Ghost-free High Dynamic Range Imaging with Context-aware Transformer 10 Aug 2022 · 3 repositories · arXiv:2208.05114Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Non-Contrastive Self-supervised Learning for Utterance-Level Information Extraction from Speech 10 Aug 2022 · 0 repositories · arXiv:2208.05445