Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers, page 13
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 13 of 22: papers 1,201 to 1,300 of 2,144, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SEDA: Self-Ensembling ViT with Defensive Distillation and Adversarial Training for robust Chest X-rays Classification 15 Aug 2023 · 1 repository · arXiv:2308.07874
-
Modified Topological Image Preprocessing for Skin Lesion Classifications 13 Aug 2023 · 0 repositories · arXiv:2308.06796
-
Performance Analysis for Resource Constrained Decentralized Federated Learning Over Wireless Networks 12 Aug 2023 · 0 repositories · arXiv:2308.06496
-
Spatio-Temporal Encoding of Brain Dynamics with Surface Masked Autoencoders 10 Aug 2023 · 2 repositories · arXiv:2308.05474
-
Temporally-Adaptive Models for Efficient Video Understanding 10 Aug 2023 · 1 repository · arXiv:2308.05787Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Which Tokens to Use? Investigating Token Reduction in Vision Transformers 9 Aug 2023 · 1 repository · arXiv:2308.04657Syntology 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 3 pointer-only (licence)
-
Multiscale patch-based feature graphs for image classification 8 Aug 2023 · 1 repository
-
Temporal DINO: A Self-supervised Video Strategy to Enhance Action Prediction 8 Aug 2023 · 0 repositories · arXiv:2308.04589
-
Communication-Efficient Framework for Distributed Image Semantic Wireless Transmission 7 Aug 2023 · 0 repositories · arXiv:2308.03713
-
DiT: Efficient Vision Transformers with Dynamic Token Routing 7 Aug 2023 · 1 repository · arXiv:2308.03409
-
FLIQS: One-Shot Mixed-Precision Floating-Point and Integer Quantization Search 7 Aug 2023 · 0 repositories · arXiv:2308.03290
-
Mask Frozen-DETR: High Quality Instance Segmentation with One GPU 7 Aug 2023 · 0 repositories · arXiv:2308.03747
-
Part-Aware Transformer for Generalizable Person Re-identification 7 Aug 2023 · 1 repository · arXiv:2308.03322Syntology official (archive's flag): 10 ran · 10 ran (of which 5 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
High-Resolution Vision Transformers for Pixel-Level Identification of Structural Components and Damage 6 Aug 2023 · 0 repositories · arXiv:2308.03006
-
MCTformer+: Multi-Class Token Transformer for Weakly Supervised Semantic Segmentation 6 Aug 2023 · 1 repository · arXiv:2308.03005
-
M2Former: Multi-Scale Patch Selection for Fine-Grained Visual Recognition 4 Aug 2023 · 0 repositories · arXiv:2308.02161
-
DINO-CXR: A self supervised method based on vision transformer for chest X-ray classification 1 Aug 2023 · 0 repositories · arXiv:2308.00475
-
Improving Pixel-based MIM by Reducing Wasted Modeling Capability 1 Aug 2023 · 1 repository · arXiv:2308.00261Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
ViT2EEG: Leveraging Hybrid Pretrained Vision Transformers for EEG Data 1 Aug 2023 · 2 repositories · arXiv:2308.00454
-
StylePrompter: All Styles Need Is Attention 30 Jul 2023 · 1 repository · arXiv:2307.16151
-
CoVid-19 Detection leveraging Vision Transformers and Explainable AI 29 Jul 2023 · 0 repositories · arXiv:2307.16033
-
HandMIM: Pose-Aware Self-Supervised Learning for 3D Hand Mesh Estimation 29 Jul 2023 · 0 repositories · arXiv:2307.16061
-
Self-Supervised Graph Transformer for Deepfake Detection 27 Jul 2023 · 0 repositories · arXiv:2307.15019
-
Enhanced Security against Adversarial Examples Using a Random Ensemble of Encrypted Vision Transformer Models 26 Jul 2023 · 0 repositories · arXiv:2307.13985
-
MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation 26 Jul 2023 · 2 repositories · arXiv:2307.14460Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Understanding Deep Neural Networks via Linear Separability of Hidden Layers 26 Jul 2023 · 0 repositories · arXiv:2307.13962
-
Visual Prompt Flexible-Modal Face Anti-Spoofing 26 Jul 2023 · 0 repositories · arXiv:2307.13958
-
Conditional Cross Attention Network for Multi-Space Embedding without Entanglement in Only a SINGLE Network 25 Jul 2023 · 0 repositories · arXiv:2307.13254
-
Multi-Granularity Prediction with Learnable Fusion for Scene Text Recognition 25 Jul 2023 · 2 repositories · arXiv:2307.13244
-
A Good Student is Cooperative and Reliable: CNN-Transformer Collaborative Learning for Semantic Segmentation 24 Jul 2023 · 0 repositories · arXiv:2307.12574
-
AMAE: Adaptation of Pre-Trained Masked Autoencoder for Dual-Distribution Anomaly Detection in Chest X-Rays 24 Jul 2023 · 0 repositories · arXiv:2307.12721
-
Sparse then Prune: Toward Efficient Vision Transformers 22 Jul 2023 · 1 repository · arXiv:2307.11988
-
Latent-OFER: Detect, Mask, and Reconstruct with Latent Vectors for Occluded Facial Expression Recognition 21 Jul 2023 · 1 repository · arXiv:2307.11404Syntology official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Actor-agnostic Multi-label Action Recognition with Multi-modal Query 20 Jul 2023 · 1 repository · arXiv:2307.10763
-
Reverse Knowledge Distillation: Training a Large Model using a Small One for Retinal Image Matching on Limited Data 20 Jul 2023 · 1 repository · arXiv:2307.10698
-
The Role of Entropy and Reconstruction in Multi-View Self-Supervised Learning 20 Jul 2023 · 1 repository · arXiv:2307.10907Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Towards General Game Representations: Decomposing Games Pixels into Content and Style 20 Jul 2023 · 0 repositories · arXiv:2307.11141
-
A Step Towards Worldwide Biodiversity Assessment: The BIOSCAN-1M Insect Dataset 19 Jul 2023 · 2 repositories · arXiv:2307.10455
-
Human Action Recognition in Still Images Using ConViT 18 Jul 2023 · 0 repositories · arXiv:2307.08994
-
Light-Weight Vision Transformer with Parallel Local and Global Self-Attention 18 Jul 2023 · 0 repositories · arXiv:2307.09120
-
MOCA: Self-supervised Representation Learning by Predicting Masked Online Codebook Assignments 18 Jul 2023 · 1 repository · arXiv:2307.09361Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Study of Vision Transformers for Covid-19 Detection from Chest X-rays 17 Jul 2023 · 0 repositories · arXiv:2307.09402
-
A Survey of Techniques for Optimizing Transformer Inference 16 Jul 2023 · 0 repositories · arXiv:2307.07982
-
Dense Multitask Learning to Reconfigure Comics 16 Jul 2023 · 0 repositories · arXiv:2307.08071
-
Domain Generalisation with Bidirectional Encoder Representations from Vision Transformers 16 Jul 2023 · 0 repositories · arXiv:2307.08117
-
S2R-ViT for Multi-Agent Cooperative Perception: Bridging the Gap from Simulation to Reality 16 Jul 2023 · 0 repositories · arXiv:2307.07935
-
MaxSR: Image Super-Resolution Using Improved MaxViT 14 Jul 2023 · 0 repositories · arXiv:2307.07240
-
Deepfake Video Detection Using Generative Convolutional Vision Transformer 13 Jul 2023 · 1 repository · arXiv:2307.07036
-
Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution 12 Jul 2023 · 3 repositories · arXiv:2307.06304Syntology 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Image Reconstruction using Enhanced Vision Transformer 11 Jul 2023 · 0 repositories · arXiv:2307.05616
-
Non-Hierarchical Transformers for Pedestrian Segmentation 11 Jul 2023 · 0 repositories · arXiv:2311.02506
-
PIGEON: Predicting Image Geolocations 11 Jul 2023 · 1 repository · arXiv:2307.05845Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Distill-SODA: Distilling Self-Supervised Vision Transformer for Source-Free Open-Set Domain Adaptation in Computational Pathology 10 Jul 2023 · 2 repositories · arXiv:2307.04596
-
Cross-modal Orthogonal High-rank Augmentation for RGB-Event Transformer-trackers 9 Jul 2023 · 2 repositories · arXiv:2307.04129Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Distilling Self-Supervised Vision Transformers for Weakly-Supervised Few-Shot Classification & Segmentation 7 Jul 2023 · 0 repositories · arXiv:2307.03407
-
HoughLaneNet: Lane Detection with Deep Hough Transform and Dynamic Convolution 7 Jul 2023 · 0 repositories · arXiv:2307.03494
-
Weakly-supervised Contrastive Learning for Unsupervised Object Discovery 7 Jul 2023 · 1 repository · arXiv:2307.03376
-
MAE-DFER: Efficient Masked Autoencoder for Self-supervised Dynamic Facial Expression Recognition 5 Jul 2023 · 1 repository · arXiv:2307.02227Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 7 pointer-only (licence)
-
Make A Long Image Short: Adaptive Token Length for Vision Transformers 5 Jul 2023 · 0 repositories · arXiv:2307.02092
-
Deep Features for Contactless Fingerprint Presentation Attack Detection: Can They Be Generalized? 4 Jul 2023 · 0 repositories · arXiv:2307.01845
-
HVTSurv: Hierarchical Vision Transformer for Patient-Level Survival Prediction from Whole Slide Image 30 Jun 2023 · 1 repository · arXiv:2306.17373Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
CellViT: Vision Transformers for Precise Cell Segmentation and Classification 27 Jun 2023 · 3 repositories · arXiv:2306.15350Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Novel Hybrid-Learning Algorithms for Improved Millimeter-Wave Imaging Systems 27 Jun 2023 · 1 repository · arXiv:2306.15341
-
Taming Detection Transformers for Medical Object Detection 27 Jun 2023 · 0 repositories · arXiv:2306.15472
-
Towards predicting Pedestrian Evacuation Time and Density from Floorplans using a Vision Transformer 27 Jun 2023 · 1 repository · arXiv:2306.15318
-
FeSViBS: Federated Split Learning of Vision Transformer with Block Sampling 26 Jun 2023 · 1 repository · arXiv:2306.14638
-
Adaptive Window Pruning for Efficient Local Motion Deblurring 25 Jun 2023 · 0 repositories · arXiv:2306.14268
-
ProRes: Exploring Degradation-aware Visual Prompt for Universal Image Restoration 23 Jun 2023 · 1 repository · arXiv:2306.13653Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Swin-Free: Achieving Better Cross-Window Attention and Efficiency with Size-varying Window 23 Jun 2023 · 0 repositories · arXiv:2306.13776
-
Inter-Instance Similarity Modeling for Contrastive Learning 21 Jun 2023 · 1 repository · arXiv:2306.12243
-
Masking meets Supervision: A Strong Learning Alliance 20 Jun 2023 · 1 repository · arXiv:2306.11339
-
RaViTT: Random Vision Transformer Tokens 19 Jun 2023 · 0 repositories · arXiv:2306.10959
-
TeleViT: Teleconnection-driven Transformers Improve Subseasonal to Seasonal Wildfire Forecasting 19 Jun 2023 · 1 repository · arXiv:2306.10940
-
DEYOv2: Rank Feature with Greedy Matching for End-to-End Object Detection 15 Jun 2023 · 0 repositories · arXiv:2306.09165
-
Evaluating Data Attribution for Text-to-Image Models 15 Jun 2023 · 2 repositories · arXiv:2306.09345Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Seeing the Pose in the Pixels: Learning Pose-Aware Representations in Vision Transformers 15 Jun 2023 · 1 repository · arXiv:2306.09331
-
DiffAug: A Diffuse-and-Denoise Augmentation for Training Robust Classifiers 15 Jun 2023 · 0 repositories · arXiv:2306.09192
-
ViP: A Differentially Private Foundation Model for Computer Vision 15 Jun 2023 · 1 repository · arXiv:2306.08842Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Reviving Shift Equivariance in Vision Transformers 13 Jun 2023 · 0 repositories · arXiv:2306.07470
-
Semi-supervised learning made simple with self-supervised clustering 13 Jun 2023 · 1 repository · arXiv:2306.07483Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Enhancing COVID-19 Diagnosis through Vision Transformer-Based Analysis of Chest X-ray Images 12 Jun 2023 · 0 repositories · arXiv:2306.06914
-
Learning to Mask and Permute Visual Tokens for Vision Transformer Pre-Training 12 Jun 2023 · 1 repository · arXiv:2306.07346
-
MaskedFusion360: Reconstruct LiDAR Data by Querying Camera Features 12 Jun 2023 · 1 repository · arXiv:2306.07087
-
E(2)-Equivariant Vision Transformer 11 Jun 2023 · 1 repository · arXiv:2306.06722Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Vista-Morph: Unsupervised Image Registration of Visible-Thermal Facial Pairs 10 Jun 2023 · 0 repositories · arXiv:2306.06505
-
Customizing General-Purpose Foundation Models for Medical Report Generation 9 Jun 2023 · 0 repositories · arXiv:2306.05642
-
Connectional-Style-Guided Contextual Representation Learning for Brain Disease Diagnosis 8 Jun 2023 · 0 repositories · arXiv:2306.05297
-
Multi-Scale And Token Mergence: Make Your ViT More Efficient 8 Jun 2023 · 0 repositories · arXiv:2306.04897
-
TRIGS: Trojan Identification from Gradient-based Signatures 8 Jun 2023 · 1 repository · arXiv:2306.04877
-
Normalization Layers Are All That Sharpness-Aware Minimization Needs 7 Jun 2023 · 1 repository · arXiv:2306.04226
-
TEC-Net: Vision Transformer Embrace Convolutional Neural Networks for Medical Image Segmentation 7 Jun 2023 · 1 repository · arXiv:2306.04086
-
LESS: Label-efficient Multi-scale Learning for Cytological Whole Slide Image Screening 6 Jun 2023 · 0 repositories · arXiv:2306.03407
-
DenseDINO: Boosting Dense Self-Supervised Learning with Token-Based Point-Level Consistency 6 Jun 2023 · 0 repositories · arXiv:2306.04654
-
Industrial Anomaly Detection and Localization Using Weakly-Supervised Residual Transformers 6 Jun 2023 · 0 repositories · arXiv:2306.03492
-
Emergent Correspondence from Image Diffusion 6 Jun 2023 · 2 repositories · arXiv:2306.03881Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Human-imperceptible, Machine-recognizable Images 6 Jun 2023 · 1 repository · arXiv:2306.03679
-
A Vessel-Segmentation-Based CycleGAN for Unpaired Multi-modal Retinal Image Synthesis 5 Jun 2023 · 0 repositories · arXiv:2306.02901
-
Memorization Capacity of Multi-Head Attention in Transformers 3 Jun 2023 · 1 repository · arXiv:2306.02010
-
A Novel Vision Transformer with Residual in Self-attention for Biomedical Image Classification 2 Jun 2023 · 0 repositories · arXiv:2306.01594
-
Explainability of Speech Recognition Transformers via Gradient-based Attention Visualization 2 Jun 2023 · 1 repository