Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers, page 18
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 18 of 22: papers 1,701 to 1,800 of 2,144, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Non-Contrastive Self-Supervised Learning of Utterance-Level Speech Representations 10 Aug 2022 · 1 repository · arXiv:2208.05413
-
CoViT: Real-time phylogenetics for the SARS-CoV-2 pandemic using Vision Transformers 9 Aug 2022 · 1 repository · arXiv:2208.05004
-
Occlusion-Aware Instance Segmentation via BiLayer Network Architectures 8 Aug 2022 · 1 repository · arXiv:2208.04438
-
Analysing the Memorability of a Procedural Crime-Drama TV Series, CSI 6 Aug 2022 · 0 repositories · arXiv:2208.03479
-
DropKey 4 Aug 2022 · 0 repositories · arXiv:2208.02646
-
Self-Ensembling Vision Transformer (SEViT) for Robust Medical Image Classification 4 Aug 2022 · 1 repository · arXiv:2208.02851
-
Multi-Feature Vision Transformer via Self-Supervised Representation Learning for Improvement of COVID-19 Diagnosis 3 Aug 2022 · 1 repository · arXiv:2208.01843
-
SSformer: A Lightweight Transformer for Semantic Segmentation 3 Aug 2022 · 1 repository · arXiv:2208.02034
-
A Novel Transformer Network with Shifted Window Cross-Attention for Spatiotemporal Weather Forecasting 2 Aug 2022 · 0 repositories · arXiv:2208.01252
-
Two-Stream Transformer Architecture for Long Video Understanding 2 Aug 2022 · 0 repositories · arXiv:2208.01753
-
TransDeepLab: Convolution-Free Transformer-based DeepLab v3+ for Medical Image Segmentation 1 Aug 2022 · 1 repository · arXiv:2208.00713
-
DnSwin: Toward Real-World Denoising via Continuous Wavelet Sliding-Transformer 28 Jul 2022 · 1 repository · arXiv:2207.13861
-
Deep Clustering with Features from Self-Supervised Pretraining 27 Jul 2022 · 0 repositories · arXiv:2207.13364
-
Group DETR: Fast DETR Training with Group-Wise One-to-Many Assignment 26 Jul 2022 · 2 repositories · arXiv:2207.13085
-
V²L: Leveraging Vision and Vision-language Models into Large-scale Product Retrieval 26 Jul 2022 · 1 repository · arXiv:2207.12994
-
Jigsaw-ViT: Learning Jigsaw Puzzles in Vision Transformer 25 Jul 2022 · 1 repository · arXiv:2207.11971
-
Affective Behaviour Analysis Using Pretrained Model with Facial Priori 24 Jul 2022 · 1 repository · arXiv:2207.11679
-
Online Continual Learning with Contrastive Vision Transformer 24 Jul 2022 · 0 repositories · arXiv:2207.13516
-
Applying Spatiotemporal Attention to Identify Distracted and Drowsy Driving with Vision Transformers 22 Jul 2022 · 0 repositories · arXiv:2207.12148
-
Emotion Separation and Recognition from a Facial Expression by Generating the Poker Face with Vision Transformers 22 Jul 2022 · 0 repositories · arXiv:2207.11081
-
Focused Decoding Enables 3D Anatomical Detection by Transformers 21 Jul 2022 · 1 repository · arXiv:2207.10774
-
Locality Guidance for Improving Vision Transformers on Tiny Datasets 20 Jul 2022 · 1 repository · arXiv:2207.10026Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
MeshMAE: Masked Autoencoders for 3D Mesh Data Analysis 20 Jul 2022 · 0 repositories · arXiv:2207.10228
-
Unsupervised Industrial Anomaly Detection via Pattern Generative and Contrastive Networks 20 Jul 2022 · 0 repositories · arXiv:2207.09792
-
ViGAT: Bottom-up event recognition and explanation in video using factorized graph attention network 20 Jul 2022 · 1 repository · arXiv:2207.09927
-
Multi-manifold Attention for Vision Transformers 18 Jul 2022 · 0 repositories · arXiv:2207.08569
-
TokenMix: Rethinking Image Mixing for Data Augmentation in Vision Transformers 18 Jul 2022 · 1 repository · arXiv:2207.08409Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Explainable vision transformer enabled convolutional neural network for plant disease identification: PlantXViT 16 Jul 2022 · 0 repositories · arXiv:2207.07919
-
SSMTL++: Revisiting Self-Supervised Multi-Task Learning for Video Anomaly Detection 16 Jul 2022 · 0 repositories · arXiv:2207.08003
-
Convolutional Bypasses Are Better Vision Transformer Adapters 14 Jul 2022 · 1 repository · arXiv:2207.07039Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Current Trends in Deep Learning for Earth Observation: An Open-source Benchmark Arena for Image Classification 14 Jul 2022 · 2 repositories · arXiv:2207.07189
-
Deepfake Video Detection with Spatiotemporal Dropout Transformer 14 Jul 2022 · 0 repositories · arXiv:2207.06612
-
iColoriT: Towards Propagating Local Hint to the Right Region in Interactive Colorization by Leveraging Vision Transformer 14 Jul 2022 · 1 repository · arXiv:2207.06831
-
Scene Text Recognition with Permuted Autoregressive Sequence Models 14 Jul 2022 · 2 repositories · arXiv:2207.06966Syntology official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
eX-ViT: A Novel eXplainable Vision Transformer for Weakly Supervised Semantic Segmentation 12 Jul 2022 · 0 repositories · arXiv:2207.05358
-
Image and Model Transformation with Secret Key for Vision Transformer 12 Jul 2022 · 0 repositories · arXiv:2207.05366
-
MSP-Former: Multi-Scale Projection Transformer for Single Image Desnowing 12 Jul 2022 · 0 repositories · arXiv:2207.05621
-
Next-ViT: Next Generation Vision Transformer for Efficient Deployment in Realistic Industrial Scenarios 12 Jul 2022 · 5 repositories · arXiv:2207.05501Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Trusted Multi-Scale Classification Framework for Whole Slide Image 12 Jul 2022 · 0 repositories · arXiv:2207.05290
-
Dual Vision Transformer 11 Jul 2022 · 1 repository · arXiv:2207.04976
-
Wave-ViT: Unifying Wavelet and Transformers for Visual Representation Learning 11 Jul 2022 · 3 repositories · arXiv:2207.04978Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Consecutive Pretraining: A Knowledge Transfer Learning Strategy with Relevant Unlabeled Data for Remote Sensing Domain 8 Jul 2022 · 1 repository · arXiv:2207.03860
-
Cross-Attention Transformer for Video Interpolation 8 Jul 2022 · 1 repository · arXiv:2207.04132
-
Deep learning based Hand gesture recognition system and design of a Human-Machine Interface 7 Jul 2022 · 0 repositories · arXiv:2207.03112
-
CNN-based Local Vision Transformer for COVID-19 Diagnosis 5 Jul 2022 · 0 repositories · arXiv:2207.02027
-
Improving Semantic Segmentation in Transformers using Hierarchical Inter-Level Attention 5 Jul 2022 · 0 repositories · arXiv:2207.02126
-
Unified Object Detector for Different Modalities based on Vision Transformers 3 Jul 2022 · 1 repository · arXiv:2207.01071
-
Dissecting Self-Supervised Learning Methods for Surgical Computer Vision 1 Jul 2022 · 1 repository · arXiv:2207.00449
-
Polarized Color Image Denoising using Pocoformer 1 Jul 2022 · 0 repositories · arXiv:2207.00215
-
No Reason for No Supervision: Improved Generalization in Supervised Models 30 Jun 2022 · 1 repository · arXiv:2206.15369Syntology 7 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Multi-Channel Vision Transformer for Epileptic Seizure Prediction 29 Jun 2022 · 0 repositories
-
Cross-Forgery Analysis of Vision Transformers and CNNs for Deepfake Image Detection 28 Jun 2022 · 2 repositories · arXiv:2206.13829
-
Vision Transformer for Contrastive Clustering 26 Jun 2022 · 1 repository · arXiv:2206.12925
-
Derivative-Informed Neural Operator: An Efficient Framework for High-Dimensional Parametric Derivative Learning 21 Jun 2022 · 2 repositories · arXiv:2206.10745
-
Vicinity Vision Transformer 21 Jun 2022 · 1 repository · arXiv:2206.10552Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Global Context Vision Transformers 20 Jun 2022 · 8 repositories · arXiv:2206.09959Syntology official (archive's flag): 12 ran · 21 ran (of which 8 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 4 where Syntology's instrument failed) · 15 unverified (of 36 harvested samples) · 15 pointer-only (licence)
-
EATFormer: Improving Vision Transformer Inspired by Evolutionary Algorithm 19 Jun 2022 · 1 repository · arXiv:2206.09325Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Rectify ViT Shortcut Learning by Visual Saliency 17 Jun 2022 · 0 repositories · arXiv:2206.08567
-
Adapting Self-Supervised Vision Transformers by Probing Attention-Conditioned Masking Consistency 16 Jun 2022 · 1 repository · arXiv:2206.08222
-
OmniMAE: Single Model Masked Pretraining on Images and Videos 16 Jun 2022 · 1 repository · arXiv:2206.08356Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Patch-level Representation Learning for Self-supervised Vision Transformers 16 Jun 2022 · 1 repository · arXiv:2206.07990Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples)
-
Rethinking Generalization in Few-Shot Classification 15 Jun 2022 · 1 repository · arXiv:2206.07267Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Improving generalization by mimicking the human visual diet 15 Jun 2022 · 1 repository · arXiv:2206.07802Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring Adversarial Attacks and Defenses in Vision Transformers trained with DINO 14 Jun 2022 · 1 repository · arXiv:2206.06761
-
Stand-Alone Inter-Frame Attention in Video Models 14 Jun 2022 · 1 repository · arXiv:2206.06931
-
TransVG++: End-to-End Visual Grounding with Language Conditioned Vision Transformer 14 Jun 2022 · 1 repository · arXiv:2206.06619
-
Multimodal Learning with Transformers: A Survey 13 Jun 2022 · 0 repositories · arXiv:2206.06488
-
SeATrans: Learning Segmentation-Assisted diagnosis model via Transformer 12 Jun 2022 · 0 repositories · arXiv:2206.05763
-
Kaggle Kinship Recognition Challenge: Introduction of Convolution-Free Model to boost conventional 11 Jun 2022 · 0 repositories · arXiv:2206.05488
-
Positional Label for Self-Supervised Vision Transformer 10 Jun 2022 · 0 repositories · arXiv:2206.04981
-
Neural Prompt Search 9 Jun 2022 · 1 repository · arXiv:2206.04673Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Mask DINO: Towards A Unified Transformer-based Framework for Object Detection and Segmentation 6 Jun 2022 · 10 repositories · arXiv:2206.02777Syntology community repositories only · 11 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Sports Re-ID: Improving Re-Identification Of Players In Broadcast Videos Of Team Sports 6 Jun 2022 · 1 repository · arXiv:2206.02373
-
Federated Adversarial Training with Transformers 5 Jun 2022 · 0 repositories · arXiv:2206.02131
-
Patcher: Patch Transformers with Mixture of Experts for Precise Medical Image Segmentation 3 Jun 2022 · 1 repository · arXiv:2206.01741
-
Optimizing Relevance Maps of Vision Transformers Improves Robustness 2 Jun 2022 · 1 repository · arXiv:2206.01161
-
A comparative study between vision transformers and CNNs in digital pathology 1 Jun 2022 · 0 repositories · arXiv:2206.00389
-
Stargazer: A transformer-based driver action detection system for intelligent transportation 1 Jun 2022 · 1 repository
-
Decomposing NeRF for Editing via Feature Field Distillation 31 May 2022 · 1 repository · arXiv:2205.15585Syntology 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Surface Analysis with Vision Transformers 31 May 2022 · 1 repository · arXiv:2205.15836Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Few-Shot Diffusion Models 30 May 2022 · 1 repository · arXiv:2205.15463
-
HiViT: Hierarchical Vision Transformer Meets Masked Image Modeling 30 May 2022 · 1 repository · arXiv:2205.14949
-
Zero-Shot and Few-Shot Learning for Lung Cancer Multi-Label Classification using Vision Transformer 30 May 2022 · 0 repositories · arXiv:2205.15290
-
MDMLP: Image Classification from Scratch on Small Datasets with MLP 28 May 2022 · 2 repositories · arXiv:2205.14477
-
Momentum Stiefel Optimizer, with Applications to Suitably-Orthogonal Attention, and Optimal Transport 27 May 2022 · 1 repository · arXiv:2205.14173Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
VIDI: A Video Dataset of Incidents 26 May 2022 · 1 repository · arXiv:2205.13277
-
Eye-gaze-guided Vision Transformer for Rectifying Shortcut Learning 25 May 2022 · 0 repositories · arXiv:2205.12466
-
MoCoViT: Mobile Convolutional Vision Transformer 25 May 2022 · 1 repository · arXiv:2205.12635
-
Privacy-Preserving Image Classification Using Vision Transformer 24 May 2022 · 0 repositories · arXiv:2205.12041
-
Super Vision Transformer 23 May 2022 · 1 repository · arXiv:2205.11397
-
Vision Transformers in 2022: An Update on Tiny ImageNet 21 May 2022 · 1 repository · arXiv:2205.10660Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Learning to Count Anything: Reference-less Class-agnostic Counting with Weak Supervision 20 May 2022 · 2 repositories · arXiv:2205.10203Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Mask-guided Vision Transformer (MG-ViT) for Few-Shot Learning 20 May 2022 · 0 repositories · arXiv:2205.09995
-
Uniform Masking: Enabling MAE Pre-training for Pyramid-based Vision Transformers with Locality 20 May 2022 · 1 repository · arXiv:2205.10063
-
A graph-transformer for whole slide image classification 19 May 2022 · 1 repository · arXiv:2205.09671Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
HoVer-Trans: Anatomy-aware HoVer-Transformer for ROI-free Breast Cancer Diagnosis in Ultrasound Images 17 May 2022 · 0 repositories · arXiv:2205.08390
-
POViT: Vision Transformer for Multi-objective Design and Characterization of Nanophotonic Devices 17 May 2022 · 0 repositories · arXiv:2205.09045
-
Vision Transformer Adapter for Dense Predictions 17 May 2022 · 2 repositories · arXiv:2205.08534
-
Transformer Scale Gate for Semantic Segmentation 14 May 2022 · 0 repositories · arXiv:2205.07056
-
Simple Open-Vocabulary Object Detection with Vision Transformers 12 May 2022 · 2 repositories · arXiv:2205.06230