Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers, page 15
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 15 of 22: papers 1,401 to 1,500 of 2,144, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Core-Periphery Principle Guided Redesign of Self-Attention in Transformers 27 Mar 2023 · 0 repositories · arXiv:2303.15569
-
Leveraging Hidden Positives for Unsupervised Semantic Segmentation 27 Mar 2023 · 1 repository · arXiv:2303.15014Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
MoViT: Memorizing Vision Transformers for Medical Image Analysis 27 Mar 2023 · 0 repositories · arXiv:2303.15553
-
Transformer-based Multi-Instance Learning for Weakly Supervised Object Detection 27 Mar 2023 · 0 repositories · arXiv:2303.14999
-
Multi-view knowledge distillation transformer for human action recognition 25 Mar 2023 · 0 repositories · arXiv:2303.14358
-
Towards Accurate Post-Training Quantization for Vision Transformer 25 Mar 2023 · 0 repositories · arXiv:2303.14341
-
FastViT: A Fast Hybrid Vision Transformer using Structural Reparameterization 24 Mar 2023 · 6 repositories · arXiv:2303.14189Syntology community repositories only · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; the one sample that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Image Deblurring by Exploring In-depth Properties of Transformer 24 Mar 2023 · 1 repository · arXiv:2303.15198
-
Boosting Convolution with Efficient MLP-Permutation for Volumetric Medical Image Segmentation 23 Mar 2023 · 1 repository · arXiv:2303.13111
-
MMFormer: Multimodal Transformer Using Multiscale Self-Attention for Remote Sensing Image Classification 23 Mar 2023 · 0 repositories · arXiv:2303.13101
-
MonoATT: Online Monocular 3D Object Detection with Adaptive Token Transformer 23 Mar 2023 · 0 repositories · arXiv:2303.13018
-
Patch-Mix Transformer for Unsupervised Domain Adaptation: A Game Perspective 23 Mar 2023 · 0 repositories · arXiv:2303.13434
-
Scaled Quantization for the Vision Transformer 23 Mar 2023 · 0 repositories · arXiv:2303.13601
-
Top-Down Visual Attention from Analysis by Synthesis 23 Mar 2023 · 1 repository · arXiv:2303.13043
-
Zero-guidance Segmentation Using Zero Segment Labels 23 Mar 2023 · 1 repository · arXiv:2303.13396
-
FeatureNeRF: Learning Generalizable NeRFs by Distilling Foundation Models 22 Mar 2023 · 1 repository · arXiv:2303.12786Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Machine Learning for Brain Disorders: Transformers and Visual Transformers 21 Mar 2023 · 0 repositories · arXiv:2303.12068
-
The Multiscale Surface Vision Transformer 21 Mar 2023 · 1 repository · arXiv:2303.11909
-
GeoMIM: Towards Better 3D Knowledge Transfer via Masked Image Modeling for Multi-view 3D Understanding 20 Mar 2023 · 1 repository · arXiv:2303.11325
-
Vision Transformer-based Model for Severity Quantification of Lung Pneumonia Using Chest X-ray Images 18 Mar 2023 · 1 repository · arXiv:2303.11935
-
Pedestrain detection for low-light vision proposal 17 Mar 2023 · 0 repositories · arXiv:2303.12725
-
Rehearsal-Free Domain Continual Face Anti-Spoofing: Generalize More and Forget Less 16 Mar 2023 · 0 repositories · arXiv:2303.09914Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
DeepMIM: Deep Supervision for Masked Image Modeling 15 Mar 2023 · 1 repository · arXiv:2303.08817
-
Query-guided Attention in Vision Transformers for Localizing Objects Using a Single Sketch 15 Mar 2023 · 0 repositories · arXiv:2303.08784
-
Efficiently Training Vision Transformers on Structural MRI Scans for Alzheimer's Disease Detection 14 Mar 2023 · 0 repositories · arXiv:2303.08216
-
Quaternion Orthogonal Transformer for Facial Expression Recognition in the Wild 14 Mar 2023 · 1 repository · arXiv:2303.07831
-
Extending global-local view alignment for self-supervised learning with remote sensing imagery 12 Mar 2023 · 1 repository · arXiv:2303.06670Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Stabilizing Transformer Training by Preventing Attention Entropy Collapse 11 Mar 2023 · 1 repository · arXiv:2303.06296Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CoordViT: A Novel Method of Improve Vision Transformer-Based Speech Emotion Recognition using Coordinate Information Concatenate 10 Mar 2023 · 0 repositories
-
Human Pose Estimation from Ambiguous Pressure Recordings with Spatio-temporal Masked Transformers 10 Mar 2023 · 0 repositories · arXiv:2303.05691
-
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection 9 Mar 2023 · 10 repositories · arXiv:2303.05499Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Mimic before Reconstruct: Enhancing Masked Autoencoders with Feature Mimicking 9 Mar 2023 · 1 repository · arXiv:2303.05475
-
Centroid-centered Modeling for Efficient Vision Transformer Pre-training 8 Mar 2023 · 1 repository · arXiv:2303.04664
-
SANDFORMER: CNN and Transformer under Gated Fusion for Sand Dust Image Restoration 8 Mar 2023 · 0 repositories · arXiv:2303.04365
-
SGDViT: Saliency-Guided Dynamic Vision Transformer for UAV Tracking 8 Mar 2023 · 1 repository · arXiv:2303.04378
-
X-Pruner: eXplainable Pruning for Vision Transformers 8 Mar 2023 · 1 repository · arXiv:2303.04935
-
Weakly Supervised Caveline Detection For AUV Navigation Inside Underwater Caves 7 Mar 2023 · 0 repositories · arXiv:2303.03670
-
UniHCP: A Unified Model for Human-Centric Perceptions 6 Mar 2023 · 1 repository · arXiv:2303.02936Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 10 pointer-only (licence)
-
DeepMAD: Mathematical Architecture Design for Deep Convolutional Neural Network 5 Mar 2023 · 1 repository · arXiv:2303.02165
-
Training-Free Acceleration of ViTs with Delayed Spatial Merging 4 Mar 2023 · 1 repository · arXiv:2303.02331
-
Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot Learners 3 Mar 2023 · 3 repositories · arXiv:2303.02151Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Retinal Image Restoration using Transformer and Cycle-Consistent Generative Adversarial Network 3 Mar 2023 · 1 repository · arXiv:2303.01939
-
Token Contrast for Weakly-Supervised Semantic Segmentation 2 Mar 2023 · 1 repository · arXiv:2303.01267Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AMIGO: Sparse Multi-Modal Graph Transformer with Shared-Context Processing for Representation Learning of Giga-pixel Images 1 Mar 2023 · 1 repository · arXiv:2303.00865
-
DC-Former: Diverse and Compact Transformer for Person Re-Identification 28 Feb 2023 · 1 repository · arXiv:2302.14335Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Remote Sensing Scene Classification with Masked Image Modeling (MIM) 28 Feb 2023 · 0 repositories · arXiv:2302.14256
-
Spatially-Adaptive Feature Modulation for Efficient Image Super-Resolution 27 Feb 2023 · 1 repository · arXiv:2302.13800Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
UMIFormer: Mining the Correlations between Similar Tokens for Multi-View 3D Reconstruction 27 Feb 2023 · 1 repository · arXiv:2302.13987
-
A Convolutional Vision Transformer for Semantic Segmentation of Side-Scan Sonar Data 24 Feb 2023 · 1 repository · arXiv:2302.12416
-
StudyFormer : Attention-Based and Dynamic Multi View Classifier for X-ray images 23 Feb 2023 · 0 repositories · arXiv:2302.11840
-
A residual dense vision transformer for medical image super-resolution with segmentation-based perceptual loss fine-tuning 22 Feb 2023 · 1 repository · arXiv:2302.11184
-
Deep Active Learning in the Presence of Label Noise: A Survey 22 Feb 2023 · 0 repositories · arXiv:2302.11075
-
Bokeh Rendering Based on Adaptive Depth Calibration Network 21 Feb 2023 · 0 repositories · arXiv:2302.10808
-
SF2Former: Amyotrophic Lateral Sclerosis Identification From Multi-center MRI Data Using Spatial and Frequency Fusion Transformer 21 Feb 2023 · 1 repository · arXiv:2302.10859
-
VITAL: Vision Transformer Neural Networks for Accurate Smartphone Heterogeneity Resilient Indoor Localization 18 Feb 2023 · 0 repositories · arXiv:2302.09443
-
EnfoMax: Domain Entropy and Mutual Information Maximization for Domain Generalized Face Anti-spoofing 17 Feb 2023 · 0 repositories · arXiv:2302.08674
-
ViTA: A Vision Transformer Inference Accelerator for Edge Applications 17 Feb 2023 · 0 repositories · arXiv:2302.09108
-
Efficiency 360: Efficient Vision Transformers 16 Feb 2023 · 1 repository · arXiv:2302.08374
-
TcGAN: Semantic-Aware and Structure-Preserved GANs with Individual Vision Transformer for Fast Arbitrary One-Shot Image Generation 16 Feb 2023 · 0 repositories · arXiv:2302.08047
-
TFormer: A Transmission-Friendly ViT Model for IoT Devices 15 Feb 2023 · 0 repositories · arXiv:2302.07734
-
DiffFashion: Reference-based Fashion Design with Structure-aware Transfer by Diffusion Models 14 Feb 2023 · 1 repository · arXiv:2302.06826
-
A Comprehensive Study of Modern Architectures and Regularization Approaches on CheXpert5000 13 Feb 2023 · 0 repositories · arXiv:2302.06684
-
Anticipating Next Active Objects for Egocentric Videos 13 Feb 2023 · 0 repositories · arXiv:2302.06358
-
VITR: Augmenting Vision Transformers with Relation-Focused Learning for Cross-Modal Information Retrieval 13 Feb 2023 · 0 repositories · arXiv:2302.06350
-
Generalized Few-Shot Continual Learning with Contrastive Mixture of Adapters 12 Feb 2023 · 1 repository · arXiv:2302.05936
-
Self-supervised pseudo-colorizing of masked cells 12 Feb 2023 · 2 repositories · arXiv:2302.05968
-
Rethinking Vision Transformer and Masked Autoencoder in Multimodal Face Anti-Spoofing 11 Feb 2023 · 0 repositories · arXiv:2302.05744
-
Reversible Vision Transformers 9 Feb 2023 · 4 repositories · arXiv:2302.04869Syntology official (archive's flag): 22 ran · 24 ran (of which 5 constructed an object rather than computing a result; 23 with no instrument failure: 0 honoured, 0 violated, 23 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 29 harvested samples) · 26 pointer-only (licence)
-
AIM: Adapting Image Models for Efficient Video Action Recognition 6 Feb 2023 · 1 repository · arXiv:2302.03024
-
V1T: large-scale mouse V1 response prediction using a Vision Transformer 6 Feb 2023 · 1 repository · arXiv:2302.03023Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 17 harvested samples)
-
Vision Transformer-based Feature Extraction for Generalized Zero-Shot Learning 2 Feb 2023 · 0 repositories · arXiv:2302.00875
-
Image-Based Vehicle Classification by Synergizing Features from Supervised and Self-Supervised Learning Paradigms 1 Feb 2023 · 0 repositories · arXiv:2302.00648
-
DepGraph: Towards Any Structural Pruning 30 Jan 2023 · 1 repository · arXiv:2301.12900Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
PhaVIP: Phage VIrion Protein classification based on chaos game representation and Vision Transformer 29 Jan 2023 · 1 repository · arXiv:2301.12422
-
Aerial Image Object Detection With Vision Transformer Detector (ViTDet) 28 Jan 2023 · 1 repository · arXiv:2301.12058
-
Voting from Nearest Tasks: Meta-Vote Pruning of Pre-trained Models for Downstream Tasks 27 Jan 2023 · 0 repositories · arXiv:2301.11560
-
Compact Transformer Tracker with Correlative Masked Modeling 26 Jan 2023 · 1 repository · arXiv:2301.10938Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Facial Expression Recognition using Squeeze and Excitation-powered Swin Transformers 26 Jan 2023 · 0 repositories · arXiv:2301.10906
-
Out of Distribution Performance of State of Art Vision Model 25 Jan 2023 · 0 repositories · arXiv:2301.10750
-
A Simple Recipe for Competitive Low-compute Self supervised Vision Models 23 Jan 2023 · 0 repositories · arXiv:2301.09451
-
Combined Use of Federated Learning and Image Encryption for Privacy-Preserving Image Classification with Vision Transformer 23 Jan 2023 · 0 repositories · arXiv:2301.09255
-
Exploring the Synergy Between Vision-Language Pretraining and ChatGPT for Artwork Captioning: A Preliminary Study 21 Jan 2023 · 1 repository
-
Image Memorability Prediction with Vision Transformers 20 Jan 2023 · 0 repositories · arXiv:2301.08647
-
MedSegDiff-V2: Diffusion based Medical Image Segmentation with Transformer 19 Jan 2023 · 2 repositories · arXiv:2301.11798Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 2 honoured, 0 violated, 10 with no contract checked; 5 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 10 pointer-only (licence)
-
Efficient Activation Function Optimization through Surrogate Modeling 13 Jan 2023 · 2 repositories · arXiv:2301.05785Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
ViTs for SITS: Vision Transformers for Satellite Image Time Series 12 Jan 2023 · 3 repositories · arXiv:2301.04944Syntology official (archive's flag): 3 ran · 11 ran (of which 3 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
Super-resolution of Ray-tracing Channel Simulation via Attention Mechanism based Deep Learning Model 11 Jan 2023 · 0 repositories · arXiv:2301.04479
-
Head-Free Lightweight Semantic Segmentation with Linear Transformer 11 Jan 2023 · 1 repository · arXiv:2301.04648
-
Dynamic Grained Encoder for Vision Transformers 10 Jan 2023 · 1 repository · arXiv:2301.03831Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Enabling Augmented Segmentation and Registration in Ultrasound-Guided Spinal Surgery via Realistic Ultrasound Synthesis from Diagnostic CT Volume 5 Jan 2023 · 0 repositories · arXiv:2301.01940
-
Single-round Self-supervised Distributed Learning using Vision Transformer 5 Jan 2023 · 0 repositories · arXiv:2301.02064
-
TinyMIM: An Empirical Study of Distilling MIM Pre-trained Models 3 Jan 2023 · 2 repositories · arXiv:2301.01296Syntology official (archive's flag): 2 ran · 8 ran (of which 5 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
A Simple Vision Transformer for Weakly Semi-supervised 3D Object Detection 1 Jan 2023 · 0 repositories
-
Adaptive and Background-Aware Vision Transformer for Real-Time UAV Tracking 1 Jan 2023 · 1 repository
-
Adversarial Normalization: I Can Visualize Everything (ICE) 1 Jan 2023 · 1 repository
-
AttentionShift: Iteratively Estimated Part-Based Attention Map for Pointly Supervised Instance Segmentation 1 Jan 2023 · 0 repositories
-
Automated Knowledge Distillation via Monte Carlo Tree Search 1 Jan 2023 · 0 repositories
-
Bit-Shrinking: Limiting Instantaneous Sharpness for Improving Post-Training Quantization 1 Jan 2023 · 0 repositories
-
Building Vision Transformers with Hierarchy Aware Feature Aggregation 1 Jan 2023 · 0 repositories
-
BUS: Efficient and Effective Vision-Language Pre-Training with Bottom-Up Patch Summarization. 1 Jan 2023 · 0 repositories