Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers, page 16
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 16 of 22: papers 1,501 to 1,600 of 2,144, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CO-PILOT: Dynamic Top-Down Point Cloud with Conditional Neighborhood Aggregation for Multi-Gigapixel Histopathology Image Representation 1 Jan 2023 · 0 repositories
-
DropKey for Vision Transformer 1 Jan 2023 · 0 repositories
-
Dynamic Inference With Grounding Based Vision and Language Models 1 Jan 2023 · 0 repositories
-
FDViT: Improve the Hierarchical Architecture of Vision Transformer 1 Jan 2023 · 0 repositories
-
Goal-Guided Transformer-Enabled Reinforcement Learning for Efficient Autonomous Navigation 1 Jan 2023 · 1 repository · arXiv:2301.00362
-
Improving CLIP Fine-tuning Performance 1 Jan 2023 · 1 repository
-
Masked Auto-Encoders Meet Generative Adversarial Networks and Beyond 1 Jan 2023 · 1 repository
-
R2Former: Unified Retrieval and Reranking Transformer for Place Recognition 1 Jan 2023 · 1 repository
-
SelfME: Self-Supervised Motion Learning for Micro-Expression Recognition 1 Jan 2023 · 0 repositories
-
Sparse Multi-Modal Graph Transformer With Shared-Context Processing for Representation Learning of Giga-Pixel Images 1 Jan 2023 · 0 repositories
-
Towards Stable Human Pose Estimation via Cross-View Fusion and Foot Stabilization 1 Jan 2023 · 0 repositories
-
Trap Attention: Monocular Depth Estimation With Manual Traps 1 Jan 2023 · 1 repository
-
OVO: One-shot Vision Transformer Search with Online distillation 28 Dec 2022 · 0 repositories · arXiv:2212.13766
-
RevealED: Uncovering Pro-Eating Disorder Content on Twitter Using Deep Learning 28 Dec 2022 · 0 repositories · arXiv:2212.13949
-
Exploring Efficiency of Vision Transformers for Self-Supervised Monocular Depth Estimation 27 Dec 2022 · 1 repository
-
A Close Look at Spatial Modeling: From Attention to Convolution 23 Dec 2022 · 1 repository · arXiv:2212.12552
-
PanoViT: Vision Transformer for Room Layout Estimation from a Single Panoramic Image 23 Dec 2022 · 0 repositories · arXiv:2212.12156
-
SupeRGB-D: Zero-shot Instance Segmentation in Cluttered Indoor Environments 22 Dec 2022 · 1 repository · arXiv:2212.11922
-
Investigation of Network Architecture for Multimodal Head-and-Neck Tumor Segmentation 21 Dec 2022 · 0 repositories · arXiv:2212.10724
-
Unified Framework for Histopathology Image Augmentation and Classification via Generative Models 20 Dec 2022 · 0 repositories · arXiv:2212.09977
-
BEATs: Audio Pre-Training with Acoustic Tokenizers 18 Dec 2022 · 4 repositories · arXiv:2212.09058Syntology community repositories only · 16 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 4 pointer-only (licence)
-
Rethinking Cooking State Recognition with Vision Transformers 16 Dec 2022 · 1 repository · arXiv:2212.08586
-
Detecting Bone Lesions in X-Ray Under Diverse Acquisition Conditions 15 Dec 2022 · 0 repositories · arXiv:2212.07792
-
GPViT: A High Resolution Non-Hierarchical Vision Transformer with Group Propagation 13 Dec 2022 · 2 repositories · arXiv:2212.06795
-
PromptCAL: Contrastive Affinity Learning via Auxiliary Prompts for Generalized Novel Category Discovery 11 Dec 2022 · 1 repository · arXiv:2212.05590Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Dynamic Test-Time Augmentation via Differentiable Functions 9 Dec 2022 · 1 repository · arXiv:2212.04681
-
Sparse Upcycling: Training Mixture-of-Experts from Dense Checkpoints 9 Dec 2022 · 1 repository · arXiv:2212.05055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ViTPose++: Vision Transformer for Generic Body Pose Estimation 7 Dec 2022 · 2 repositories · arXiv:2212.04246Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Event-based Monocular Dense Depth Estimation with Recurrent Transformers 6 Dec 2022 · 0 repositories · arXiv:2212.02791
-
FacT: Factor-Tuning for Lightweight Adaptation on Vision Transformer 6 Dec 2022 · 1 repository · arXiv:2212.03145Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Semantic-aware Message Broadcasting for Efficient Unsupervised Domain Adaptation 6 Dec 2022 · 1 repository · arXiv:2212.02739
-
3D-LatentMapper: View Agnostic Single-View Reconstruction of 3D Shapes 5 Dec 2022 · 0 repositories · arXiv:2212.02184
-
Exploring Stochastic Autoregressive Image Modeling for Visual Representation 3 Dec 2022 · 1 repository · arXiv:2212.01610
-
Part-based Face Recognition with Vision Transformers 30 Nov 2022 · 1 repository · arXiv:2212.00057Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Attribute De-biased Vision Transformer (AD-ViT) for Long-Term Person Re-identification 29 Nov 2022 · 1 repository
-
Metal-conscious Embedding for CBCT Projection Inpainting 29 Nov 2022 · 0 repositories · arXiv:2211.16219
-
NoisyQuant: Noisy Bias-Enhanced Post-Training Activation Quantization for Vision Transformers 29 Nov 2022 · 1 repository · arXiv:2211.16056Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Transformer-based Hand Gesture Recognition via High-Density EMG Signals: From Instantaneous Recognition to Fusion of Motor Unit Spike Trains 29 Nov 2022 · 1 repository · arXiv:2212.00743
-
Semantic-Aware Local-Global Vision Transformer 27 Nov 2022 · 0 repositories · arXiv:2211.14705
-
Degenerate Swin to Win: Plain Window-based Transformer without Sophisticated Operations 25 Nov 2022 · 0 repositories · arXiv:2211.14255
-
Spatial-Temporal Attention Network for Open-Set Fine-Grained Image Recognition 25 Nov 2022 · 0 repositories · arXiv:2211.13940
-
TAOTF: A Two-stage Approximately Orthogonal Training Framework in Deep Neural Networks 25 Nov 2022 · 0 repositories · arXiv:2211.13902
-
Efficient Zero-shot Visual Search via Target and Context-aware Transformer 24 Nov 2022 · 0 repositories · arXiv:2211.13470
-
CODA-Prompt: COntinual Decomposed Attention-based Prompting for Rehearsal-Free Continual Learning 23 Nov 2022 · 2 repositories · arXiv:2211.13218Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Data Augmentation Vision Transformer for Fine-grained Image Classification 23 Nov 2022 · 0 repositories · arXiv:2211.12879
-
Identification of Surface Defects on Solar PV Panels and Wind Turbine Blades using Attention based Deep Learning Model 23 Nov 2022 · 0 repositories · arXiv:2211.15374
-
SVFormer: Semi-supervised Video Transformer for Action Recognition 23 Nov 2022 · 1 repository · arXiv:2211.13222
-
Generalizable Industrial Visual Anomaly Detection with Self-Induction Vision Transformer 22 Nov 2022 · 0 repositories · arXiv:2211.12311
-
MagicPony: Learning Articulated 3D Animals in the Wild 22 Nov 2022 · 0 repositories · arXiv:2211.12497
-
Teach-DETR: Better Training DETR with Teachers 22 Nov 2022 · 1 repository · arXiv:2211.11953Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 1 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
Transformer Based Multi-Grained Features for Unsupervised Person Re-Identification 22 Nov 2022 · 1 repository · arXiv:2211.12280
-
Computer Vision for Transit Travel Time Prediction: An End-to-End Framework Using Roadside Urban Imagery 22 Nov 2022 · 0 repositories · arXiv:2211.12322
-
Frozen Overparameterization: A Double Descent Perspective on Transfer Learning of Deep Neural Networks 20 Nov 2022 · 0 repositories · arXiv:2211.11074
-
CroCo v2: Improved Cross-view Completion Pre-training for Stereo Matching and Optical Flow 18 Nov 2022 · 1 repository · arXiv:2211.10408
-
Weighted Ensemble Self-Supervised Learning 18 Nov 2022 · 0 repositories · arXiv:2211.09981
-
CPT-V: A Contrastive Approach to Post-Training Quantization of Vision Transformers 17 Nov 2022 · 0 repositories · arXiv:2211.09643
-
Data-Centric Debugging: mitigating model failures via targeted data collection 17 Nov 2022 · 0 repositories · arXiv:2211.09859
-
Detecting Arbitrary Keypoints on Limbs and Skis with Sparse Partly Correct Segmentation Masks 17 Nov 2022 · 1 repository · arXiv:2211.09446
-
How to Fine-Tune Vision Models with SGD 17 Nov 2022 · 0 repositories · arXiv:2211.09359
-
Differentially Private Optimizers Can Learn Adversarially Robust Models 16 Nov 2022 · 0 repositories · arXiv:2211.08942
-
DeS3: Adaptive Attention-driven Self and Soft Shadow Removal using ViT Similarity 15 Nov 2022 · 1 repository · arXiv:2211.08089
-
Fcaformer: Forward Cross Attention in Hybrid Vision Transformer 14 Nov 2022 · 2 repositories · arXiv:2211.07198
-
Demystify Self-Attention in Vision Transformers from a Semantic Perspective: Analysis and Application 13 Nov 2022 · 0 repositories · arXiv:2211.08543
-
SSL4EO-S12: A Large-Scale Multi-Modal, Multi-Temporal Dataset for Self-Supervised Learning in Earth Observation 13 Nov 2022 · 4 repositories · arXiv:2211.07044
-
AU-Aware Vision Transformers for Biased Facial Expression Recognition 12 Nov 2022 · 0 repositories · arXiv:2211.06609
-
DEYO: DETR with YOLO for Step-by-Step Object Detection 12 Nov 2022 · 0 repositories · arXiv:2211.06588
-
End-to-End Machine Learning Framework for Facial AU Detection in Intensive Care Units 12 Nov 2022 · 0 repositories · arXiv:2211.06570
-
Masked Vision-Language Transformers for Scene Text Recognition 9 Nov 2022 · 1 repository · arXiv:2211.04785
-
Pure Transformer with Integrated Experts for Scene Text Recognition 9 Nov 2022 · 0 repositories · arXiv:2211.04963
-
Training a Vision Transformer from scratch in less than 24 hours with 1 GPU 9 Nov 2022 · 1 repository · arXiv:2211.05187Syntology 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
Transformers Meet Small Datasets 9 Nov 2022 · 0 repositories
-
Pushing the limits of self-supervised speaker verification using regularized distillation framework 8 Nov 2022 · 1 repository · arXiv:2211.04168
-
Splitting expands the application range of Vision Transformer -- variable Vision Transformer (vViT) 8 Nov 2022 · 0 repositories · arXiv:2211.03992
-
CoNMix for Source-free Single and Multi-target Domain Adaptation 7 Nov 2022 · 1 repository · arXiv:2211.03876Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MogaNet: Multi-order Gated Aggregation Network 7 Nov 2022 · 7 repositories · arXiv:2211.03295Syntology official (archive's flag): 12 ran · 12 ran (of which 7 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples)
-
Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining 7 Nov 2022 · 0 repositories · arXiv:2211.03594
-
RCDPT: Radar-Camera fusion Dense Prediction Transformer 4 Nov 2022 · 1 repository · arXiv:2211.02432
-
Evaluating a Synthetic Image Dataset Generated with Stable Diffusion 3 Nov 2022 · 0 repositories · arXiv:2211.01777
-
Rethinking Hierarchies in Pre-trained Plain Vision Transformer 3 Nov 2022 · 0 repositories · arXiv:2211.01785
-
Scaling Multimodal Pre-Training via Cross-Modality Gradient Harmonization 3 Nov 2022 · 0 repositories · arXiv:2211.02077
-
Attention-based Neural Cellular Automata 2 Nov 2022 · 0 repositories · arXiv:2211.01233
-
RegCLR: A Self-Supervised Framework for Tabular Representation Learning in the Wild 2 Nov 2022 · 0 repositories · arXiv:2211.01165
-
WITT: A Wireless Image Transmission Transformer for Semantic Communications 2 Nov 2022 · 2 repositories · arXiv:2211.00937
-
VID-Trans-ReID: Enhanced Video Transformers for Person Re-identification 1 Nov 2022 · 1 repository
-
ViT-DeiT: An Ensemble Model for Breast Cancer Histopathological Images Classification 1 Nov 2022 · 0 repositories · arXiv:2211.00749
-
ViT-LSLA: Vision Transformer with Light Self-Limited-Attention 31 Oct 2022 · 0 repositories · arXiv:2210.17115
-
Exemplar Guided Deep Neural Network for Spatial Transcriptomics Analysis of Gene Expression Prediction 30 Oct 2022 · 1 repository · arXiv:2210.16721
-
Foreign Object Debris Detection for Airport Pavement Images based on Self-supervised Localization and Vision Transformer 30 Oct 2022 · 1 repository · arXiv:2210.16901
-
ViTASD: Robust Vision Transformer Baselines for Autism Spectrum Disorder Facial Diagnosis 30 Oct 2022 · 1 repository · arXiv:2210.16943
-
Differentially Private CutMix for Split Learning with Vision Transformer 28 Oct 2022 · 0 repositories · arXiv:2210.15986
-
Elastic Weight Consolidation Improves the Robustness of Self-Supervised Learning Methods under Transfer 28 Oct 2022 · 0 repositories · arXiv:2210.16365
-
Federated Learning for Chronic Obstructive Pulmonary Disease Classification with Partial Personalized Attention Mechanism 28 Oct 2022 · 0 repositories · arXiv:2210.16142
-
HYDRA-HGR: A Hybrid Transformer-based Architecture for Fusion of Macroscopic and Microscopic Neural Drive Information 27 Oct 2022 · 0 repositories · arXiv:2211.02619
-
Masked Transformer for image Anomaly Localization 27 Oct 2022 · 0 repositories · arXiv:2210.15540
-
Masked Vision-Language Transformer in Fashion 27 Oct 2022 · 1 repository · arXiv:2210.15110
-
Spatio-Temporal Hybrid Fusion of CAE and SWIn Transformers for Lung Cancer Malignancy Prediction 27 Oct 2022 · 0 repositories · arXiv:2210.15297
-
M³ViT: Mixture-of-Experts Vision Transformer for Efficient Multi-task Learning with Model-Accelerator Co-design 26 Oct 2022 · 1 repository · arXiv:2210.14793Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Explicitly Increasing Input Information Density for Vision Transformers on Small Datasets 25 Oct 2022 · 1 repository · arXiv:2210.14319Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Minutiae-Guided Fingerprint Embeddings via Vision Transformers 25 Oct 2022 · 0 repositories · arXiv:2210.13994
-
UIA-ViT: Unsupervised Inconsistency-Aware Method based on Vision Transformer for Face Forgery Detection 23 Oct 2022 · 0 repositories · arXiv:2210.12752