Methods › Computer Vision › Vision Transformers › Vision Transformer › Papers, page 9
Vision Transformer
Papers archive 2025-07-28
archive papers tagged: 2,144 · with a code link: 1,051 · where Syntology ran a sample: 328 (286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (328 of 2,144 tagged: 286 with a run with no instrument failure, 42 where every run was a failure of Syntology's instrument)
Page 9 of 22: papers 801 to 900 of 2,144, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Towards Large-Scale Training of Pathology Foundation Models 24 Mar 2024 · 2 repositories · arXiv:2404.15217Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Once for Both: Single Stage of Importance and Sparsity Search for Vision Transformer Compression 23 Mar 2024 · 1 repository · arXiv:2403.15835Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding 22 Mar 2024 · 0 repositories · arXiv:2403.15004
-
Learning with SASQuaTCh: a Novel Variational Quantum Transformer Architecture with Kernel-Based Self-Attention 21 Mar 2024 · 0 repositories · arXiv:2403.14753
-
SpikingResformer: Bridging ResNet and Vision Transformer in Spiking Neural Networks 21 Mar 2024 · 2 repositories · arXiv:2403.14302Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Token Transformation Matters: Towards Faithful Post-hoc Explanation for Vision Transformer 21 Mar 2024 · 0 repositories · arXiv:2403.14552
-
Unsupervised Audio-Visual Segmentation with Modality Alignment 21 Mar 2024 · 0 repositories · arXiv:2403.14203
-
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining 20 Mar 2024 · 2 repositories · arXiv:2403.13430Syntology official (archive's flag): 1 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer 20 Mar 2024 · 0 repositories · arXiv:2403.13570
-
Retina Vision Transformer (RetinaViT): Introducing Scaled Patches into Vision Transformers 20 Mar 2024 · 1 repository · arXiv:2403.13677
-
Rotary Position Embedding for Vision Transformer 20 Mar 2024 · 2 repositories · arXiv:2403.13298Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Emotion Recognition Using Transformers with Masked Learning 19 Mar 2024 · 1 repository · arXiv:2403.13731
-
Improved EATFormer: A Vision Transformer for Medical Image Classification 19 Mar 2024 · 0 repositories · arXiv:2403.13167
-
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs 18 Mar 2024 · 0 repositories · arXiv:2403.11999
-
SETA: Semantic-Aware Token Augmentation for Domain Generalization 18 Mar 2024 · 1 repository · arXiv:2403.11792
-
From Pixels to Predictions: Spectrogram and Vision Transformer for Better Time Series Forecasting 17 Mar 2024 · 0 repositories · arXiv:2403.11047
-
TAG: Guidance-free Open-Vocabulary Semantic Segmentation 17 Mar 2024 · 1 repository · arXiv:2403.11197
-
Magic Tokens: Select Diverse Tokens for Multi-modal Object Re-Identification 15 Mar 2024 · 2 repositories · arXiv:2403.10254Syntology official (archive's flag): 3 ran · 9 ran (of which 5 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers 15 Mar 2024 · 1 repository · arXiv:2403.10030Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Process-and-Forward: Deep Joint Source-Channel Coding Over Cooperative Relay Networks 15 Mar 2024 · 1 repository · arXiv:2403.10613
-
ViTCN: Vision Transformer Contrastive Network For Reasoning 15 Mar 2024 · 0 repositories · arXiv:2403.09962
-
Derivative-informed neural operator acceleration of geometric MCMC for infinite-dimensional Bayesian inverse problems 13 Mar 2024 · 1 repository · arXiv:2403.08220Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
METER: a mobile vision transformer architecture for monocular depth estimation 13 Mar 2024 · 1 repository · arXiv:2403.08368
-
SAP: Corrective Machine Unlearning with Scaled Activation Projection for Label Noise Robustness 13 Mar 2024 · 1 repository · arXiv:2403.08618
-
ViT-CoMer: Vision Transformer with Convolutional Multi-scale Feature Interaction for Dense Predictions 13 Mar 2024 · 2 repositories
-
A Survey of Vision Transformers in Autonomous Driving: Current Trends and Future Directions 12 Mar 2024 · 0 repositories · arXiv:2403.07542
-
ViT-CoMer: Vision Transformer with Convolutional Multi-scale Feature Interaction for Dense Predictions 12 Mar 2024 · 1 repository · arXiv:2403.07392Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Attacking Transformers with Feature Diversity Adversarial Perturbation 10 Mar 2024 · 0 repositories · arXiv:2403.07942
-
Towards In-Vehicle Multi-Task Facial Attribute Recognition: Investigating Synthetic Data and Vision Foundation Models 10 Mar 2024 · 0 repositories · arXiv:2403.06088
-
General surgery vision transformer: A video pre-trained foundation model for general surgery 9 Mar 2024 · 1 repository · arXiv:2403.05949
-
Segmentation Guided Sparse Transformer for Under-Display Camera Image Restoration 9 Mar 2024 · 0 repositories · arXiv:2403.05906
-
Self-Supervised Multiple Instance Learning for Acute Myeloid Leukemia Classification 8 Mar 2024 · 0 repositories · arXiv:2403.05379
-
Spatial-aware Transformer-GRU Framework for Enhanced Glaucoma Diagnosis from 3D OCT Imaging 8 Mar 2024 · 1 repository · arXiv:2403.05702
-
ACC-ViT : Atrous Convolution's Comeback in Vision Transformers 7 Mar 2024 · 0 repositories · arXiv:2403.04200
-
AO-DETR: Anti-Overlapping DETR for X-Ray Prohibited Items Detection 7 Mar 2024 · 1 repository · arXiv:2403.04309
-
AUFormer: Vision Transformers are Parameter-Efficient Facial Action Unit Detectors 7 Mar 2024 · 1 repository · arXiv:2403.04697
-
Yi: Open Foundation Models by 01.AI 7 Mar 2024 · 1 repository · arXiv:2403.04652Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples)
-
Multi-modal Deep Learning 6 Mar 2024 · 0 repositories · arXiv:2403.03385
-
ARNN: Attentive Recurrent Neural Network for Multi-channel EEG Signals to Identify Epileptic Seizures 5 Mar 2024 · 1 repository · arXiv:2403.03276
-
Lightweight Object Detection: A Study Based on YOLOv7 Integrated with ShuffleNetv2 and Vision Transformer 4 Mar 2024 · 0 repositories · arXiv:2403.01736
-
NiNformer: A Network in Network Transformer with Token Mixing Generated Gating Function 4 Mar 2024 · 1 repository · arXiv:2403.02411
-
Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like Architectures 4 Mar 2024 · 1 repository · arXiv:2403.02308Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 19 harvested samples) · 5 pointer-only (licence)
-
LUM-ViT: Learnable Under-sampling Mask Vision Transformer for Bandwidth Limited Optical Signal Acquisition 3 Mar 2024 · 1 repository · arXiv:2403.01412Syntology official (archive's flag): 5 ran · 5 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Deformable One-shot Face Stylization via DINO Semantic Guidance 1 Mar 2024 · 1 repository · arXiv:2403.00459
-
VisionLLaMA: A Unified LLaMA Backbone for Vision Tasks 1 Mar 2024 · 1 repository · arXiv:2403.00522Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Simple yet Effective Network based on Vision Transformer for Camouflaged Object and Salient Object Detection 29 Feb 2024 · 1 repository · arXiv:2402.18922
-
Loss-Free Machine Unlearning 29 Feb 2024 · 0 repositories · arXiv:2402.19308
-
RSAM-Seg: A SAM-based Approach with Prior Knowledge Integration for Remote Sensing Image Semantic Segmentation 29 Feb 2024 · 0 repositories · arXiv:2402.19004
-
STC-ViT: Spatio Temporal Continuous Vision Transformer for Weather Forecasting 28 Feb 2024 · 0 repositories · arXiv:2402.17966
-
Objective and Interpretable Breast Cosmesis Evaluation with Attention Guided Denoising Diffusion Anomaly Detection Model 28 Feb 2024 · 0 repositories · arXiv:2402.18362
-
Investigating the Robustness of Vision Transformers against Label Noise in Medical Image Classification 26 Feb 2024 · 0 repositories · arXiv:2402.16734
-
Exploring the Power of Pure Attention Mechanisms in Blind Room Parameter Estimation 25 Feb 2024 · 0 repositories · arXiv:2402.16003
-
One-stage Prompt-based Continual Learning 25 Feb 2024 · 0 repositories · arXiv:2402.16189
-
Attention-aware Semantic Communications for Collaborative Inference 23 Feb 2024 · 1 repository · arXiv:2404.07217
-
Multi-HMR: Multi-Person Whole-Body Human Mesh Recovery in a Single Shot 22 Feb 2024 · 1 repository · arXiv:2402.14654Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Self-supervised Visualisation of Medical Image Datasets 22 Feb 2024 · 1 repository · arXiv:2402.14566
-
EffLoc: Lightweight Vision Transformer for Efficient 6-DOF Camera Relocalization 21 Feb 2024 · 0 repositories · arXiv:2402.13537
-
ASCEND: Accurate yet Efficient End-to-End Stochastic Computing Acceleration of Vision Transformer 20 Feb 2024 · 0 repositories · arXiv:2402.12820
-
DINOBot: Robot Manipulation via Retrieval and Alignment with Vision Foundation Models 20 Feb 2024 · 0 repositories · arXiv:2402.13181
-
Quantum Embedding with Transformer for High-dimensional Data 20 Feb 2024 · 0 repositories · arXiv:2402.12704
-
FiT: Flexible Vision Transformer for Diffusion Model 19 Feb 2024 · 2 repositories · arXiv:2402.12376Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 3 pointer-only (licence)
-
Stealing the Invisible: Unveiling Pre-Trained CNN Models through Adversarial Examples and Timing Side-Channels 19 Feb 2024 · 0 repositories · arXiv:2402.11953
-
A Decoding Scheme with Successive Aggregation of Multi-Level Features for Light-Weight Semantic Segmentation 17 Feb 2024 · 0 repositories · arXiv:2402.11201
-
FViT: A Focal Vision Transformer with Gabor Filter 17 Feb 2024 · 1 repository · arXiv:2402.11303Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
BioFusionNet: Deep Learning-Based Survival Risk Stratification in ER+ Breast Cancer Through Multifeature and Multimodal Data Fusion 16 Feb 2024 · 1 repository · arXiv:2402.10717
-
Weak-Mamba-UNet: Visual Mamba Makes CNN and ViT Work Better for Scribble-based Medical Image Segmentation 16 Feb 2024 · 2 repositories · arXiv:2402.10887Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
Preserving Data Privacy for ML-driven Applications in Open Radio Access Networks 15 Feb 2024 · 0 repositories · arXiv:2402.09710
-
HEAL-ViT: Vision Transformers on a spherical mesh for medium-range weather forecasting 14 Feb 2024 · 0 repositories · arXiv:2403.17016
-
Reducing Texture Bias of Deep Neural Networks via Edge Enhancing Diffusion 14 Feb 2024 · 1 repository · arXiv:2402.09530
-
Comparative Analysis of ImageNet Pre-Trained Deep Learning Models and DINOv2 in Medical Imaging Classification 12 Feb 2024 · 1 repository · arXiv:2402.07595
-
Deciphering Heartbeat Signatures: A Vision Transformer Approach to Explainable Atrial Fibrillation Detection from ECG Signals 12 Feb 2024 · 0 repositories · arXiv:2402.09474
-
A Random Ensemble of Encrypted Vision Transformers for Adversarially Robust Defense 11 Feb 2024 · 0 repositories · arXiv:2402.07183
-
GeoFormer: A Vision and Sequence Transformer-based Approach for Greenhouse Gas Monitoring 11 Feb 2024 · 0 repositories · arXiv:2402.07164
-
Semi-Mamba-UNet: Pixel-Level Contrastive and Pixel-Level Cross-Supervised Visual Mamba-based UNet for Semi-Supervised Medical Image Segmentation 11 Feb 2024 · 1 repository · arXiv:2402.07245Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Masked LoGoNet: Fast and Accurate 3D Image Analysis for Medical Domain 9 Feb 2024 · 0 repositories · arXiv:2402.06190
-
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers 8 Feb 2024 · 2 repositories · arXiv:2402.05602
-
Memory Consolidation Enables Long-Context Video Understanding 8 Feb 2024 · 0 repositories · arXiv:2402.05861
-
On Convolutional Vision Transformers for Yield Prediction 8 Feb 2024 · 0 repositories · arXiv:2402.05557
-
Question Aware Vision Transformer for Multimodal Reasoning 8 Feb 2024 · 0 repositories · arXiv:2402.05472
-
Parameter-tuning-free data entry error unlearning with adaptive selective synaptic dampening 6 Feb 2024 · 1 repository · arXiv:2402.10098
-
Pre-training of Lightweight Vision Transformers on Small Datasets with Minimally Scaled Images 6 Feb 2024 · 0 repositories · arXiv:2402.03752
-
NeRCC: Nested-Regression Coded Computing for Resilient Distributed Prediction Serving Systems 6 Feb 2024 · 0 repositories · arXiv:2402.04377
-
Cross-Domain Few-Shot Object Detection via Enhanced Open-Set Object Detector 5 Feb 2024 · 2 repositories · arXiv:2402.03094Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
Focal Modulation Networks for Interpretable Sound Classification 5 Feb 2024 · 0 repositories · arXiv:2402.02754
-
Just Cluster It: An Approach for Exploration in High-Dimensions using Clustering and Pre-Trained Representations 5 Feb 2024 · 1 repository · arXiv:2402.03138Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Time-, Memory- and Parameter-Efficient Visual Adaptation 5 Feb 2024 · 0 repositories · arXiv:2402.02887Syntology 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
PROSAC: Provably Safe Certification for Machine Learning Models under Adversarial Attacks 4 Feb 2024 · 0 repositories · arXiv:2402.02629
-
3D Lymphoma Segmentation on PET/CT Images via Multi-Scale Information Fusion with Cross-Attention 4 Feb 2024 · 0 repositories · arXiv:2402.02349
-
ParZC: Parametric Zero-Cost Proxies for Efficient NAS 3 Feb 2024 · 0 repositories · arXiv:2402.02105
-
A Probabilistic Model Behind Self-Supervised Learning 2 Feb 2024 · 1 repository · arXiv:2402.01399Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
ALERT-Transformer: Bridging Asynchronous and Synchronous Machine Learning for Real-Time Event-based Spatio-Temporal Data 2 Feb 2024 · 0 repositories · arXiv:2402.01393
-
Faster Inference of Integer SWIN Transformer by Removing the GELU Activation 2 Feb 2024 · 0 repositories · arXiv:2402.01169
-
Dendritic Learning-incorporated Vision Transformer for Image Recognition 1 Feb 2024 · 1 repository
-
Hybrid Quantum Vision Transformers for Event Classification in High Energy Physics 1 Feb 2024 · 0 repositories · arXiv:2402.00776
-
From Training-Free to Adaptive: Empirical Insights into MLLMs' Understanding of Detection Information 31 Jan 2024 · 0 repositories · arXiv:2401.17981
-
Leveraging Swin Transformer for Local-to-Global Weakly Supervised Semantic Segmentation 31 Jan 2024 · 1 repository · arXiv:2401.17828
-
OptiState: State Estimation of Legged Robots using Gated Networks with Transformer-based Vision and Kalman Filtering 30 Jan 2024 · 1 repository · arXiv:2401.16719
-
ViTree: Single-path Neural Tree for Step-wise Interpretable Fine-grained Visual Categorization 30 Jan 2024 · 0 repositories · arXiv:2401.17050
-
Cutup and Detect: Human Fall Detection on Cutup Untrimmed Videos Using a Large Foundational Video Understanding Model 29 Jan 2024 · 0 repositories · arXiv:2401.16280
-
SHViT: Single-Head Vision Transformer with Memory Efficient Macro Design 29 Jan 2024 · 1 repository · arXiv:2401.16456Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)