Methods › Computer Vision › Vision Transformers › Swin Transformer › Papers, page 4
Swin Transformer
Papers archive 2025-07-28
archive papers tagged: 416 · with a code link: 207 · where Syntology ran a sample: 58 (50 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (58 of 416 tagged: 50 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 4 of 5: papers 301 to 400 of 416, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
HST: Hierarchical Swin Transformer for Compressed Image Super-resolution 21 Aug 2022 · 3 repositories · arXiv:2208.09885
-
Shifted Windows Transformers for Medical Image Quality Assessment 11 Aug 2022 · 0 repositories · arXiv:2208.06034
-
Multi-scale Feature Aggregation for Crowd Counting 10 Aug 2022 · 0 repositories · arXiv:2208.05256
-
SSformer: A Lightweight Transformer for Semantic Segmentation 3 Aug 2022 · 1 repository · arXiv:2208.02034
-
STrajNet: Multi-modal Hierarchical Transformer for Occupancy Flow Field Prediction in Autonomous Driving 31 Jul 2022 · 1 repository · arXiv:2208.00394
-
Bodily Behaviors in Social Interaction: Novel Annotations and State-of-the-Art Evaluation 26 Jul 2022 · 0 repositories · arXiv:2207.12817
-
Combining Self-Training and Hybrid Architecture for Semi-supervised Abdominal Organ Segmentation 23 Jul 2022 · 2 repositories · arXiv:2207.11512
-
High-Resolution Swin Transformer for Automatic Medical Image Segmentation 23 Jul 2022 · 1 repository · arXiv:2207.11553
-
Applying Spatiotemporal Attention to Identify Distracted and Drowsy Driving with Vision Transformers 22 Jul 2022 · 0 repositories · arXiv:2207.12148
-
Cost Aggregation with 4D Convolutional Swin Transformer for Few-Shot Segmentation 22 Jul 2022 · 1 repository · arXiv:2207.10866Syntology official (archive's flag): 10 ran · 12 ran (of which 7 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 8 unverified (of 20 harvested samples) · 2 pointer-only (licence)
-
Video Swin Transformers for Egocentric Video Understanding @ Ego4D Challenges 2022 22 Jul 2022 · 0 repositories · arXiv:2207.11329
-
HiFormer: Hierarchical Multi-scale Representations Using Transformers for Medical Image Segmentation 18 Jul 2022 · 1 repository · arXiv:2207.08518
-
Structural Prior Guided Generative Adversarial Transformers for Low-Light Image Enhancement 16 Jul 2022 · 0 repositories · arXiv:2207.07828
-
Current Trends in Deep Learning for Earth Observation: An Open-source Benchmark Arena for Image Classification 14 Jul 2022 · 2 repositories · arXiv:2207.07189
-
Self-attention on Multi-Shifted Windows for Scene Segmentation 10 Jul 2022 · 1 repository · arXiv:2207.04403
-
Back to the Basics: Revisiting Out-of-Distribution Detection Baselines 7 Jul 2022 · 2 repositories · arXiv:2207.03061Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
More ConvNets in the 2020s: Scaling up Kernels Beyond 51x51 using Sparsity 7 Jul 2022 · 1 repository · arXiv:2207.03620
-
Improving Semantic Segmentation in Transformers using Hierarchical Inter-Level Attention 5 Jul 2022 · 0 repositories · arXiv:2207.02126
-
Deep Reinforcement Learning with Swin Transformers 30 Jun 2022 · 1 repository · arXiv:2206.15269
-
Bag of Tricks for Long-Tail Visual Recognition of Animal Species in Camera-Trap Images 24 Jun 2022 · 1 repository · arXiv:2206.12458
-
Global Context Vision Transformers 20 Jun 2022 · 8 repositories · arXiv:2206.09959Syntology official (archive's flag): 12 ran · 21 ran (of which 8 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 0 violated, 17 with no contract checked; 4 where Syntology's instrument failed) · 15 unverified (of 36 harvested samples) · 15 pointer-only (licence)
-
Rethinking Generalization in Few-Shot Classification 15 Jun 2022 · 1 repository · arXiv:2206.07267Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
SwinCheX: Multi-label classification on chest X-ray images with transformers 9 Jun 2022 · 1 repository · arXiv:2206.04246
-
Blind Face Restoration: Benchmark Datasets and a Baseline Model 8 Jun 2022 · 2 repositories · arXiv:2206.03697
-
Tutel: Adaptive Mixture-of-Experts at Scale 7 Jun 2022 · 2 repositories · arXiv:2206.03382Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
HiViT: Hierarchical Vision Transformer Meets Masked Image Modeling 30 May 2022 · 1 repository · arXiv:2205.14949
-
Green Hierarchical Vision Transformer for Masked Image Modeling 26 May 2022 · 1 repository · arXiv:2205.13515Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
MixMAE: Mixed and Masked Autoencoder for Efficient Pretraining of Hierarchical Vision Transformers 26 May 2022 · 1 repository · arXiv:2205.13137
-
MSTRIQ: No Reference Image Quality Assessment Based on Swin Transformer with Multi-Stage Fusion 20 May 2022 · 0 repositories · arXiv:2205.10101
-
MulT: An End-to-End Multitask Learning Transformer 17 May 2022 · 1 repository · arXiv:2205.08303Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
SwinIQA: Learned Swin Distance for Compressed Image Quality Assessment 9 May 2022 · 1 repository · arXiv:2205.04264
-
Reinforced Swin-Convs Transformer for Underwater Image Enhancement 1 May 2022 · 1 repository · arXiv:2205.00434Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
One Model to Synthesize Them All: Multi-contrast Multi-scale Transformer for Missing Data Imputation 28 Apr 2022 · 0 repositories · arXiv:2204.13738
-
SwinFuse: A Residual Swin Transformer Fusion Network for Infrared and Visible Images 25 Apr 2022 · 1 repository · arXiv:2204.11436
-
MANIQA: Multi-dimension Attention Network for No-Reference Image Quality Assessment 19 Apr 2022 · 2 repositories · arXiv:2204.08958Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
BSRT: Improving Burst Super-Resolution with Swin Transformer and Flow-Guided Deformable Alignment 18 Apr 2022 · 1 repository · arXiv:2204.08332Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
An Extendable, Efficient and Effective Transformer-based Object Detector 17 Apr 2022 · 1 repository · arXiv:2204.07962Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Residual Swin Transformer Channel Attention Network for Image Demosaicing 14 Apr 2022 · 0 repositories · arXiv:2204.07098
-
SwinNet: Swin Transformer drives edge-aware RGB-D and RGB-T salient object detection 12 Apr 2022 · 1 repository · arXiv:2204.05585
-
Panoptic-PartFormer: Learning a Unified Model for Panoptic Part Segmentation 10 Apr 2022 · 1 repository · arXiv:2204.04655
-
Does Robustness on ImageNet Transfer to Downstream Tasks? 8 Apr 2022 · 0 repositories · arXiv:2204.03934
-
Vision Transformers for Single Image Dehazing 8 Apr 2022 · 1 repository · arXiv:2204.03883Syntology 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Unified Contrastive Learning in Image-Text-Label Space 7 Apr 2022 · 1 repository · arXiv:2204.03610Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
MixFormer: Mixing Features across Windows and Dimensions 6 Apr 2022 · 3 repositories · arXiv:2204.02557Syntology community repositories only · 10 ran (of which 3 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
Unleashing Vanilla Vision Transformer with Masked Image Modeling for Object Detection 6 Apr 2022 · 2 repositories · arXiv:2204.02964
-
MatteFormer: Transformer-Based Image Matting via Prior-Tokens 29 Mar 2022 · 1 repository · arXiv:2203.15662Syntology official (archive's flag): 4 ran · 5 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
RSTT: Real-time Spatial Temporal Transformer for Space-Time Video Super-Resolution 27 Mar 2022 · 1 repository · arXiv:2203.14186Syntology official (archive's flag): 8 ran · 8 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Practical Blind Image Denoising via Swin-Conv-UNet and Data Synthesis 24 Mar 2022 · 2 repositories · arXiv:2203.13278Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 5 pointer-only (licence)
-
Video Instance Segmentation via Multi-scale Spatio-temporal Split Attention Transformer 24 Mar 2022 · 1 repository · arXiv:2203.13253
-
The Devil Is in the Details: Window-based Attention for Image Compression 16 Mar 2022 · 2 repositories · arXiv:2203.08450Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 3 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 15 harvested samples) · 11 pointer-only (licence)
-
Scaling Up Your Kernels to 31x31: Revisiting Large Kernel Design in CNNs 13 Mar 2022 · 8 repositories · arXiv:2203.06717Syntology official (archive's flag): 4 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
DFTR: Depth-supervised Fusion Transformer for Salient Object Detection 12 Mar 2022 · 0 repositories · arXiv:2203.06429
-
One-stage Video Instance Segmentation: From Frame-in Frame-out to Clip-in Clip-out 12 Mar 2022 · 0 repositories · arXiv:2203.06421
-
PHTrans: Parallelly Aggregating Global and Local Representations for Medical Image Segmentation 9 Mar 2022 · 2 repositories · arXiv:2203.04568
-
SUNet: Swin Transformer UNet for Image Denoising 28 Feb 2022 · 2 repositories · arXiv:2202.14009
-
Using Multi-scale SwinTransformer-HTC with Data augmentation in CoNIC Challenge 28 Feb 2022 · 0 repositories · arXiv:2202.13588
-
Real-World Blind Super-Resolution via Feature Matching with Implicit High-Resolution Priors 26 Feb 2022 · 2 repositories · arXiv:2202.13142
-
S3T: Self-Supervised Pre-training with Swin Transformer for Music Classification 21 Feb 2022 · 1 repository · arXiv:2202.10139
-
Mixing and Shifting: Exploiting Global and Local Dependencies in Vision MLPs 14 Feb 2022 · 2 repositories · arXiv:2202.06510Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 6 pointer-only (licence)
-
BViT: Broad Attention based Vision Transformer 13 Feb 2022 · 1 repository · arXiv:2202.06268
-
Generalised Image Outpainting with U-Transformer 27 Jan 2022 · 1 repository · arXiv:2201.11403
-
DSFormer: A Dual-domain Self-supervised Transformer for Accelerated Multi-contrast MRI Reconstruction 26 Jan 2022 · 0 repositories · arXiv:2201.10776
-
When Shift Operation Meets Vision Transformer: An Extremely Simple Alternative to Attention Mechanism 26 Jan 2022 · 2 repositories · arXiv:2201.10801Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 7 harvested samples)
-
Fast MRI Reconstruction: How Powerful Transformers Are? 23 Jan 2022 · 1 repository · arXiv:2201.09400
-
Q-ViT: Fully Differentiable Quantization for Vision Transformer 19 Jan 2022 · 1 repository · arXiv:2201.07703Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Swin-Pose: Swin Transformer Based Human Pose Estimation 19 Jan 2022 · 0 repositories · arXiv:2201.07384
-
Swin Transformer for Fast MRI 10 Jan 2022 · 2 repositories · arXiv:2201.03230Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Swin Transformer coupling CNNs Makes Strong Contextual Encoders for VHR Image Road Extraction 10 Jan 2022 · 0 repositories · arXiv:2201.03178
-
PyramidTNT: Improved Transformer-in-Transformer Baselines with Pyramid Architecture 4 Jan 2022 · 1 repository · arXiv:2201.00978
-
Swin UNETR: Swin Transformers for Semantic Segmentation of Brain Tumors in MRI Images 4 Jan 2022 · 3 repositories · arXiv:2201.01266
-
Vision Transformer with Deformable Attention 3 Jan 2022 · 2 repositories · arXiv:2201.00520Syntology official (archive's flag): 2 ran · 8 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Vision Transformer for Small-Size Datasets 27 Dec 2021 · 5 repositories · arXiv:2112.13492Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Raw Produce Quality Detection with Shifted Window Self-Attention 24 Dec 2021 · 0 repositories · arXiv:2112.13845
-
ELSA: Enhanced Local Self-Attention for Vision Transformer 23 Dec 2021 · 1 repository · arXiv:2112.12786Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SeMask: Semantically Masked Transformers for Semantic Segmentation 23 Dec 2021 · 1 repository · arXiv:2112.12782
-
iSegFormer: Interactive Segmentation via Transformers with Application to 3D Knee MR Images 21 Dec 2021 · 1 repository · arXiv:2112.11325
-
StyleSwin: Transformer-based GAN for High-resolution Image Generation 20 Dec 2021 · 1 repository · arXiv:2112.10762Syntology official (archive's flag): 2 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
5th Place Solution for VSPW 2021 Challenge 13 Dec 2021 · 0 repositories · arXiv:2112.06379
-
HRFormer: High-Resolution Vision Transformer for Dense Predict 1 Dec 2021 · 2 repositories
-
End-to-End Referring Video Object Segmentation with Multimodal Transformers 29 Nov 2021 · 2 repositories · arXiv:2111.14821Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
Multi-domain Integrative Swin Transformer network for Sparse-View Tomographic Reconstruction 28 Nov 2021 · 0 repositories · arXiv:2111.14831
-
SWAT: Spatial Structure Within and Among Tokens 26 Nov 2021 · 1 repository · arXiv:2111.13677
-
Global Interaction Modelling in Vision Transformer via Super Tokens 25 Nov 2021 · 0 repositories · arXiv:2111.13156
-
DBIA: Data-free Backdoor Injection Attack against Transformer Networks 22 Nov 2021 · 1 repository · arXiv:2111.11870
-
Lightweight Transformer Backbone for Medical Object Detection 22 Nov 2021 · 0 repositories · arXiv:2111.11546
-
Swin Transformer V2: Scaling Up Capacity and Resolution 18 Nov 2021 · 23 repositories · arXiv:2111.09883Syntology community repositories only · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 3 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 13 unverified (of 30 harvested samples) · 4 pointer-only (licence)
-
Transformer-based Image Compression 12 Nov 2021 · 0 repositories · arXiv:2111.06707
-
Hepatic vessel segmentation based on 3D swin-transformer with inductive biased multi-head self-attention 5 Nov 2021 · 0 repositories · arXiv:2111.03368
-
Vis-TOP: Visual Transformer Overlay Processor 21 Oct 2021 · 0 repositories · arXiv:2110.10957
-
HRFormer: High-Resolution Transformer for Dense Prediction 18 Oct 2021 · 1 repository · arXiv:2110.09408Syntology official (archive's flag): 7 ran · 7 ran (of which 7 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 7 samples that ran constructed an object rather than computing a result (of 11 harvested samples)
-
COVID-19 Detection in Chest X-ray Images Using Swin-Transformer and Transformer in Transformer 16 Oct 2021 · 0 repositories · arXiv:2110.08427
-
Satellite Image Semantic Segmentation 12 Oct 2021 · 1 repository · arXiv:2110.05812
-
ViDT: An Efficient and Effective Fully Transformer-based Object Detector 8 Oct 2021 · 1 repository · arXiv:2110.03921
-
3rd Place Scheme on Instance Segmentation Track of ICCV 2021 VIPriors Challenges 1 Oct 2021 · 0 repositories · arXiv:2110.00242
-
An object-centric sensitivity analysis of deep learning based instance segmentation 29 Sep 2021 · 0 repositories
-
Video Forgery Detection Using Multiple Cues on Fusion of EfficientNet and Swin Transformer 29 Sep 2021 · 0 repositories
-
UNetFormer: A UNet-like Transformer for Efficient Semantic Segmentation of Remote Sensing Urban Scene Imagery 18 Sep 2021 · 1 repository · arXiv:2109.08937Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Semi-Supervised Wide-Angle Portraits Correction by Multi-Scale Transformer 14 Sep 2021 · 1 repository · arXiv:2109.08024
-
Sparse MLP for Image Recognition: Is Self-Attention Really Necessary? 12 Sep 2021 · 2 repositories · arXiv:2109.05422
-
SwinIR: Image Restoration Using Swin Transformer 23 Aug 2021 · 9 repositories · arXiv:2108.10257Syntology official (archive's flag): 3 ran · 30 ran (of which 14 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 14 where Syntology's instrument failed) · 15 unverified (of 45 harvested samples) · 5 pointer-only (licence)