Methods › Computer Vision › Vision Transformers › Swin Transformer › Papers, page 2
Swin Transformer
Papers archive 2025-07-28
archive papers tagged: 416 · with a code link: 207 · where Syntology ran a sample: 58 (50 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (58 of 416 tagged: 50 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 2 of 5: papers 101 to 200 of 416, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Image-to-LaTeX Converter for Mathematical Formulas and Text 7 Aug 2024 · 1 repository · arXiv:2408.04015
-
SwinShadow: Shifted Window for Ambiguous Adjacent Shadow Detection 7 Aug 2024 · 1 repository · arXiv:2408.03521
-
POA: Pre-training Once for Models of All Sizes 2 Aug 2024 · 1 repository · arXiv:2408.01031
-
FSSC: Federated Learning of Transformer Neural Networks for Semantic Image Communication 31 Jul 2024 · 0 repositories · arXiv:2407.21507
-
Semantic Successive Refinement: A Generative AI-aided Semantic Communication Framework 31 Jul 2024 · 0 repositories · arXiv:2408.05112
-
Domain Generalized Recaptured Screen Image Identification Using SWIN Transformer 24 Jul 2024 · 0 repositories · arXiv:2407.17170
-
Improving Representation of High-frequency Components for Medical Visual Foundation Models 19 Jul 2024 · 1 repository · arXiv:2407.14651Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 1 pointer-only (licence)
-
Deep(er) Reconstruction of Imaging Cherenkov Detectors with Swin Transformers and Normalizing Flow Models 10 Jul 2024 · 1 repository · arXiv:2407.07376
-
Cross-Modal Spherical Aggregation for Weakly Supervised Remote Sensing Shadow Removal 25 Jun 2024 · 1 repository · arXiv:2406.17469
-
SUM: Saliency Unification through Mamba for Visual Attention Modeling 25 Jun 2024 · 1 repository · arXiv:2406.17815Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples)
-
MixTex: Unambiguous Recognition Should Not Rely Solely on Real Data 24 Jun 2024 · 1 repository · arXiv:2406.17148
-
SwinStyleformer is a favorable choice for image inversion 19 Jun 2024 · 0 repositories · arXiv:2406.13153
-
Diffusion-Based Adaptation for Classification of Unknown Degraded Images 17 Jun 2024 · 1 repository
-
QMamba: On First Exploration of Vision Mamba for Image Quality Assessment 13 Jun 2024 · 1 repository · arXiv:2406.09546
-
A Robust Pipeline for Classification and Detection of Bleeding Frames in Wireless Capsule Endoscopy using Swin Transformer and RT-DETR 12 Jun 2024 · 0 repositories · arXiv:2406.08046
-
SuperFormer: Volumetric Transformer Architectures for MRI Super-Resolution 5 Jun 2024 · 1 repository · arXiv:2406.03359
-
Robust Image Semantic Coding with Learnable CSI Fusion Masking over MIMO Fading Channels 30 May 2024 · 0 repositories · arXiv:2406.07389
-
YotoR-You Only Transform One Representation 30 May 2024 · 0 repositories · arXiv:2405.19629
-
MDS-ViTNet: Improving saliency prediction for Eye-Tracking with Vision Transformer 29 May 2024 · 1 repository · arXiv:2405.19501
-
Self-Supervised Modality-Agnostic Pre-Training of Swin Transformers 21 May 2024 · 1 repository · arXiv:2405.12781
-
Ground-based image deconvolution with Swin Transformer UNet 13 May 2024 · 0 repositories · arXiv:2405.07842
-
Super-Resolving Blurry Images with Events 11 May 2024 · 0 repositories · arXiv:2405.06918
-
SatSwinMAE: Efficient Autoencoding for Multiscale Time-series Satellite Imagery 3 May 2024 · 0 repositories · arXiv:2405.02512
-
Technical report on target classification in SAR track 3 May 2024 · 0 repositories · arXiv:2405.02361
-
Transformers Fusion across Disjoint Samples for Hyperspectral Image Classification 2 May 2024 · 0 repositories · arXiv:2405.01095
-
StrideNET: Swin Transformer for Terrain Recognition with Dynamic Roughness Extraction 20 Apr 2024 · 0 repositories · arXiv:2404.13270
-
Towards Robust Ferrous Scrap Material Classification with Deep Learning and Conformal Prediction 19 Apr 2024 · 0 repositories · arXiv:2404.13002
-
Deep Learning and LLM-based Methods Applied to Stellar Lightcurve Classification 16 Apr 2024 · 1 repository · arXiv:2404.10757
-
ODFormer: Semantic Fundus Image Segmentation Using Transformer for Optic Nerve Head Detection 15 Apr 2024 · 0 repositories · arXiv:2405.09552
-
Rethinking Low-Rank Adaptation in Vision: Exploring Head-Level Responsiveness across Diverse Tasks 13 Apr 2024 · 0 repositories · arXiv:2404.08894
-
Scalability in Building Component Data Annotation: Enhancing Facade Material Classification with Synthetic Data 12 Apr 2024 · 0 repositories · arXiv:2404.08557
-
Post-hurricane building damage assessment using street-view imagery and structured data: A multi-modal deep learning approach 11 Apr 2024 · 0 repositories · arXiv:2404.07399
-
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields 1 Apr 2024 · 1 repository · arXiv:2404.01300Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs 28 Mar 2024 · 3 repositories · arXiv:2403.19588Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Residual Dense Swin Transformer for Continuous Depth-Independent Ultrasound Imaging 25 Mar 2024 · 1 repository · arXiv:2403.16384
-
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding 22 Mar 2024 · 0 repositories · arXiv:2403.15004
-
High-confidence pseudo-labels for domain adaptation in COVID-19 detection 20 Mar 2024 · 0 repositories · arXiv:2403.13509
-
Attention-Enhanced Hybrid Feature Aggregation Network for 3D Brain Tumor Segmentation 15 Mar 2024 · 1 repository · arXiv:2403.09942
-
A Spatio-temporal Aligned SUNet Model for Low-light Video Enhancement 4 Mar 2024 · 0 repositories · arXiv:2403.02408
-
Enhancing Retinal Vascular Structure Segmentation in Images With a Novel Design Two-Path Interactive Fusion Module Model 3 Mar 2024 · 1 repository · arXiv:2403.01362
-
Automated segmentation of lesions and organs at risk on [68Ga]Ga-PSMA-11 PET/CT images using self-supervised learning with Swin UNETR 29 Feb 2024 · 2 repositories
-
BEFUnet: A Hybrid CNN-Transformer Architecture for Precise Medical Image Segmentation 13 Feb 2024 · 1 repository · arXiv:2402.08793
-
Faster Inference of Integer SWIN Transformer by Removing the GELU Activation 2 Feb 2024 · 0 repositories · arXiv:2402.01169
-
Leveraging Swin Transformer for Local-to-Global Weakly Supervised Semantic Segmentation 31 Jan 2024 · 1 repository · arXiv:2401.17828
-
A New Method for Vehicle Logo Recognition Based on Swin Transformer 27 Jan 2024 · 0 repositories · arXiv:2401.15458
-
Adversarial Augmentation Training Makes Action Recognition Models More Robust to Realistic Video Distribution Shifts 21 Jan 2024 · 1 repository · arXiv:2401.11406
-
LRP-QViT: Mixed-Precision Vision Transformer Quantization via Layer-wise Relevance Propagation 20 Jan 2024 · 0 repositories · arXiv:2401.11243
-
Trapped in texture bias? A large scale comparison of deep instance segmentation 17 Jan 2024 · 1 repository · arXiv:2401.09109
-
B-Cos Aligned Transformers Learn Human-Interpretable Features 16 Jan 2024 · 0 repositories · arXiv:2401.08868
-
Importance-Aware Image Segmentation-based Semantic Communication for Autonomous Driving 16 Jan 2024 · 0 repositories · arXiv:2401.10153
-
Video Quality Assessment Based on Swin TransformerV2 and Coarse to Fine Strategy 16 Jan 2024 · 0 repositories · arXiv:2401.08522
-
DedustNet: A Frequency-dominated Swin Transformer-based Wavelet Network for Agricultural Dust Removal 9 Jan 2024 · 0 repositories · arXiv:2401.04750
-
A novel method to enhance pneumonia detection via a model-level ensembling of CNN and vision transformer 4 Jan 2024 · 0 repositories · arXiv:2401.02358
-
MOC-RVQ: Multilevel Codebook-Assisted Digital Generative Semantic Communication 2 Jan 2024 · 1 repository · arXiv:2401.01272
-
ParameterNet: Parameters Are All You Need for Large-scale Visual Pretraining of Mobile Networks 1 Jan 2024 · 0 repositories
-
Image Super-resolution Reconstruction Network based on Enhanced Swin Transformer via Alternating Aggregation of Local-Global Features 30 Dec 2023 · 0 repositories · arXiv:2401.00241
-
C2T-Net: Channel-Aware Cross-Fused Transformer-Style Networks for Pedestrian Attribute Recognition 26 Dec 2023 · 1 repository
-
SCUNet++: Swin-UNet and CNN Bottleneck Hybrid Architecture with Multi-Fusion Dense Skip Connection for Pulmonary Embolism CT Image Segmentation 22 Dec 2023 · 1 repository · arXiv:2312.14705
-
TPTNet: A Data-Driven Temperature Prediction Model Based on Turbulent Potential Temperature 22 Dec 2023 · 0 repositories · arXiv:2312.14980
-
Research on Multilingual Natural Scene Text Detection Algorithm 18 Dec 2023 · 0 repositories · arXiv:2312.11153
-
Factorization Vision Transformer: Modeling Long Range Dependency with Local Window Cost 14 Dec 2023 · 1 repository · arXiv:2312.08614
-
Polyper: Boundary Sensitive Polyp Segmentation 14 Dec 2023 · 1 repository · arXiv:2312.08735
-
Intelligent Anomaly Detection for Lane Rendering Using Transformer with Self-Supervised Pre-Training and Customized Fine-Tuning 7 Dec 2023 · 0 repositories · arXiv:2312.04398
-
Self-training solutions for the ICCV 2023 GeoNet Challenge 28 Nov 2023 · 1 repository · arXiv:2311.16843
-
EAFP-Med: An Efficient Adaptive Feature Processing Module Based on Prompts for Medical Image Detection 27 Nov 2023 · 0 repositories · arXiv:2311.15540
-
Progressive Learning with Visual Prompt Tuning for Variable-Rate Image Compression 23 Nov 2023 · 0 repositories · arXiv:2311.13846
-
BenthIQ: a Transformer-Based Benthic Classification Model for Coral Restoration 22 Nov 2023 · 0 repositories · arXiv:2311.13661
-
Feature Extraction for Generative Medical Imaging Evaluation: New Evidence Against an Evolving Trend 22 Nov 2023 · 2 repositories · arXiv:2311.13717
-
PMP-Swin: Multi-Scale Patch Message Passing Swin Transformer for Retinal Disease Classification 20 Nov 2023 · 0 repositories · arXiv:2311.11669
-
Inspecting Explainability of Transformer Models with Additional Statistical Information 19 Nov 2023 · 0 repositories · arXiv:2311.11378
-
Wildfire Smoke Detection with Cross Contrast Patch Embedding 16 Nov 2023 · 1 repository · arXiv:2311.10116
-
Two Stream Scene Understanding on Graph Embedding 12 Nov 2023 · 0 repositories · arXiv:2311.06746
-
Distilling Knowledge from CNN-Transformer Models for Enhanced Human Action Recognition 2 Nov 2023 · 0 repositories · arXiv:2311.01283
-
Detecting Visual Cues in the Intensive Care Unit and Association with Patient Clinical Status 1 Nov 2023 · 0 repositories · arXiv:2311.00565
-
FaultSeg Swin-UNETR: Transformer-Based Self-Supervised Pretraining Model for Fault Recognition 27 Oct 2023 · 0 repositories · arXiv:2310.17974
-
Multimodal Transformer Using Cross-Channel attention for Object Detection in Remote Sensing Images 21 Oct 2023 · 1 repository · arXiv:2310.13876
-
Auxiliary Features-Guided Super Resolution for Monte Carlo Rendering 20 Oct 2023 · 0 repositories · arXiv:2310.13235
-
DIAR: Deep Image Alignment and Reconstruction using Swin Transformers 17 Oct 2023 · 0 repositories · arXiv:2310.11605
-
Mask wearing object detection algorithm based on improved YOLOv5 16 Oct 2023 · 0 repositories · arXiv:2310.10245
-
COVID-19 detection using ViT transformer-based approach from Computed Tomography Images 12 Oct 2023 · 1 repository · arXiv:2310.08165
-
Selective Feature Adapter for Dense Vision Transformers 3 Oct 2023 · 0 repositories · arXiv:2310.01843
-
Self-distilled Masked Attention guided masked image modeling with noise Regularized Teacher (SMART) for medical image analysis 2 Oct 2023 · 0 repositories · arXiv:2310.01209
-
Egocentric RGB+Depth Action Recognition in Industry-Like Settings 25 Sep 2023 · 1 repository · arXiv:2309.13962
-
OSNet & MNetO: Two Types of General Reconstruction Architectures for Linear Computed Tomography in Multi-Scenarios 21 Sep 2023 · 0 repositories · arXiv:2309.11858
-
Learning Dynamic MRI Reconstruction with Convolutional Network Assisted Reconstruction Swin Transformer 19 Sep 2023 · 0 repositories · arXiv:2309.10227
-
Neural network-based coronary dominance classification of RCA angiograms 13 Sep 2023 · 0 repositories · arXiv:2309.06958
-
MS-UNet-v2: Adaptive Denoising Method and Training Strategy for Medical Image Segmentation with Small Training Data 7 Sep 2023 · 0 repositories · arXiv:2309.03686
-
Improving diagnosis and prognosis of lung cancer using vision transformers: A scoping review 6 Sep 2023 · 0 repositories · arXiv:2309.02783
-
DAT++: Spatially Dynamic Vision Transformer with Deformable Attention 4 Sep 2023 · 1 repository · arXiv:2309.01430Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 6 pointer-only (licence)
-
Multi-dimension unified Swin Transformer for 3D Lesion Segmentation in Multiple Anatomical Locations 4 Sep 2023 · 0 repositories · arXiv:2309.01823
-
ACC-UNet: A Completely Convolutional UNet model for the 2020s 25 Aug 2023 · 1 repository · arXiv:2308.13680
-
SG-Former: Self-guided Transformer with Evolving Token Reallocation 23 Aug 2023 · 1 repository · arXiv:2308.12216Syntology official (archive's flag): 8 ran · 8 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Small Object Detection for Birds with Swin Transformer 22 Aug 2023 · 0 repositories
-
SwinFace: A Multi-task Transformer for Face Recognition, Expression Recognition, Age Estimation and Attribute Estimation 22 Aug 2023 · 1 repository · arXiv:2308.11509
-
SwinV2DNet: Pyramid and Self-Supervision Compounded Feature Learning for Remote Sensing Images Change Detection 22 Aug 2023 · 0 repositories · arXiv:2308.11159
-
LDCSF: Local depth convolution-based Swim framework for classifying multi-label histopathology images 21 Aug 2023 · 0 repositories · arXiv:2308.10446
-
SwinLSTM:Improving Spatiotemporal Prediction Accuracy using Swin Transformer and LSTM 19 Aug 2023 · 1 repository · arXiv:2308.09891Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 9 harvested samples)
-
SwinJSCC: Taming Swin Transformer for Deep Joint Source-Channel Coding 18 Aug 2023 · 2 repositories · arXiv:2308.09361
-
SST: A Simplified Swin Transformer-based Model for Taxi Destination Prediction based on Existing Trajectory 15 Aug 2023 · 0 repositories · arXiv:2308.07555
-
SCSC: Spatial Cross-scale Convolution Module to Strengthen both CNNs and Transformers 14 Aug 2023 · 0 repositories · arXiv:2308.07110