Methods › General › Regularization › Label Smoothing › Papers, page 110
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 110 of 144: papers 10,901 to 11,000 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Scaling Up Your Kernels to 31x31: Revisiting Large Kernel Design in CNNs 13 Mar 2022 · 8 repositories · arXiv:2203.06717Syntology official (archive's flag): 4 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Sparse Local Patch Transformer for Robust Face Alignment and Landmarks Inherent Relation Learning 13 Mar 2022 · 1 repository · arXiv:2203.06541
-
DFTR: Depth-supervised Fusion Transformer for Salient Object Detection 12 Mar 2022 · 0 repositories · arXiv:2203.06429
-
Joint CNN and Transformer Network via weakly supervised Learning for efficient crowd counting 12 Mar 2022 · 0 repositories · arXiv:2203.06388
-
One-stage Video Instance Segmentation: From Frame-in Frame-out to Clip-in Clip-out 12 Mar 2022 · 0 repositories · arXiv:2203.06421
-
Wasserstein Adversarial Transformer for Cloud Workload Prediction 12 Mar 2022 · 1 repository · arXiv:2203.06501
-
Block-Recurrent Transformers 11 Mar 2022 · 3 repositories · arXiv:2203.07852Syntology official: no sample here; runs from other or unrecorded repositories · 18 ran (of which 2 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 3 violated, 4 with no contract checked; 9 where Syntology's instrument failed) · 5 unverified (of 23 harvested samples) · 4 pointer-only (licence)
-
Font Shape-to-Impression Translation 11 Mar 2022 · 0 repositories · arXiv:2203.05808
-
PathSAGE: Spatial Graph Attention Neural Networks With Random Path Sampling 11 Mar 2022 · 0 repositories · arXiv:2203.05793
-
The Role of ImageNet Classes in Fréchet Inception Distance 11 Mar 2022 · 2 repositories · arXiv:2203.06026
-
Transformer-based Streaming ASR with Cumulative Attention 11 Mar 2022 · 0 repositories · arXiv:2203.05736
-
Visualizing and Understanding Patch Interactions in Vision Transformer 11 Mar 2022 · 0 repositories · arXiv:2203.05922
-
Look Backward and Forward: Self-Knowledge Distillation with Bidirectional Decoder for Neural Machine Translation 10 Mar 2022 · 0 repositories · arXiv:2203.05248
-
Parameter-Free Attentive Scoring for Speaker Verification 10 Mar 2022 · 1 repository · arXiv:2203.05642
-
StyleBabel: Artistic Style Tagging and Captioning 10 Mar 2022 · 0 repositories · arXiv:2203.05321
-
TrueType Transformer: Character and Font Style Recognition in Outline Format 10 Mar 2022 · 1 repository · arXiv:2203.05338
-
Anti-Oversmoothing in Deep Vision Transformers via the Fourier Domain Analysis: From Theory to Practice 9 Mar 2022 · 1 repository · arXiv:2203.05962Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Autonomous Mosquito Habitat Detection Using Satellite Imagery and Convolutional Neural Networks for Disease Risk Mapping 9 Mar 2022 · 1 repository · arXiv:2203.04463
-
Coarse-to-Fine Sparse Transformer for Hyperspectral Image Reconstruction 9 Mar 2022 · 1 repository · arXiv:2203.04845Syntology official (archive's flag): 13 ran · 13 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Multiscale Convolutional Transformer with Center Mask Pretraining for Hyperspectral Image Classification 9 Mar 2022 · 0 repositories · arXiv:2203.04771
-
PHTrans: Parallelly Aggregating Global and Local Representations for Medical Image Segmentation 9 Mar 2022 · 2 repositories · arXiv:2203.04568
-
Region-Aware Face Swapping 9 Mar 2022 · 0 repositories · arXiv:2203.04564
-
The evolution, evolvability and engineering of gene regulatory DNA 9 Mar 2022 · 1 repository
-
Uni4Eye: Unified 2D and 3D Self-supervised Pre-training via Masked Image Modeling Transformer for Ophthalmic Image Classification 9 Mar 2022 · 0 repositories · arXiv:2203.04614
-
CaSS: A Channel-aware Self-supervised Representation Learning Framework for Multivariate Time Series Classification 8 Mar 2022 · 0 repositories · arXiv:2203.04298
-
DuMLP-Pin: A Dual-MLP-dot-product Permutation-invariant Network for Set Feature Extraction 8 Mar 2022 · 1 repository · arXiv:2203.04007
-
Dynamic Group Transformer: A General Vision Transformer Backbone with Dynamic Group Attention 8 Mar 2022 · 0 repositories · arXiv:2203.03937
-
Graph Attention Transformer Network for Multi-Label Image Classification 8 Mar 2022 · 1 repository · arXiv:2203.04049
-
Joint rotational invariance and adversarial training of a dual-stream Transformer yields state of the art Brain-Score for Area V4 8 Mar 2022 · 1 repository · arXiv:2203.06649
-
Lane Detection with Versatile AtrousFormer and Local Semantic Guidance 8 Mar 2022 · 0 repositories · arXiv:2203.04067
-
Measuring the Mixing of Contextual Information in the Transformer 8 Mar 2022 · 2 repositories · arXiv:2203.04212Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Monocular Robot Navigation with Self-Supervised Pretrained Vision Transformers 7 Mar 2022 · 0 repositories · arXiv:2203.03682
-
SkillNet-NLU: A Sparsely Activated Model for General-Purpose Natural Language Understanding 7 Mar 2022 · 0 repositories · arXiv:2203.03312
-
Stepwise Feature Fusion: Local Guides Global 7 Mar 2022 · 1 repository · arXiv:2203.03635
-
Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer 7 Mar 2022 · 7 repositories · arXiv:2203.03466Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Conditional Bilingual Mutual Information Based Adaptive Training for Neural Machine Translation 6 Mar 2022 · 1 repository · arXiv:2203.02951
-
Exploring Dual-task Correlation for Pose Guided Person Image Generation 6 Mar 2022 · 1 repository · arXiv:2203.02910Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Focus on the Target's Vocabulary: Masked Label Smoothing for Machine Translation 6 Mar 2022 · 2 repositories · arXiv:2203.02889
-
Learnable Irrelevant Modality Dropout for Multimodal Action Recognition on Modality-Specific Annotated Videos 6 Mar 2022 · 0 repositories · arXiv:2203.03014
-
Multi-class Token Transformer for Weakly Supervised Semantic Segmentation 6 Mar 2022 · 1 repository · arXiv:2203.02891Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
PanFormer: a Transformer Based Model for Pan-sharpening 6 Mar 2022 · 1 repository · arXiv:2203.02916
-
Better Supervisory Signals by Observing Learning Paths 4 Mar 2022 · 1 repository · arXiv:2203.02485Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
ClarET: Pre-training a Correlation-Aware Context-To-Event Transformer for Event-Centric Generation and Classification 4 Mar 2022 · 1 repository · arXiv:2203.02225
-
DiT: Self-supervised Pre-training for Document Image Transformer 4 Mar 2022 · 4 repositories · arXiv:2203.02378Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
LiteTransformerSearch: Training-free Neural Architecture Search for Efficient Language Models 4 Mar 2022 · 1 repository · arXiv:2203.02094Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
UVCGAN: UNet Vision Transformer cycle-consistent GAN for unpaired image-to-image translation 4 Mar 2022 · 2 repositories · arXiv:2203.02557
-
LGT-Net: Indoor Panoramic Room Layout Estimation with Geometry-Aware Transformer Network 3 Mar 2022 · 1 repository · arXiv:2203.01824Syntology official (archive's flag): 17 ran · 29 ran (of which 5 constructed an object rather than computing a result; 20 with no instrument failure: 2 honoured, 1 violated, 17 with no contract checked; 9 where Syntology's instrument failed) · 12 unverified (of 41 harvested samples)
-
Multi-Tailed Vision Transformer for Efficient Inference 3 Mar 2022 · 0 repositories · arXiv:2203.01587
-
ViTransPAD: Video Transformer using convolution and self-attention for Face Presentation Attack Detection 3 Mar 2022 · 0 repositories · arXiv:2203.01562
-
3DCTN: 3D Convolution-Transformer Network for Point Cloud Classification 2 Mar 2022 · 1 repository · arXiv:2203.00828
-
Aggregated Pyramid Vision Transformer: Split-transform-merge Strategy for Image Recognition without Convolutions 2 Mar 2022 · 0 repositories · arXiv:2203.00960
-
Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic Segmentation 2 Mar 2022 · 1 repository · arXiv:2203.01452
-
Contextual Attention Network: Transformer Meets U-Net 2 Mar 2022 · 3 repositories · arXiv:2203.01932
-
D^2ETR: Decoder-Only DETR with Computationally Efficient Cross-Scale Attention 2 Mar 2022 · 0 repositories · arXiv:2203.00860
-
DN-DETR: Accelerate DETR Training by Introducing Query DeNoising 2 Mar 2022 · 17 repositories · arXiv:2203.01305Syntology official (archive's flag): 10 ran · 17 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 11 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 9 pointer-only (licence)
-
FastFold: Reducing AlphaFold Training Time from 11 Days to 67 Hours 2 Mar 2022 · 1 repository · arXiv:2203.00854
-
MTet: Multi-domain Translation for English-Vietnamese 2 Mar 2022 · 1 repository
-
Protecting Celebrities from DeepFake with Identity Consistency Transformer 2 Mar 2022 · 1 repository · arXiv:2203.01318Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Temporal Context Matters: Enhancing Single Image Prediction with Disease Progression Representations 2 Mar 2022 · 0 repositories · arXiv:2203.01933
-
X-Trans2Cap: Cross-Modal Knowledge Transfer using Transformer for 3D Dense Captioning 2 Mar 2022 · 1 repository · arXiv:2203.00843
-
DeepNet: Scaling Transformers to 1,000 Layers 1 Mar 2022 · 6 repositories · arXiv:2203.00555Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 3 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Fast-R2D2: A Pretrained Recursive Neural Network based on Pruned CKY for Grammar Induction and Text Representation 1 Mar 2022 · 2 repositories · arXiv:2203.00281Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
Temporal Perceiver: A General Architecture for Arbitrary Boundary Detection 1 Mar 2022 · 0 repositories · arXiv:2203.00307
-
Transformer Grammars: Augmenting Transformer Language Models with Syntactic Inductive Biases at Scale 1 Mar 2022 · 0 repositories · arXiv:2203.00633
-
Wearable Sensor-Based Human Activity Recognition with Transformer Model 1 Mar 2022 · 1 repository
-
A Data-scalable Transformer for Medical Image Segmentation: Architecture, Model Efficiency, and Benchmark 28 Feb 2022 · 2 repositories · arXiv:2203.00131Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
CTformer: Convolution-free Token2Token Dilated Vision Transformer for Low-dose CT Denoising 28 Feb 2022 · 2 repositories · arXiv:2202.13517
-
DropIT: Dropping Intermediate Tensors for Memory-Efficient DNN Training 28 Feb 2022 · 1 repository · arXiv:2202.13808
-
Filter-enhanced MLP is All You Need for Sequential Recommendation 28 Feb 2022 · 2 repositories · arXiv:2202.13556
-
LiLT: A Simple yet Effective Language-Independent Layout Transformer for Structured Document Understanding 28 Feb 2022 · 5 repositories · arXiv:2202.13669Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
LISA: Learning Interpretable Skill Abstractions from Language 28 Feb 2022 · 1 repository · arXiv:2203.00054
-
SUNet: Swin Transformer UNet for Image Denoising 28 Feb 2022 · 2 repositories · arXiv:2202.14009
-
DXM-TransFuse U-net: Dual Cross-Modal Transformer Fusion U-net for Automated Nerve Identification 27 Feb 2022 · 0 repositories · arXiv:2202.13304
-
Spatial-Temporal Attention Fusion Network for short-term passenger flow prediction on holidays in urban rail transit systems 27 Feb 2022 · 0 repositories · arXiv:2203.00007
-
Short-term passenger flow prediction for multi-traffic modes: A Transformer and residual network based multi-task learning method 27 Feb 2022 · 0 repositories · arXiv:2203.00422
-
An End-to-End Transformer Model for Crowd Localization 26 Feb 2022 · 1 repository · arXiv:2202.13065
-
Real-World Blind Super-Resolution via Feature Matching with Implicit High-Resolution Priors 26 Feb 2022 · 2 repositories · arXiv:2202.13142
-
A Hardware-Aware System for Accelerating Deep Neural Network Optimization 25 Feb 2022 · 0 repositories · arXiv:2202.12954
-
Self-Supervised and Interpretable Anomaly Detection using Network Transformers 25 Feb 2022 · 0 repositories · arXiv:2202.12997
-
Instantaneous Physiological Estimation using Video Transformers 24 Feb 2022 · 1 repository · arXiv:2202.12368
-
openFEAT: Improving Speaker Identification by Open-set Few-shot Embedding Adaptation with Transformer 24 Feb 2022 · 0 repositories · arXiv:2202.12349
-
Transformers in Medical Image Analysis: A Review 24 Feb 2022 · 0 repositories · arXiv:2202.12165
-
A Differential Attention Fusion Model Based on Transformer for Time Series Forecasting 23 Feb 2022 · 0 repositories · arXiv:2202.11402
-
FastRPB: a Scalable Relative Positional Encoding for Long Sequence Tasks 23 Feb 2022 · 1 repository · arXiv:2202.11364
-
Paying U-Attention to Textures: Multi-Stage Hourglass Vision Transformer for Universal Texture Synthesis 23 Feb 2022 · 0 repositories · arXiv:2202.11703
-
Refining the state-of-the-art in Machine Translation, optimizing NMT for the JA <-> EN language pair by leveraging personal domain expertise 23 Feb 2022 · 0 repositories · arXiv:2202.11669
-
Think Global, Act Local: Dual-scale Graph Transformer for Vision-and-Language Navigation 23 Feb 2022 · 1 repository · arXiv:2202.11742Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
A New Generation of Perspective API: Efficient Multilingual Character-level Transformers 22 Feb 2022 · 0 repositories · arXiv:2202.11176
-
GroupViT: Semantic Segmentation Emerges from Text Supervision 22 Feb 2022 · 6 repositories · arXiv:2202.11094
-
One-shot Scene Graph Generation 22 Feb 2022 · 1 repository · arXiv:2202.10824
-
Social Computational Design Method for Generating Product Shapes with GAN and Transformer Models 22 Feb 2022 · 0 repositories · arXiv:2202.10774
-
Socialformer: Social Network Inspired Long Document Modeling for Document Ranking 22 Feb 2022 · 1 repository · arXiv:2202.10870
-
Embarrassingly Simple Performance Prediction for Abductive Natural Language Inference 21 Feb 2022 · 1 repository · arXiv:2202.10408
-
Rethinking the Zigzag Flattening for Image Reading 21 Feb 2022 · 0 repositories · arXiv:2202.10240
-
S3T: Self-Supervised Pre-training with Swin Transformer for Music Classification 21 Feb 2022 · 1 repository · arXiv:2202.10139
-
ViTAEv2: Vision Transformer Advanced by Exploring Inductive Bias for Image Recognition and Beyond 21 Feb 2022 · 8 repositories · arXiv:2202.10108Syntology community repositories only · 22 ran (of which 8 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 6 where Syntology's instrument failed) · 3 unverified (of 25 harvested samples) · 4 pointer-only (licence)
-
ARM3D: Attention-based relation module for indoor 3D object detection 20 Feb 2022 · 1 repository · arXiv:2202.09715
-
Do Transformers know symbolic rules, and would we know if they did? 19 Feb 2022 · 0 repositories · arXiv:2203.00162
-
TransDreamer: Reinforcement Learning with Transformer World Models 19 Feb 2022 · 0 repositories · arXiv:2202.09481Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Survey of Vision-Language Pre-Trained Models 18 Feb 2022 · 0 repositories · arXiv:2202.10936