Methods › General › Regularization › Label Smoothing › Papers, page 121
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 121 of 144: papers 12,001 to 12,100 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
H-Transformer-1D: Fast One-Dimensional Hierarchical Attention for Sequences 25 Jul 2021 · 2 repositories · arXiv:2107.11906Syntology 9 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Class-Incremental Domain Adaptation with Smoothing and Calibration for Surgical Report Generation 23 Jul 2021 · 1 repository · arXiv:2107.11091
-
Similarity Based Label Smoothing For Dialogue Generation 23 Jul 2021 · 0 repositories · arXiv:2107.11481
-
Confidence-Aware Scheduled Sampling for Neural Machine Translation 22 Jul 2021 · 1 repository · arXiv:2107.10427
-
EAN: Event Adaptive Network for Enhanced Action Recognition 22 Jul 2021 · 1 repository · arXiv:2107.10771
-
FNetAR: Mixing Tokens with Autoregressive Fourier Transforms 22 Jul 2021 · 1 repository · arXiv:2107.10932
-
Query2Label: A Simple Transformer Way to Multi-Label Classification 22 Jul 2021 · 3 repositories · arXiv:2107.10834Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Tsformer: Time series Transformer for tourism demand forecasting 22 Jul 2021 · 0 repositories · arXiv:2107.10977
-
Audio Captioning Transformer 21 Jul 2021 · 1 repository · arXiv:2107.09817
-
CycleMLP: A MLP-like Architecture for Dense Prediction 21 Jul 2021 · 8 repositories · arXiv:2107.10224Syntology official (archive's flag): 2 ran · 10 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples) · 2 pointer-only (licence)
-
Multi-Stream Transformers 21 Jul 2021 · 1 repository · arXiv:2107.10342
-
A Multi-scale Graph Network with Multi-head Attention for Histopathology Image Diagnosis 20 Jul 2021 · 0 repositories
-
CREW: Computation Reuse and Efficient Weight Storage for Hardware-accelerated MLPs and RNNs 20 Jul 2021 · 0 repositories · arXiv:2107.09408
-
Generative Video Transformer: Can Objects be the Words? 20 Jul 2021 · 0 repositories · arXiv:2107.09240
-
Learning ULMFiT and Self-Distillation with Calibration for Medical Dialogue System 20 Jul 2021 · 0 repositories · arXiv:2107.09625
-
Weakly Supervised Global-Local Feature Learning for Cervical Cytology Image Analysis 20 Jul 2021 · 0 repositories
-
Image Fusion Transformer 19 Jul 2021 · 1 repository · arXiv:2107.09011
-
Learning Attributed Graph Representations with Communicative Message Passing Transformer 19 Jul 2021 · 1 repository · arXiv:2107.08773
-
LeViT-UNet: Make Faster Encoders with Transformer for Medical Image Segmentation 19 Jul 2021 · 2 repositories · arXiv:2107.08623Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Long-term series forecasting with Query Selector -- efficient model of sparse attention 19 Jul 2021 · 2 repositories · arXiv:2107.08687
-
Residual Tree Aggregation of Layers for Neural Machine Translation 19 Jul 2021 · 0 repositories · arXiv:2107.14590
-
Sequence-to-Sequence Piano Transcription with Transformers 19 Jul 2021 · 2 repositories · arXiv:2107.09142Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Discriminative Semantic Ranker for Question Retrieval 18 Jul 2021 · 0 repositories · arXiv:2107.08345
-
TFix: Learning to Fix Coding Errors with a Text-to-Text Transformer 18 Jul 2021 · 1 repository
-
Dynamic Transformer for Efficient Machine Translation on Embedded Devices 17 Jul 2021 · 0 repositories · arXiv:2107.08199
-
Neural Search: Learning Query and Product Representations in Fashion E-commerce 17 Jul 2021 · 0 repositories · arXiv:2107.08291
-
Is attention to bounding boxes all you need for pedestrian action prediction? 16 Jul 2021 · 0 repositories · arXiv:2107.08031
-
Learning Sparse Interaction Graphs of Partially Detected Pedestrians for Trajectory Prediction 15 Jul 2021 · 1 repository · arXiv:2107.07056
-
STAR: Sparse Transformer-based Action Recognition 15 Jul 2021 · 1 repository · arXiv:2107.07089
-
Transformer-based Machine Learning for Fast SAT Solvers and Logic Synthesis 15 Jul 2021 · 1 repository · arXiv:2107.07116
-
A Note on Learning Rare Events in Molecular Dynamics using LSTM and Transformer 14 Jul 2021 · 1 repository · arXiv:2107.06573
-
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines 14 Jul 2021 · 1 repository · arXiv:2107.06925
-
Indonesia's Fake News Detection using Transformer Network 14 Jul 2021 · 1 repository · arXiv:2107.06796
-
Real Time Pear Fruit Detection and Counting Using YOLOv4 Models and Deep SORT 14 Jul 2021 · 3 repositories
-
Serialized Multi-Layer Multi-Head Attention for Neural Speaker Embedding 14 Jul 2021 · 0 repositories · arXiv:2107.06493
-
HAT: Hierarchical Aggregation Transformers for Person Re-identification 13 Jul 2021 · 1 repository · arXiv:2107.05946
-
Real-Time Pothole Detection Using Deep Learning 13 Jul 2021 · 0 repositories · arXiv:2107.06356
-
The Piano Inpainting Application 13 Jul 2021 · 2 repositories · arXiv:2107.05944
-
MECT: Multi-Metadata Embedding based Cross-Transformer for Chinese Named Entity Recognition 12 Jul 2021 · 1 repository · arXiv:2107.05418
-
BERT-like Pre-training for Symbolic Piano Music Classification Tasks 12 Jul 2021 · 1 repository · arXiv:2107.05223
-
MOOCRep: A Unified Pre-trained Embedding of MOOC Entities 12 Jul 2021 · 1 repository · arXiv:2107.05154
-
TransAttUnet: Multi-level Attention-guided U-Net with Transformer for Medical Image Segmentation 12 Jul 2021 · 1 repository · arXiv:2107.05274
-
Visual Transformer with Statistical Test for COVID-19 Classification 12 Jul 2021 · 0 repositories · arXiv:2107.05334
-
Transformers with multi-modal features and post-fusion context for e-commerce session-based recommendation 11 Jul 2021 · 0 repositories · arXiv:2107.05124
-
Consensual Collaborative Training And Knowledge Distillation Based Facial Expression Recognition Under Noisy Annotations 10 Jul 2021 · 3 repositories · arXiv:2107.04746
-
Few-Shot Domain Adaptation with Polymorphic Transformers 10 Jul 2021 · 1 repository · arXiv:2107.04805
-
Local-to-Global Self-Attention in Vision Transformers 10 Jul 2021 · 0 repositories · arXiv:2107.04735
-
Can Deep Neural Networks Predict Data Correlations from Column Names? 9 Jul 2021 · 1 repository · arXiv:2107.04553
-
Affect Expression Behaviour Analysis in the Wild using Consensual Collaborative Training 8 Jul 2021 · 1 repository · arXiv:2107.05736
-
Calliope -- A Polyphonic Music Transformer 8 Jul 2021 · 0 repositories · arXiv:2107.05546
-
Learning to Delegate for Large-scale Vehicle Routing 8 Jul 2021 · 1 repository · arXiv:2107.04139Syntology official: harvested, nothing ran · 0 ran · 6 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
E-PixelHop: An Enhanced PixelHop Method for Object Classification 7 Jul 2021 · 0 repositories · arXiv:2107.02966
-
Efficient Transformer for Direct Speech Translation 7 Jul 2021 · 0 repositories · arXiv:2107.03069
-
Learning Vision Transformer with Squeeze and Excitation for Facial Expression Recognition 7 Jul 2021 · 0 repositories · arXiv:2107.03107
-
MACCIF-TDNN: Multi aspect aggregation of channel and context interdependence features in TDNN-based speaker verification 7 Jul 2021 · 0 repositories · arXiv:2107.03104
-
Scopeformer: n-CNN-ViT Hybrid Model for Intracranial Hemorrhage Classification 7 Jul 2021 · 0 repositories · arXiv:2107.04575
-
Trans4Trans: Efficient Transformer for Transparent Object Segmentation to Help Visually Impaired People Navigate in the Real World 7 Jul 2021 · 1 repository · arXiv:2107.03172
-
Transformer Network for Significant Stenosis Detection in CCTA of Coronary Arteries 7 Jul 2021 · 1 repository · arXiv:2107.03035
-
Automatic size and pose homogenization with spatial transformer network to improve and accelerate pediatric segmentation 6 Jul 2021 · 0 repositories · arXiv:2107.02655
-
COVID-19 Pneumonia Severity Prediction using Hybrid Convolution-Attention Neural Architectures 6 Jul 2021 · 0 repositories · arXiv:2107.02672
-
Detecting Hypo-plastic Left Heart Syndrome in Fetal Ultrasound via Disease-specific Atlas Maps 6 Jul 2021 · 0 repositories · arXiv:2107.02643
-
Feature Fusion Vision Transformer for Fine-Grained Visual Categorization 6 Jul 2021 · 1 repository · arXiv:2107.02341
-
Point Cloud Registration using Representative Overlapping Points 6 Jul 2021 · 1 repository · arXiv:2107.02583Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples)
-
Long-Short Transformer: Efficient Transformers for Language and Vision 5 Jul 2021 · 3 repositories · arXiv:2107.02192Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 2 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 2 pointer-only (licence)
-
Test-Time Personalization with a Transformer for Human Pose Estimation 5 Jul 2021 · 0 repositories · arXiv:2107.02133
-
Vision Xformers: Efficient Attention for Image Classification 5 Jul 2021 · 2 repositories · arXiv:2107.02239
-
What Makes for Hierarchical Vision Transformer? 5 Jul 2021 · 0 repositories · arXiv:2107.02174
-
End-to-end Neural Coreference Resolution Revisited: A Simple yet Effective Baseline 4 Jul 2021 · 0 repositories · arXiv:2107.01700
-
Introducing Self-Attention to Target Attentive Graph Neural Networks 4 Jul 2021 · 1 repository · arXiv:2107.01516
-
Can Transformers Jump Around Right in Natural Language? Assessing Performance Transfer from SCAN 3 Jul 2021 · 0 repositories · arXiv:2107.01366
-
Supervised Off-Policy Ranking 3 Jul 2021 · 1 repository · arXiv:2107.01360Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Case Relation Transformer: A Crossmodal Language Generation Model for Fetching Instructions 2 Jul 2021 · 0 repositories · arXiv:2107.00789
-
Cross-view Geo-localization with Evolving Transformer 2 Jul 2021 · 0 repositories · arXiv:2107.00842
-
Online Metro Origin-Destination Prediction via Heterogeneous Information Aggregation 2 Jul 2021 · 1 repository · arXiv:2107.00946
-
R2D2: Recursive Transformer based on Differentiable Tree for Interpretable Hierarchical Language Modeling 2 Jul 2021 · 1 repository · arXiv:2107.00967
-
Relaxed Attention: A Simple Method to Boost Performance of End-to-End Automatic Speech Recognition 2 Jul 2021 · 1 repository · arXiv:2107.01275
-
Solving Machine Learning Problems 2 Jul 2021 · 1 repository · arXiv:2107.01238
-
Target-dependent UNITER: A Transformer-Based Multimodal Language Comprehension Model for Domestic Service Robots 2 Jul 2021 · 0 repositories · arXiv:2107.00811
-
Transformer-F: A Transformer network with effective methods for learning universal sentence representation 2 Jul 2021 · 0 repositories · arXiv:2107.00653
-
UTNet: A Hybrid Transformer Architecture for Medical Image Segmentation 2 Jul 2021 · 1 repository · arXiv:2107.00781
-
Visual Relationship Forecasting in Videos 2 Jul 2021 · 0 repositories · arXiv:2107.01181
-
Action Transformer: A Self-Attention Model for Short-Time Pose-Based Human Action Recognition 1 Jul 2021 · 4 repositories · arXiv:2107.00606
-
Cross-Lingual Transfer Learning for Statistical Type Inference 1 Jul 2021 · 0 repositories · arXiv:2107.00157
-
CSWin Transformer: A General Vision Transformer Backbone with Cross-Shaped Windows 1 Jul 2021 · 7 repositories · arXiv:2107.00652Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Focal Self-attention for Local-Global Interactions in Vision Transformers 1 Jul 2021 · 3 repositories · arXiv:2107.00641Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
GlyphCRM: Bidirectional Encoder Representation for Chinese Character with its Glyph 1 Jul 2021 · 0 repositories · arXiv:2107.00395
-
Multimodal Graph-based Transformer Framework for Biomedical Relation Extraction 1 Jul 2021 · 1 repository · arXiv:2107.00596
-
A Generative Model for Raw Audio Using Transformer Architectures 30 Jun 2021 · 0 repositories · arXiv:2106.16036
-
Dual Aspect Self-Attention based on Transformer for Remaining Useful Life Prediction 30 Jun 2021 · 1 repository · arXiv:2106.15842
-
An Efficient Cervical Whole Slide Image Analysis Framework Based on Multi-scale Semantic and Location Deep Features 29 Jun 2021 · 1 repository · arXiv:2106.15113
-
FastPitchFormant: Source-filter based Decomposed Modeling for Speech Synthesis 29 Jun 2021 · 1 repository · arXiv:2106.15123
-
GeoT: A Geometry-aware Transformer for Reliable Molecular Property Prediction and Chemically Interpretable Representation Learning 29 Jun 2021 · 1 repository · arXiv:2106.15516
-
Hierarchical Context-Aware Transformers for Non-Autoregressive Text to Speech 29 Jun 2021 · 0 repositories · arXiv:2106.15144
-
Looking Outside the Window: Wide-Context Transformer for the Semantic Segmentation of High-Resolution Remote Sensing Images 29 Jun 2021 · 1 repository · arXiv:2106.15754
-
Multi-Exit Vision Transformer for Dynamic Inference 29 Jun 2021 · 0 repositories · arXiv:2106.15183
-
Digging Errors in NMT: Evaluating and Understanding Model Errors from Partial Hypothesis Space 29 Jun 2021 · 0 repositories · arXiv:2106.15217
-
Unified Questioner Transformer for Descriptive Question Generation in Goal-Oriented Visual Dialogue 29 Jun 2021 · 1 repository · arXiv:2106.15550
-
A Knowledge-Grounded Dialog System Based on Pre-Trained Language Models 28 Jun 2021 · 0 repositories · arXiv:2106.14444
-
Achieving Real-Time Object Detection on MobileDevices with Neural Pruning Search 28 Jun 2021 · 0 repositories · arXiv:2106.14943
-
Complexity-based partitioning of CSFI problem instances with Transformers 28 Jun 2021 · 0 repositories · arXiv:2106.14481