Methods › General › Regularization › Label Smoothing › Papers, page 44
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 44 of 144: papers 4,301 to 4,400 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Word Alignment as Preference for Machine Translation 15 May 2024 · 0 repositories · arXiv:2405.09223
-
A Comprehensive Survey of Large Language Models and Multimodal Large Language Models in Medicine 14 May 2024 · 0 repositories · arXiv:2405.08603
-
A Timely Survey on Vision Transformer for Deepfake Detection 14 May 2024 · 0 repositories · arXiv:2405.08463
-
Airport Delay Prediction with Temporal Fusion Transformers 14 May 2024 · 0 repositories · arXiv:2405.08293
-
Beyond Scaling Laws: Understanding Transformer Performance with Associative Memory 14 May 2024 · 0 repositories · arXiv:2405.08707
-
DGCformer: Deep Graph Clustering Transformer for Multivariate Time Series Forecasting 14 May 2024 · 0 repositories · arXiv:2405.08440
-
Harnessing the power of longitudinal medical imaging for eye disease prognosis using Transformer-based sequence modeling 14 May 2024 · 1 repository · arXiv:2405.08780
-
Improving Transformers with Dynamically Composable Multi-Head Attention 14 May 2024 · 2 repositories · arXiv:2405.08553Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
Reinformer: Max-Return Sequence Modeling for Offline RL 14 May 2024 · 1 repository · arXiv:2405.08740Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 4 harvested samples) · 4 pointer-only (licence)
-
Rethinking Scanning Strategies with Vision Mamba in Semantic Segmentation of Remote Sensing Imagery: An Experimental Study 14 May 2024 · 0 repositories · arXiv:2405.08493
-
RMT-BVQA: Recurrent Memory Transformer-based Blind Video Quality Assessment for Enhanced Video Content 14 May 2024 · 0 repositories · arXiv:2405.08621
-
TFWT: Tabular Feature Weighting with Transformer 14 May 2024 · 0 repositories · arXiv:2405.08403
-
WaterMamba: Visual State Space Model for Underwater Image Enhancement 14 May 2024 · 0 repositories · arXiv:2405.08419
-
Can Language Models Explain Their Own Classification Behavior? 13 May 2024 · 1 repository · arXiv:2405.07436
-
CDFormer:When Degradation Prediction Embraces Diffusion Model for Blind Image Super-Resolution 13 May 2024 · 1 repository · arXiv:2405.07648Syntology official (archive's flag): 10 ran · 11 ran (of which 7 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 17 harvested samples) · 1 pointer-only (licence)
-
Coding historical causes of death data with Large Language Models 13 May 2024 · 1 repository · arXiv:2405.07560
-
Fighter flight trajectory prediction based on spatio-temporal graphcial attention network 13 May 2024 · 0 repositories · arXiv:2405.08034
-
Ground-based image deconvolution with Swin Transformer UNet 13 May 2024 · 0 repositories · arXiv:2405.07842
-
Decision Mamba Architectures 13 May 2024 · 2 repositories · arXiv:2405.07943
-
HybridHash: Hybrid Convolutional and Self-Attention Deep Hashing for Image Retrieval 13 May 2024 · 1 repository · arXiv:2405.07524
-
MetaReflection: Learning Instructions for Language Agents using Past Reflections 13 May 2024 · 0 repositories · arXiv:2405.13009
-
Plot2Code: A Comprehensive Benchmark for Evaluating Multi-modal Large Language Models in Code Generation from Scientific Plots 13 May 2024 · 0 repositories · arXiv:2405.07990
-
Quantifying and Optimizing Global Faithfulness in Persona-driven Role-playing 13 May 2024 · 1 repository · arXiv:2405.07726Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
RESTAD: REconstruction and Similarity based Transformer for time series Anomaly Detection 13 May 2024 · 1 repository · arXiv:2405.07509
-
SambaNova SN40L: Scaling the AI Memory Wall with Dataflow and Composition of Experts 13 May 2024 · 0 repositories · arXiv:2405.07518
-
CaFA: Global Weather Forecasting with Factorized Attention on Sphere 12 May 2024 · 1 repository · arXiv:2405.07395Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 2 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
HGTDR: Advancing Drug Repurposing with Heterogeneous Graph Transformers 12 May 2024 · 0 repositories · arXiv:2405.08031
-
Humor Mechanics: Advancing Humor Generation with Multistep Reasoning 12 May 2024 · 1 repository · arXiv:2405.07280
-
L(u)PIN: LLM-based Political Ideology Nowcasting 12 May 2024 · 0 repositories · arXiv:2405.07320
-
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis 12 May 2024 · 1 repository · arXiv:2405.07248Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
MedConceptsQA: Open Source Medical Concepts QA Benchmark 12 May 2024 · 1 repository · arXiv:2405.07348Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AraSpell: A Deep Learning Approach for Arabic Spelling Correction 11 May 2024 · 1 repository · arXiv:2405.06981
-
Automating Thematic Analysis: How LLMs Analyse Controversial Topics 11 May 2024 · 0 repositories · arXiv:2405.06919
-
Identifying Key Terms in Prompts for Relevance Evaluation with GPT Models 11 May 2024 · 0 repositories · arXiv:2405.06931
-
Length-Aware Multi-Kernel Transformer for Long Document Classification 11 May 2024 · 1 repository · arXiv:2405.07052
-
Multi-agent Traffic Prediction via Denoised Endpoint Distribution 11 May 2024 · 0 repositories · arXiv:2405.07041
-
QMViT: A Mushroom is worth 16x16 Words 11 May 2024 · 0 repositories · arXiv:2407.04708
-
Replication Study and Benchmarking of Real-Time Object Detection Models 11 May 2024 · 1 repository · arXiv:2405.06911
-
RoTHP: Rotary Position Embedding-based Transformer Hawkes Process 11 May 2024 · 0 repositories · arXiv:2405.06985
-
Super-Resolving Blurry Images with Events 11 May 2024 · 0 repositories · arXiv:2405.06918
-
TacoERE: Cluster-aware Compression for Event Relation Extraction 11 May 2024 · 0 repositories · arXiv:2405.06890
-
A Lightweight Sparse Focus Transformer for Remote Sensing Image Change Captioning 10 May 2024 · 1 repository · arXiv:2405.06598
-
CANAL -- Cyber Activity News Alerting Language Model: Empirical Approach vs. Expensive LLM 10 May 2024 · 0 repositories · arXiv:2405.06772
-
Dual-Task Vision Transformer for Rapid and Accurate Intracerebral Hemorrhage CT Image Classification 10 May 2024 · 0 repositories · arXiv:2405.06814
-
Large Language Model in Financial Regulatory Interpretation 10 May 2024 · 0 repositories · arXiv:2405.06808
-
Mesh Denoising Transformer 10 May 2024 · 0 repositories · arXiv:2405.06536
-
Multimodal LLMs Struggle with Basic Visual Network Analysis: a VNA Benchmark 10 May 2024 · 1 repository · arXiv:2405.06634
-
A Mixture of Experts Approach to 3D Human Motion Prediction 9 May 2024 · 1 repository · arXiv:2405.06088
-
Artificial Intelligence as the New Hacker: Developing Agents for Offensive Security 9 May 2024 · 0 repositories · arXiv:2406.07561
-
Bidirectional Progressive Transformer for Interaction Intention Anticipation 9 May 2024 · 0 repositories · arXiv:2405.05552
-
Can large language models understand uncommon meanings of common words? 9 May 2024 · 0 repositories · arXiv:2405.05741
-
Digital Diagnostics: The Potential Of Large Language Models In Recognizing Symptoms Of Common Illnesses 9 May 2024 · 0 repositories · arXiv:2405.06712
-
Ditto: Quantization-aware Secure Inference of Transformers upon MPC 9 May 2024 · 1 repository · arXiv:2405.05525
-
Enhancing Suicide Risk Detection on Social Media through Semi-Supervised Deep Label Smoothing 9 May 2024 · 0 repositories · arXiv:2405.05795
-
HMT: Hierarchical Memory Transformer for Long Context Language Processing 9 May 2024 · 1 repository · arXiv:2405.06067Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
Letter to the Editor: What are the legal and ethical considerations of submitting radiology reports to ChatGPT? 9 May 2024 · 0 repositories · arXiv:2405.05647
-
Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers 9 May 2024 · 2 repositories · arXiv:2405.05945
-
People cannot distinguish GPT-4 from a human in a Turing test 9 May 2024 · 0 repositories · arXiv:2405.08007
-
Self-Supervised Learning of Time Series Representation via Diffusion Process and Imputation-Interpolation-Forecasting Mask 9 May 2024 · 2 repositories · arXiv:2405.05959
-
Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media 9 May 2024 · 0 repositories · arXiv:2405.05760
-
Smurfs: Leveraging Multiple Proficiency Agents with Context-Efficiency for Tool Planning 9 May 2024 · 1 repository · arXiv:2405.05955Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Vision-Language Modeling with Regularized Spatial Transformer Networks for All Weather Crosswind Landing of Aircraft 9 May 2024 · 0 repositories · arXiv:2405.05574
-
VM-DDPM: Vision Mamba Diffusion for Medical Image Synthesis 9 May 2024 · 0 repositories · arXiv:2405.05667
-
ACORN: Aspect-wise Commonsense Reasoning Explanation Evaluation 8 May 2024 · 1 repository · arXiv:2405.04818
-
Few-Shot Class Incremental Learning via Robust Transformer Approach 8 May 2024 · 1 repository · arXiv:2405.05984
-
HMANet: Hybrid Multi-Axis Aggregation Network for Image Super-Resolution 8 May 2024 · 1 repository · arXiv:2405.05001
-
Open Source Language Models Can Provide Feedback: Evaluating LLMs' Ability to Help Students Using GPT-4-As-A-Judge 8 May 2024 · 1 repository · arXiv:2405.05253
-
Seeds of Stereotypes: A Large-Scale Textual Analysis of Race and Gender Associations with Diseases in Online Sources 8 May 2024 · 0 repositories · arXiv:2405.05049
-
You Only Cache Once: Decoder-Decoder Architectures for Language Models 8 May 2024 · 1 repository · arXiv:2405.05254
-
D-TrAttUnet: Toward Hybrid CNN-Transformer Architecture for Generic and Subtle Segmentation in Medical Images 7 May 2024 · 1 repository · arXiv:2405.04169
-
Enhancing the Efficiency and Accuracy of Underlying Asset Reviews in Structured Finance: The Application of Multi-agent Framework 7 May 2024 · 1 repository · arXiv:2405.04294
-
Folded Context Condensation in Path Integral Formalism for Infinite Context Transformers 7 May 2024 · 0 repositories · arXiv:2405.04620
-
HAFFormer: A Hierarchical Attention-Free Framework for Alzheimer's Disease Detection From Spontaneous Speech 7 May 2024 · 0 repositories · arXiv:2405.03952
-
Learning Linear Block Error Correction Codes 7 May 2024 · 1 repository · arXiv:2405.04050
-
Masked Graph Transformer for Large-Scale Recommendation 7 May 2024 · 0 repositories · arXiv:2405.04028
-
NaturalCodeBench: Examining Coding Performance Mismatch on HumanEval and Natural User Prompts 7 May 2024 · 1 repository · arXiv:2405.04520
-
POV Learning: Individual Alignment of Multimodal Models using Human Perception 7 May 2024 · 0 repositories · arXiv:2405.04443
-
Predictive Modeling with Temporal Graphical Representation on Electronic Health Records 7 May 2024 · 1 repository · arXiv:2405.03943Syntology official (archive's flag): 10 ran · 10 ran (of which 3 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
S3Former: Self-supervised High-resolution Transformer for Solar PV Profiling 7 May 2024 · 0 repositories · arXiv:2405.04489
-
Structured Click Control in Transformer-based Interactive Segmentation 7 May 2024 · 1 repository · arXiv:2405.04009Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Vision Mamba: A Comprehensive Survey and Taxonomy 7 May 2024 · 1 repository · arXiv:2405.04404
-
xLSTM: Extended Long Short-Term Memory 7 May 2024 · 5 repositories · arXiv:2405.04517Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
AlphaMath Almost Zero: Process Supervision without Process 6 May 2024 · 1 repository · arXiv:2405.03553
-
Anchored Answers: Unravelling Positional Bias in GPT-2's Multiple-Choice Questions 6 May 2024 · 1 repository · arXiv:2405.03205
-
Class-relevant Patch Embedding Selection for Few-Shot Image Classification 6 May 2024 · 0 repositories · arXiv:2405.03722
-
CRA5: Extreme Compression of ERA5 for Portable Global Climate and Weather Research via an Efficient Variational Transformer 6 May 2024 · 1 repository · arXiv:2405.03376
-
Dual Relation Mining Network for Zero-Shot Learning 6 May 2024 · 0 repositories · arXiv:2405.03613
-
Enhancing DETRs Variants through Improved Content Query and Similar Query Aggregation 6 May 2024 · 0 repositories · arXiv:2405.03318
-
Exploring the Frontiers of Softmax: Provable Optimization, Applications in Diffusion Model, and Beyond 6 May 2024 · 0 repositories · arXiv:2405.03251
-
GREEN: Generative Radiology Report Evaluation and Error Notation 6 May 2024 · 0 repositories · arXiv:2405.03595
-
Intra-task Mutual Attention based Vision Transformer for Few-Shot Learning 6 May 2024 · 0 repositories · arXiv:2405.03109
-
Investigating Personalized Driving Behaviors in Dilemma Zones: Analysis and Prediction of Stop-or-Go Decisions 6 May 2024 · 0 repositories · arXiv:2405.03873
-
MAmmoTH2: Scaling Instructions from the Web 6 May 2024 · 0 repositories · arXiv:2405.03548
-
Modality Prompts for Arbitrary Modality Salient Object Detection 6 May 2024 · 0 repositories · arXiv:2405.03351
-
ReCycle: Fast and Efficient Long Time Series Forecasting with Residual Cyclic Transformers 6 May 2024 · 1 repository · arXiv:2405.03429
-
Salient Object Detection From Arbitrary Modalities 6 May 2024 · 1 repository · arXiv:2405.03352
-
SocialFormer: Social Interaction Modeling with Edge-enhanced Heterogeneous Graph Transformers for Trajectory Prediction 6 May 2024 · 0 repositories · arXiv:2405.03809
-
Transformer-based RGB-T Tracking with Channel and Spatial Feature Fusion 6 May 2024 · 1 repository · arXiv:2405.03177
-
Transformer models as an efficient replacement for statistical test suites to evaluate the quality of random numbers 6 May 2024 · 0 repositories · arXiv:2405.03904
-
Can Large Language Models Make the Grade? An Empirical Study Evaluating LLMs Ability to Mark Short Answer Questions in K-12 Education 5 May 2024 · 0 repositories · arXiv:2405.02985