Methods › General › Regularization › Label Smoothing › Papers, page 3
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 3 of 144: papers 201 to 300 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Attention-Aided MMSE for OFDM Channel Estimation: Learning Linear Filters with Attention 31 May 2025 · 0 repositories · arXiv:2506.00452
-
Channel-Imposed Fusion: A Simple yet Effective Method for Medical Time Series Classification 31 May 2025 · 0 repositories · arXiv:2506.00337
-
Evaluating Robot Policies in a World Model 31 May 2025 · 0 repositories · arXiv:2506.00613
-
Machine vs Machine: Using AI to Tackle Generative AI Threats in Assessment 31 May 2025 · 0 repositories · arXiv:2506.02046
-
Translate With Care: Addressing Gender Bias, Neutrality, and Reasoning in Large Language Model Translations 31 May 2025 · 1 repository · arXiv:2506.00748
-
Cross-Attention Speculative Decoding 30 May 2025 · 0 repositories · arXiv:2505.24544
-
D2AF: A Dual-Driven Annotation and Filtering Framework for Visual Grounding 30 May 2025 · 0 repositories · arXiv:2505.24372
-
Leveraging Intermediate Features of Vision Transformer for Face Anti-Spoofing 30 May 2025 · 0 repositories · arXiv:2505.24402
-
Mamba Knockout for Unraveling Factual Information Flow 30 May 2025 · 1 repository · arXiv:2505.24244Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer 30 May 2025 · 1 repository · arXiv:2505.24378
-
Mixture-of-Experts for Personalized and Semantic-Aware Next Location Prediction 30 May 2025 · 0 repositories · arXiv:2505.24597
-
PCIE_Pose Solution for EgoExo4D Pose and Proficiency Estimation Challenge 30 May 2025 · 0 repositories · arXiv:2505.24411
-
PersianMedQA: Language-Centric Evaluation of LLMs in the Persian Medical Domain 30 May 2025 · 0 repositories · arXiv:2506.00250
-
SPPSFormer: High-quality Superpoint-based Transformer for Roof Plane Instance Segmentation from Point Clouds 30 May 2025 · 0 repositories · arXiv:2505.24475
-
Adversarial Semantic and Label Perturbation Attack for Pedestrian Attribute Recognition 29 May 2025 · 2 repositories · arXiv:2505.23313
-
ATLAS: Learning to Optimally Memorize the Context at Test Time 29 May 2025 · 0 repositories · arXiv:2505.23735
-
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time 29 May 2025 · 0 repositories · arXiv:2505.23729
-
CF-DETR: Coarse-to-Fine Transformer for Real-Time Object Detection 29 May 2025 · 0 repositories · arXiv:2505.23317
-
DA-VPT: Semantic-Guided Visual Prompt Tuning for Vision Transformers 29 May 2025 · 1 repository · arXiv:2505.23694Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Differential Gated Self-Attention 29 May 2025 · 0 repositories · arXiv:2505.24054
-
Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach 29 May 2025 · 0 repositories · arXiv:2505.23953
-
Equivariant Spherical Transformer for Efficient Molecular Modeling 29 May 2025 · 0 repositories · arXiv:2505.23086
-
From Images to Signals: Are Large Vision Models Useful for Time Series Analysis? 29 May 2025 · 0 repositories · arXiv:2505.24030
-
HyperPointFormer: Multimodal Fusion in 3D Space with Dual-Branch Cross-Attention Transformers 29 May 2025 · 1 repository · arXiv:2505.23206
-
Learning to Regulate: A New Event-Level Dataset of Capital Control Measures 29 May 2025 · 0 repositories · arXiv:2505.23025
-
Patient Domain Supervised Contrastive Learning for Lung Sound Classification Using Mobile Phone 29 May 2025 · 0 repositories · arXiv:2505.23132
-
Probing Association Biases in LLM Moderation Over-Sensitivity 29 May 2025 · 0 repositories · arXiv:2505.23914
-
Table-R1: Inference-Time Scaling for Table Reasoning 29 May 2025 · 1 repository · arXiv:2505.23621
-
The Warmup Dilemma: How Learning Rate Strategies Impact Speech-to-Text Model Convergence 29 May 2025 · 1 repository · arXiv:2505.23420
-
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos 29 May 2025 · 1 repository · arXiv:2505.23693
-
Attention-Enhanced Prompt Decision Transformers for UAV-Assisted Communications with AoI 28 May 2025 · 0 repositories · arXiv:2505.22170
-
HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer 28 May 2025 · 2 repositories · arXiv:2505.22705Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
MultiFormer: A Multi-Person Pose Estimation System Based on CSI and Attention Mechanism 28 May 2025 · 0 repositories · arXiv:2505.22555
-
Triple Attention Transformer Architecture for Time-Dependent Concrete Creep Prediction 28 May 2025 · 0 repositories · arXiv:2506.04243
-
Update Your Transformer to the Latest Release: Re-Basin of Task Vectors 28 May 2025 · 1 repository · arXiv:2505.22697Syntology official (archive's flag): 1 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 3 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
A domain adaptation neural network for digital twin-supported fault diagnosis 27 May 2025 · 1 repository · arXiv:2505.21046
-
AgriFM: A Multi-source Temporal Remote Sensing Foundation Model for Crop Mapping 27 May 2025 · 1 repository · arXiv:2505.21357
-
Beyond 1D: Vision Transformers and Multichannel Signal Images for PPG-to-ECG Reconstruction 27 May 2025 · 0 repositories · arXiv:2505.21767
-
Continuous-Time Attention: PDE-Guided Mechanisms for Long-Sequence Transformers 27 May 2025 · 0 repositories · arXiv:2505.20666
-
HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling 27 May 2025 · 0 repositories · arXiv:2505.20836
-
HTMNet: A Hybrid Network with Transformer-Mamba Bottleneck Multimodal Fusion for Transparent and Reflective Objects Depth Completion 27 May 2025 · 0 repositories · arXiv:2505.20904
-
Minute-Long Videos with Dual Parallelisms 27 May 2025 · 1 repository · arXiv:2505.21070
-
MoPFormer: Motion-Primitive Transformer for Wearable-Sensor Activity Recognition 27 May 2025 · 0 repositories · arXiv:2505.20744
-
Pause Tokens Strictly Increase the Expressivity of Constant-Depth Transformers 27 May 2025 · 0 repositories · arXiv:2505.21024
-
Privacy-Preserving Chest X-ray Report Generation via Multimodal Federated Learning with ViT and GPT-2 27 May 2025 · 0 repositories · arXiv:2505.21715
-
SOSBENCH: Benchmarking Safety Alignment on Scientific Knowledge 27 May 2025 · 0 repositories · arXiv:2505.21605
-
Absolute Coordinates Make Motion Generation Easy 26 May 2025 · 0 repositories · arXiv:2505.19377
-
Aggregated Structural Representation with Large Language Models for Human-Centric Layout Generation 26 May 2025 · 0 repositories · arXiv:2505.19554
-
AMQA: An Adversarial Dataset for Benchmarking Bias of LLMs in Medicine and Healthcare 26 May 2025 · 1 repository · arXiv:2505.19562
-
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages 26 May 2025 · 0 repositories · arXiv:2505.19851
-
CardioPatternFormer: Pattern-Guided Attention for Interpretable ECG Classification with Transformer Architecture 26 May 2025 · 0 repositories · arXiv:2505.20481
-
Dependency Parsing is More Parameter-Efficient with Normalization 26 May 2025 · 0 repositories · arXiv:2505.20215
-
Electrolyzers-HSI: Close-Range Multi-Scene Hyperspectral Imaging Benchmark Dataset 26 May 2025 · 0 repositories · arXiv:2505.20507
-
GoLF-NRT: Integrating Global Context and Local Geometry for Few-Shot View Synthesis 26 May 2025 · 1 repository · arXiv:2505.19813Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 9 unverified (of 24 harvested samples) · 24 pointer-only (licence)
-
Grokking ExPLAIND: Unifying Model, Data, and Training Attribution to Study Model Behavior 26 May 2025 · 1 repository · arXiv:2505.20076Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots 26 May 2025 · 1 repository · arXiv:2505.20288
-
Large Language Models' Reasoning Stalls: An Investigation into the Capabilities of Frontier Models 26 May 2025 · 0 repositories · arXiv:2505.19676
-
LeCoDe: A Benchmark Dataset for Interactive Legal Consultation Dialogue Evaluation 26 May 2025 · 0 repositories · arXiv:2505.19667
-
LlamaSeg: Image Segmentation via Autoregressive Mask Generation 26 May 2025 · 0 repositories · arXiv:2505.19422
-
Minimalist Softmax Attention Provably Learns Constrained Boolean Functions 26 May 2025 · 0 repositories · arXiv:2505.19531
-
Multi-modal brain encoding models for multi-modal stimuli 26 May 2025 · 1 repository · arXiv:2505.20027Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning 26 May 2025 · 0 repositories · arXiv:2505.19938
-
One Surrogate to Fool Them All: Universal, Transferable, and Targeted Adversarial Attacks with CLIP 26 May 2025 · 1 repository · arXiv:2505.19840
-
REARANK: Reasoning Re-ranking Agent via Reinforcement Learning 26 May 2025 · 1 repository · arXiv:2505.20046Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples) · 4 pointer-only (licence)
-
Structured Initialization for Vision Transformers 26 May 2025 · 0 repositories · arXiv:2505.19985
-
Synthetic Time Series Forecasting with Transformer Architectures: Extensive Simulation Benchmarks 26 May 2025 · 1 repository · arXiv:2505.20048
-
The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants 26 May 2025 · 1 repository · arXiv:2505.19797
-
The Missing Point in Vision Transformers for Universal Image Segmentation 26 May 2025 · 1 repository · arXiv:2505.19795
-
The Role of Video Generation in Enhancing Data-Limited Action Understanding 26 May 2025 · 0 repositories · arXiv:2505.19495
-
Training LLM-Based Agents with Synthetic Self-Reflected Trajectories and Partial Masking 26 May 2025 · 0 repositories · arXiv:2505.20023
-
Transformers in Protein: A Survey 26 May 2025 · 0 repositories · arXiv:2505.20098
-
Understanding Transformer from the Perspective of Associative Memory 26 May 2025 · 0 repositories · arXiv:2505.19488
-
VADER: A Human-Evaluated Benchmark for Vulnerability Assessment, Detection, Explanation, and Remediation 26 May 2025 · 1 repository · arXiv:2505.19395
-
A Smart Healthcare System for Monkeypox Skin Lesion Detection and Tracking 25 May 2025 · 0 repositories · arXiv:2505.19023
-
Assistant-Guided Mitigation of Teacher Preference Bias in LLM-as-a-Judge 25 May 2025 · 1 repository · arXiv:2505.19176
-
Benchmarking Large Language Models for Cyberbullying Detection in Real-World YouTube Comments 25 May 2025 · 0 repositories · arXiv:2505.18927
-
Communication-Efficient Multi-Device Inference Acceleration for Transformer Models 25 May 2025 · 1 repository · arXiv:2505.19342
-
Exploring Magnitude Preservation and Rotation Modulation in Diffusion Transformers 25 May 2025 · 0 repositories · arXiv:2505.19122
-
GhostPrompt: Jailbreaking Text-to-image Generative Models based on Dynamic Optimization 25 May 2025 · 0 repositories · arXiv:2505.18979
-
System-1.5 Reasoning: Traversal in Language and Latent Spaces with Dynamic Shortcuts 25 May 2025 · 0 repositories · arXiv:2505.18962
-
How Does Sequence Modeling Architecture Influence Base Capabilities of Pre-trained Language Models? Exploring Key Architecture Design Principles to Avoid Base Capabilities Degradation 24 May 2025 · 0 repositories · arXiv:2505.18522
-
Localizing Knowledge in Diffusion Transformers 24 May 2025 · 0 repositories · arXiv:2505.18832
-
Security Concerns for Large Language Models: A Survey 24 May 2025 · 0 repositories · arXiv:2505.18889
-
Smart Energy Guardian: A Hybrid Deep Learning Model for Detecting Fraudulent PV Generation 24 May 2025 · 0 repositories · arXiv:2505.18755
-
SW-ViT: A Spatio-Temporal Vision Transformer Network with Post Denoiser for Sequential Multi-Push Ultrasound Shear Wave Elastography 24 May 2025 · 0 repositories · arXiv:2505.18865
-
TrajMoE: Spatially-Aware Mixture of Experts for Unified Human Mobility Modeling 24 May 2025 · 0 repositories · arXiv:2505.18670
-
Unleashing Diffusion Transformers for Visual Correspondence by Modulating Massive Activations 24 May 2025 · 0 repositories · arXiv:2505.18584
-
COLORA: Efficient Fine-Tuning for Convolutional Models with a Study Case on Optical Coherence Tomography Image Classification 23 May 2025 · 0 repositories · arXiv:2505.18315
-
Contrastive Distillation of Emotion Knowledge from LLMs for Zero-Shot Emotion Recognition 23 May 2025 · 1 repository · arXiv:2505.18040
-
Direct3D-S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention 23 May 2025 · 1 repository · arXiv:2505.17412Syntology 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Explainable Anatomy-Guided AI for Prostate MRI: Foundation Models and In Silico Clinical Trials for Virtual Biopsy-based Risk Assessment 23 May 2025 · 0 repositories · arXiv:2505.17971
-
FreqU-FNet: Frequency-Aware U-Net for Imbalanced Medical Image Segmentation 23 May 2025 · 0 repositories · arXiv:2505.17544
-
Gaming Tool Preferences in Agentic LLMs 23 May 2025 · 1 repository · arXiv:2505.18135
-
Hybrid Mamba-Transformer Decoder for Error-Correcting Codes 23 May 2025 · 0 repositories · arXiv:2505.17834
-
Is It Bad to Work All the Time? Cross-Cultural Evaluation of Social Norm Biases in GPT-4 23 May 2025 · 0 repositories · arXiv:2505.18322
-
Multi-Scale Probabilistic Generation Theory: A Hierarchical Framework for Interpreting Large Language Models 23 May 2025 · 0 repositories · arXiv:2505.18244
-
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs 23 May 2025 · 1 repository · arXiv:2505.17598
-
Token Reduction Should Go Beyond Efficiency in Generative Models -- From Vision, Language to Multimodality 23 May 2025 · 1 repository · arXiv:2505.18227
-
Beamforming-Codebook-Aware Channel Knowledge Map Construction for Multi-Antenna Systems 22 May 2025 · 1 repository · arXiv:2505.16132
-
Bottlenecked Transformers: Periodic KV Cache Abstraction for Generalised Reasoning 22 May 2025 · 0 repositories · arXiv:2505.16950