Methods › General › Regularization › Label Smoothing › Papers, page 6
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 6 of 144: papers 501 to 600 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Token-free Models for Sarcasm Detection 2 May 2025 · 0 repositories · arXiv:2505.01006
-
Zero-Shot Document-Level Biomedical Relation Extraction via Scenario-based Prompt Design in Two-Stage with LLM 2 May 2025 · 0 repositories · arXiv:2505.01077
-
DARTer: Dynamic Adaptive Representation Tracker for Nighttime UAV Tracking 1 May 2025 · 0 repositories · arXiv:2505.00752
-
A Time-Series Data Augmentation Model through Diffusion and Transformer Integration 1 May 2025 · 0 repositories · arXiv:2505.03790
-
Directly Forecasting Belief for Reinforcement Learning with Delays 1 May 2025 · 1 repository · arXiv:2505.00546
-
Efficient Recommendation with Millions of Items by Dynamic Pruning of Sub-Item Embeddings 1 May 2025 · 0 repositories · arXiv:2505.00560
-
Enhancing Tropical Cyclone Path Forecasting with an Improved Transformer Network 1 May 2025 · 0 repositories · arXiv:2505.00495
-
Gateformer: Advancing Multivariate Time Series Forecasting through Temporal and Variate-Wise Attention with Gated Representations 1 May 2025 · 1 repository · arXiv:2505.00307Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Open-Source LLM-Driven Federated Transformer for Predictive IoV Management 1 May 2025 · 0 repositories · arXiv:2505.00651
-
Unlocking the Potential of Linear Networks for Irregular Multivariate Time Series Forecasting 1 May 2025 · 0 repositories · arXiv:2505.00590
-
Consistency-aware Fake Videos Detection on Short Video Platforms 30 Apr 2025 · 1 repository · arXiv:2504.21495
-
DOPE: Dual Object Perception-Enhancement Network for Vision-and-Language Navigation 30 Apr 2025 · 0 repositories · arXiv:2505.00743
-
Enhancing Security and Strengthening Defenses in Automated Short-Answer Grading Systems 30 Apr 2025 · 0 repositories · arXiv:2505.00061
-
Advance Fake Video Detection via Vision Transformers 29 Apr 2025 · 0 repositories · arXiv:2504.20669
-
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation 29 Apr 2025 · 0 repositories · arXiv:2504.20629
-
DB-GNN: Dual-Branch Graph Neural Network with Multi-Level Contrastive Learning for Jointly Identifying Within- and Cross-Frequency Coupled Brain Networks 29 Apr 2025 · 0 repositories · arXiv:2504.20744
-
Geolocating Earth Imagery from ISS: Integrating Machine Learning with Astronaut Photography for Enhanced Geographic Mapping 29 Apr 2025 · 1 repository · arXiv:2504.21194
-
In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer 29 Apr 2025 · 0 repositories · arXiv:2504.20690
-
ISDrama: Immersive Spatial Drama Generation through Multimodal Prompting 29 Apr 2025 · 0 repositories · arXiv:2504.20630
-
JaccDiv: A Metric and Benchmark for Quantifying Diversity of Generated Marketing Text in the Music Industry 29 Apr 2025 · 0 repositories · arXiv:2504.20849
-
Multimodal Large Language Models for Medicine: A Comprehensive Survey 29 Apr 2025 · 0 repositories · arXiv:2504.21051
-
SteelBlastQC: Shot-blasted Steel Surface Dataset with Interpretable Detection of Surface Defects 29 Apr 2025 · 1 repository · arXiv:2504.20510
-
Towards Understanding the Nature of Attention with Low-Rank Sparse Decomposition 29 Apr 2025 · 1 repository · arXiv:2504.20938
-
YoChameleon: Personalized Vision and Language Generation 29 Apr 2025 · 0 repositories · arXiv:2504.20998
-
A Transformer-Based Approach for Diagnosing Fault Cases in Optical Fiber Amplifiers 28 Apr 2025 · 0 repositories · arXiv:2505.06245
-
Enhancing Systematic Reviews with Large Language Models: Using GPT-4 and Kimi 28 Apr 2025 · 0 repositories · arXiv:2504.20276
-
Geometry-Informed Neural Operator Transformer 28 Apr 2025 · 1 repository · arXiv:2504.19452
-
m-KAILIN: Knowledge-Driven Agentic Scientific Corpus Distillation Framework for Biomedical Large Language Models Training 28 Apr 2025 · 0 repositories · arXiv:2504.19565
-
UNet with Axial Transformer : A Neural Weather Model for Precipitation Nowcasting 28 Apr 2025 · 1 repository · arXiv:2504.19408
-
From Inductive to Deductive: LLMs-Based Qualitative Data Analysis in Requirements Engineering 27 Apr 2025 · 1 repository · arXiv:2504.19384
-
LM-MCVT: A Lightweight Multi-modal Multi-view Convolutional-Vision Transformer Approach for 3D Object Recognition 27 Apr 2025 · 0 repositories · arXiv:2504.19256
-
TSRM: A Lightweight Temporal Feature Encoding Architecture for Time Series Forecasting and Imputation 26 Apr 2025 · 1 repository · arXiv:2504.18878
-
Why you shouldn't fully trust ChatGPT: A synthesis of this AI tool's error rates across disciplines and the software engineering lifecycle 26 Apr 2025 · 0 repositories · arXiv:2504.18858
-
MEDIBENG WHISPER TINY: A FINE-TUNED CODE-SWITCHED BENGALI-ENGLISH TRANSLATOR FOR CLINICAL APPLICATIONS 25 Apr 2025 · 1 repository
-
A Spatially-Aware Multiple Instance Learning Framework for Digital Pathology 24 Apr 2025 · 1 repository · arXiv:2504.17379
-
Masked strategies for images with small objects 24 Apr 2025 · 0 repositories · arXiv:2504.17935
-
The Sparse Frontier: Sparse Attention Trade-offs in Transformer LLMs 24 Apr 2025 · 0 repositories · arXiv:2504.17768
-
Token-Shuffle: Towards High-Resolution Image Generation with Autoregressive Models 24 Apr 2025 · 0 repositories · arXiv:2504.17789
-
A Novel Hybrid Approach Using an Attention-Based Transformer + GRU Model for Predicting Cryptocurrency Prices 23 Apr 2025 · 0 repositories · arXiv:2504.17079
-
Advanced Chest X-Ray Analysis via Transformer-Based Image Descriptors and Cross-Model Attention Mechanism 23 Apr 2025 · 0 repositories · arXiv:2504.16774
-
Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate 23 Apr 2025 · 0 repositories · arXiv:2504.16489
-
Emo Pillars: Knowledge Distillation to Support Fine-Grained Context-Aware and Context-Less Emotion Classification 23 Apr 2025 · 0 repositories · arXiv:2504.16856
-
From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories 23 Apr 2025 · 0 repositories · arXiv:2504.16449
-
Leveraging LLMs as Meta-Judges: A Multi-Agent Framework for Evaluating LLM Judgments 23 Apr 2025 · 0 repositories · arXiv:2504.17087
-
Transformers for Complex Query Answering over Knowledge Hypergraphs 23 Apr 2025 · 0 repositories · arXiv:2504.16537
-
A Large-scale Class-level Benchmark Dataset for Code Generation with LLMs 22 Apr 2025 · 0 repositories · arXiv:2504.15564
-
COBRA: Algorithm-Architecture Co-optimized Binary Transformer Accelerator for Edge Inference 22 Apr 2025 · 0 repositories · arXiv:2504.16269
-
DiTPainter: Efficient Video Inpainting with Diffusion Transformers 22 Apr 2025 · 0 repositories · arXiv:2504.15661
-
DSDNet: Raw Domain Demoiréing via Dual Color-Space Synergy 22 Apr 2025 · 0 repositories · arXiv:2504.15756
-
LongMamba: Enhancing Mamba's Long Context Capabilities via Training-Free Receptive Field Enlargement 22 Apr 2025 · 1 repository · arXiv:2504.16053
-
Quantum Doubly Stochastic Transformers 22 Apr 2025 · 0 repositories · arXiv:2504.16275Syntology 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Acquire and then Adapt: Squeezing out Text-to-Image Model for Image Restoration 21 Apr 2025 · 0 repositories · arXiv:2504.15159
-
An Efficient Aerial Image Detection with Variable Receptive Fields 21 Apr 2025 · 0 repositories · arXiv:2504.15165
-
Automated Measurement of Eczema Severity with Self-Supervised Learning 21 Apr 2025 · 0 repositories · arXiv:2504.15193
-
Distribution-aware Dataset Distillation for Efficient Image Restoration 21 Apr 2025 · 0 repositories · arXiv:2504.14826
-
ECViT: Efficient Convolutional Vision Transformer with Local-Attention and Multi-scale Stages 21 Apr 2025 · 1 repository · arXiv:2504.14825
-
Efficient Pretraining Length Scaling 21 Apr 2025 · 0 repositories · arXiv:2504.14992
-
Impact of Latent Space Dimension on IoT Botnet Detection Performance: VAE-Encoder Versus ViT-Encoder 21 Apr 2025 · 0 repositories · arXiv:2504.14879
-
Insert Anything: Image Insertion via In-Context Editing in DiT 21 Apr 2025 · 0 repositories · arXiv:2504.15009
-
LLMs as Data Annotators: How Close Are We to Human Performance 21 Apr 2025 · 0 repositories · arXiv:2504.15022
-
Mitigating Degree Bias in Graph Representation Learning with Learnable Structural Augmentation and Structural Self-Attention 21 Apr 2025 · 1 repository · arXiv:2504.15075
-
MoE Parallel Folding: Heterogeneous Parallelism Mappings for Efficient Large-Scale MoE Model Training with Megatron Core 21 Apr 2025 · 0 repositories · arXiv:2504.14960
-
Roll the dice & look before you leap: Going beyond the creative limits of next-token prediction 21 Apr 2025 · 1 repository · arXiv:2504.15266
-
Structure-guided Diffusion Transformer for Low-Light Image Enhancement 21 Apr 2025 · 0 repositories · arXiv:2504.15054
-
A Framework for Benchmarking and Aligning Task-Planning Safety in LLM-Based Embodied Agents 20 Apr 2025 · 0 repositories · arXiv:2504.14650
-
Enhancing DR Classification with Swin Transformer and Shifted Window Attention 20 Apr 2025 · 0 repositories · arXiv:2504.15317
-
MSAD-Net: Multiscale and Spatial Attention-based Dense Network for Lung Cancer Classification 20 Apr 2025 · 0 repositories · arXiv:2504.14626
-
Empirical Evaluation of Knowledge Distillation from Transformers to Subquadratic Language Models 19 Apr 2025 · 0 repositories · arXiv:2504.14366
-
Know Me, Respond to Me: Benchmarking LLMs for Dynamic User Profiling and Personalized Responses at Scale 19 Apr 2025 · 1 repository · arXiv:2504.14225Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Lightweight Road Environment Segmentation using Vector Quantization 19 Apr 2025 · 0 repositories · arXiv:2504.14113
-
SimplifyMyText: An LLM-Based System for Inclusive Plain Language Text Simplification 19 Apr 2025 · 0 repositories · arXiv:2504.14223
-
6G WavesFM: A Foundation Model for Sensing, Communication, and Localization 18 Apr 2025 · 0 repositories · arXiv:2504.14100
-
A Deep Learning-Based Supervised Transfer Learning Framework for DOA Estimation with Array Imperfections 18 Apr 2025 · 1 repository · arXiv:2504.13394
-
BeetleVerse: A study on taxonomic classification of ground beetles 18 Apr 2025 · 0 repositories · arXiv:2504.13393
-
Collective Learning Mechanism based Optimal Transport Generative Adversarial Network for Non-parallel Voice Conversion 18 Apr 2025 · 0 repositories · arXiv:2504.13791
-
CytoFM: The first cytology foundation model 18 Apr 2025 · 0 repositories · arXiv:2504.13402
-
DAM-Net: Domain Adaptation Network with Micro-Labeled Fine-Tuning for Change Detection 18 Apr 2025 · 0 repositories · arXiv:2504.13748
-
DenSe-AdViT: A novel Vision Transformer for Dense SAR Object Detection 18 Apr 2025 · 0 repositories · arXiv:2504.13638
-
HAECcity: Open-Vocabulary Scene Understanding of City-Scale Point Clouds with Superpoint Graph Clustering 18 Apr 2025 · 0 repositories · arXiv:2504.13590
-
HMPE:HeatMap Embedding for Efficient Transformer-Based Small Object Detection 18 Apr 2025 · 0 repositories · arXiv:2504.13469
-
Integrating Locality-Aware Attention with Transformers for General Geometry PDEs 18 Apr 2025 · 0 repositories · arXiv:2504.13480
-
Lightweight LiDAR-Camera 3D Dynamic Object Detection and Multi-Class Trajectory Prediction 18 Apr 2025 · 2 repositories · arXiv:2504.13647
-
Signatures of human-like processing in Transformer forward passes 18 Apr 2025 · 1 repository · arXiv:2504.14107
-
LLM Sensitivity Evaluation Framework for Clinical Diagnosis 18 Apr 2025 · 0 repositories · arXiv:2504.13475
-
LoRA-Based Continual Learning with Constraints on Critical Parameter Changes 18 Apr 2025 · 1 repository · arXiv:2504.13407
-
MMformer with Adaptive Transferable Attention: Advancing Multivariate Time Series Forecasting for Environmental Applications 18 Apr 2025 · 0 repositories · arXiv:2504.14050
-
PV-VLM: A Multimodal Vision-Language Approach Incorporating Sky Images for Intra-Hour Photovoltaic Power Forecasting 18 Apr 2025 · 0 repositories · arXiv:2504.13624
-
Retinex-guided Histogram Transformer for Mask-free Shadow Removal 18 Apr 2025 · 1 repository · arXiv:2504.14092
-
SatelliteCalculator: A Multi-Task Vision Foundation Model for Quantitative Remote Sensing Inversion 18 Apr 2025 · 0 repositories · arXiv:2504.13442
-
Towards Accurate and Interpretable Neuroblastoma Diagnosis via Contrastive Multi-scale Pathological Image Analysis 18 Apr 2025 · 1 repository · arXiv:2504.13754
-
Towards Scale-Aware Low-Light Enhancement via Structure-Guided Transformer Design 18 Apr 2025 · 1 repository · arXiv:2504.14075
-
Transformer Encoder and Multi-features Time2Vec for Financial Prediction 18 Apr 2025 · 0 repositories · arXiv:2504.13801
-
Transformers Can Overcome the Curse of Dimensionality: A Theoretical Study from an Approximation Perspective 18 Apr 2025 · 0 repositories · arXiv:2504.13558
-
Accuracy is Not Agreement: Expert-Aligned Evaluation of Crash Narrative Classification Models 17 Apr 2025 · 0 repositories · arXiv:2504.13068
-
Exploring Expert Failures Improves LLM Agent Tuning 17 Apr 2025 · 0 repositories · arXiv:2504.13145
-
Plain Transformers Can be Powerful Graph Learners 17 Apr 2025 · 0 repositories · arXiv:2504.12588
-
SSTAF: Spatial-Spectral-Temporal Attention Fusion Transformer for Motor Imagery Classification 17 Apr 2025 · 0 repositories · arXiv:2504.13220
-
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation 16 Apr 2025 · 0 repositories · arXiv:2504.11942
-
Approximation Bounds for Transformer Networks with Application to Regression 16 Apr 2025 · 0 repositories · arXiv:2504.12175
-
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts 16 Apr 2025 · 1 repository · arXiv:2504.12463Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)