Methods › General › Regularization › Label Smoothing › Papers, page 56
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 56 of 144: papers 5,501 to 5,600 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
AutoAugment Is What You Need: Enhancing Rule-based Augmentation Methods in Low-resource Regimes 8 Feb 2024 · 1 repository · arXiv:2402.05584
-
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer 8 Feb 2024 · 1 repository · arXiv:2402.05712
-
GPT-4 Generated Narratives of Life Events using a Structured Narrative Prompt: A Validation Study 8 Feb 2024 · 0 repositories · arXiv:2402.05435
-
An Examination on the Effectiveness of Divide-and-Conquer Prompting in Large Language Models 8 Feb 2024 · 0 repositories · arXiv:2402.05359
-
How do Transformers perform In-Context Autoregressive Learning? 8 Feb 2024 · 0 repositories · arXiv:2402.05787Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
How Well Can LLMs Negotiate? NegotiationArena Platform and Analysis 8 Feb 2024 · 1 repository · arXiv:2402.05863Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
In-Context Principle Learning from Mistakes 8 Feb 2024 · 1 repository · arXiv:2402.05403
-
Large Language Model Meets Graph Neural Network in Knowledge Distillation 8 Feb 2024 · 0 repositories · arXiv:2402.05894
-
Large Language Models for Psycholinguistic Plausibility Pretesting 8 Feb 2024 · 0 repositories · arXiv:2402.05455
-
Limits of Transformer Language Models on Learning to Compose Algorithms 8 Feb 2024 · 1 repository · arXiv:2402.05785Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 8 harvested samples)
-
LLMs Among Us: Generative AI Participating in Digital Discourse 8 Feb 2024 · 0 repositories · arXiv:2402.07940
-
Mamba-ND: Selective State Space Modeling for Multi-Dimensional Data 8 Feb 2024 · 1 repository · arXiv:2402.05892Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Noise Contrastive Alignment of Language Models with Explicit Rewards 8 Feb 2024 · 3 repositories · arXiv:2402.05369Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
On Convolutional Vision Transformers for Yield Prediction 8 Feb 2024 · 0 repositories · arXiv:2402.05557
-
Question Aware Vision Transformer for Multimodal Reasoning 8 Feb 2024 · 0 repositories · arXiv:2402.05472
-
Self-Alignment of Large Language Models via Monopolylogue-based Social Scene Simulation 8 Feb 2024 · 0 repositories · arXiv:2402.05699
-
Sparse-VQ Transformer: An FFN-Free Framework with Vector Quantization for Enhanced Time Series Forecasting 8 Feb 2024 · 0 repositories · arXiv:2402.05830
-
TimeArena: Shaping Efficient Multitasking Language Agents in a Time-Aware Simulation 8 Feb 2024 · 0 repositories · arXiv:2402.05733
-
Unleashing the Infinity Power of Geometry: A Novel Geometry-Aware Transformer (GOAT) for Whole Slide Histopathology Image Analysis 8 Feb 2024 · 0 repositories · arXiv:2402.05373
-
You Only Need One Color Space: An Efficient Network for Low-light Image Enhancement 8 Feb 2024 · 1 repository · arXiv:2402.05809
-
Zero-Shot Chain-of-Thought Reasoning Guided by Evolutionary Algorithms in Large Language Models 8 Feb 2024 · 0 repositories · arXiv:2402.05376
-
Can Large Language Model Agents Simulate Human Trust Behavior? 7 Feb 2024 · 1 repository · arXiv:2402.04559Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Conversational Assistants in Knowledge-Intensive Contexts: An Evaluation of LLM- versus Intent-based Systems 7 Feb 2024 · 0 repositories · arXiv:2402.04955
-
Improving Cross-Domain Low-Resource Text Generation through LLM Post-Editing: A Programmer-Interpreter Approach 7 Feb 2024 · 0 repositories · arXiv:2402.04609
-
Latent Plan Transformer for Trajectory Abstraction: Planning as Latent Space Inference 7 Feb 2024 · 1 repository · arXiv:2402.04647Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 2 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning 7 Feb 2024 · 1 repository · arXiv:2402.04833Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Navigating the Knowledge Sea: Planet-scale answer retrieval using LLMs 7 Feb 2024 · 0 repositories · arXiv:2402.05318
-
Opening the AI black box: program synthesis via mechanistic interpretability 7 Feb 2024 · 1 repository · arXiv:2402.05110Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Pseudo-labelling meets Label Smoothing for Noisy Partial Label Learning 7 Feb 2024 · 0 repositories · arXiv:2402.04835
-
Majority Kernels: An Approach to Leverage Big Model Dynamics for Efficient Small Model Training 7 Feb 2024 · 0 repositories · arXiv:2402.05033
-
StableMask: Refining Causal Masking in Decoder-only Transformer 7 Feb 2024 · 0 repositories · arXiv:2402.04779
-
Toward Accurate Camera-based 3D Object Detection via Cascade Depth Estimation and Calibration 7 Feb 2024 · 0 repositories · arXiv:2402.04883
-
TransLLaMa: LLM-based Simultaneous Translation System 7 Feb 2024 · 1 repository · arXiv:2402.04636
-
Triplet Interaction Improves Graph Transformers: Accurate Molecular Graph Learning with Triplet Graph Transformers 7 Feb 2024 · 3 repositories · arXiv:2402.04538Syntology official (archive's flag): 9 ran · 10 ran (of which 1 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
AnyTool: Self-Reflective, Hierarchical Agents for Large-Scale API Calls 6 Feb 2024 · 1 repository · arXiv:2402.04253Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Behind the Screen: Investigating ChatGPT's Dark Personality Traits and Conspiracy Beliefs 6 Feb 2024 · 0 repositories · arXiv:2402.04110
-
Breaking Symmetry When Training Transformers 6 Feb 2024 · 0 repositories · arXiv:2402.05969
-
Can Mamba Learn How to Learn? A Comparative Study on In-Context Learning Tasks 6 Feb 2024 · 2 repositories · arXiv:2402.04248Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples)
-
CAST: Clustering Self-Attention using Surrogate Tokens for Efficient Transformers 6 Feb 2024 · 0 repositories · arXiv:2402.04239
-
Comparing Abstraction in Humans and Large Language Models Using Multimodal Serial Reproduction 6 Feb 2024 · 0 repositories · arXiv:2402.03618
-
Cross Entropy versus Label Smoothing: A Neural Collapse Perspective 6 Feb 2024 · 0 repositories · arXiv:2402.03979
-
Detection Transformer for Teeth Detection, Segmentation, and Numbering in Oral Rare Diseases: Focus on Data Augmentation and Inpainting Techniques 6 Feb 2024 · 0 repositories · arXiv:2402.04408
-
Transformer based Endmember Fusion with Spatial Context for Hyperspectral Unmixing 6 Feb 2024 · 0 repositories · arXiv:2402.03835
-
Identifying Reasons for Contraceptive Switching from Real-World Data Using Large Language Models 6 Feb 2024 · 1 repository · arXiv:2402.03597
-
Iterative Prompt Refinement for Radiation Oncology Symptom Extraction Using Teacher-Student Large Language Models 6 Feb 2024 · 0 repositories · arXiv:2402.04075
-
Large Language Models As MOOCs Graders 6 Feb 2024 · 0 repositories · arXiv:2402.03776
-
Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs 6 Feb 2024 · 0 repositories · arXiv:2402.03927
-
Lens: A Foundation Model for Network Traffic 6 Feb 2024 · 0 repositories · arXiv:2402.03646
-
LLM Agents can Autonomously Hack Websites 6 Feb 2024 · 0 repositories · arXiv:2402.06664
-
Are Machines Better at Complex Reasoning? Unveiling Human-Machine Inference Gaps in Entailment Verification 6 Feb 2024 · 0 repositories · arXiv:2402.03686
-
Parameter-tuning-free data entry error unlearning with adaptive selective synaptic dampening 6 Feb 2024 · 1 repository · arXiv:2402.10098
-
Pard: Permutation-Invariant Autoregressive Diffusion for Graph Generation 6 Feb 2024 · 1 repository · arXiv:2402.03687
-
Pre-training of Lightweight Vision Transformers on Small Datasets with Minimally Scaled Images 6 Feb 2024 · 0 repositories · arXiv:2402.03752
-
Professional Agents -- Evolving Large Language Models into Autonomous Experts with Human-Level Competencies 6 Feb 2024 · 0 repositories · arXiv:2402.03628
-
Reinforcement Learning from Bagged Reward 6 Feb 2024 · 0 repositories · arXiv:2402.03771
-
Return-Aligned Decision Transformer 6 Feb 2024 · 0 repositories · arXiv:2402.03923
-
Self-Discover: Large Language Models Self-Compose Reasoning Structures 6 Feb 2024 · 3 repositories · arXiv:2402.03620Syntology 0 ran · 3 unverified (of 3 harvested samples)
-
NeRCC: Nested-Regression Coded Computing for Resilient Distributed Prediction Serving Systems 6 Feb 2024 · 0 repositories · arXiv:2402.04377
-
The Hedgehog & the Porcupine: Expressive Linear Attentions with Softmax Mimicry 6 Feb 2024 · 1 repository · arXiv:2402.04347
-
U-shaped Vision Mamba for Single Image Dehazing 6 Feb 2024 · 1 repository · arXiv:2402.04139
-
A Survey on Transformer Compression 5 Feb 2024 · 0 repositories · arXiv:2402.05964
-
Reconstruct Your Previous Conversations! Comprehensively Investigating Privacy Leakage Risks in Conversations with GPT Models 5 Feb 2024 · 1 repository · arXiv:2402.02987
-
Cross-Domain Few-Shot Object Detection via Enhanced Open-Set Object Detector 5 Feb 2024 · 2 repositories · arXiv:2402.03094Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models 5 Feb 2024 · 5 repositories · arXiv:2402.03300Syntology official (archive's flag): 8 ran · 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 10 unverified (of 24 harvested samples) · 3 pointer-only (licence)
-
DiffsFormer: A Diffusion Transformer on Stock Factor Augmentation 5 Feb 2024 · 0 repositories · arXiv:2402.06656
-
Graph-enhanced Large Language Models in Asynchronous Plan Reasoning 5 Feb 2024 · 1 repository · arXiv:2402.02805Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
HAMLET: Graph Transformer Neural Operator for Partial Differential Equations 5 Feb 2024 · 0 repositories · arXiv:2402.03541
-
Harnessing PubMed User Query Logs for Post Hoc Explanations of Recommended Similar Articles 5 Feb 2024 · 0 repositories · arXiv:2402.03484
-
Illuminate: A novel approach for depression detection with explainable analysis and proactive therapy using prompt engineering 5 Feb 2024 · 0 repositories · arXiv:2402.05127
-
Is Mamba Capable of In-Context Learning? 5 Feb 2024 · 1 repository · arXiv:2402.03170Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
MobilityGPT: Enhanced Human Mobility Modeling with a GPT model 5 Feb 2024 · 0 repositories · arXiv:2402.03264
-
SWAG: Storytelling With Action Guidance 5 Feb 2024 · 1 repository · arXiv:2402.03483
-
Toward Human-AI Alignment in Large-Scale Multi-Player Games 5 Feb 2024 · 0 repositories · arXiv:2402.03575
-
Towards Eliminating Hard Label Constraints in Gradient Inversion Attacks 5 Feb 2024 · 1 repository · arXiv:2402.03124
-
UniMem: Towards a Unified View of Long-Context Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03009
-
Aligner: Efficient Alignment by Learning to Correct 4 Feb 2024 · 0 repositories · arXiv:2402.02416
-
Evaluating Large Language Models in Analysing Classroom Dialogue 4 Feb 2024 · 0 repositories · arXiv:2402.02380
-
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation 4 Feb 2024 · 0 repositories · arXiv:2402.14594
-
INViT: A Generalizable Routing Problem Solver with Invariant Nested View Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02317Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Key-Graph Transformer for Image Restoration 4 Feb 2024 · 0 repositories · arXiv:2402.02634
-
Minusformer: Improving Time Series Forecasting by Progressively Learning Residuals 4 Feb 2024 · 1 repository · arXiv:2402.02332
-
Pathformer: Multi-scale Transformers with Adaptive Pathways for Time Series Forecasting 4 Feb 2024 · 1 repository · arXiv:2402.05956Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
PROSAC: Provably Safe Certification for Machine Learning Models under Adversarial Attacks 4 Feb 2024 · 0 repositories · arXiv:2402.02629
-
SPCTNet: A Series-Parallel CNN and Transformer Network for 3D Medical Image Segmentation 4 Feb 2024 · 0 repositories
-
Spin: An Efficient Secure Computation Framework with GPU Acceleration 4 Feb 2024 · 0 repositories · arXiv:2402.02320
-
Timer: Generative Pre-trained Transformers Are Large Time Series Models 4 Feb 2024 · 1 repository · arXiv:2402.02368Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Unified Training of Universal Time Series Forecasting Transformers 4 Feb 2024 · 1 repository · arXiv:2402.02592
-
BetterV: Controlled Verilog Generation with Discriminative Guidance 3 Feb 2024 · 0 repositories · arXiv:2402.03375
-
DiffVein: A Unified Diffusion Network for Finger Vein Segmentation and Authentication 3 Feb 2024 · 0 repositories · arXiv:2402.02060
-
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test 3 Feb 2024 · 0 repositories · arXiv:2402.02135
-
EffiBench: Benchmarking the Efficiency of Automatically Generated Code 3 Feb 2024 · 1 repository · arXiv:2402.02037Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
How well do LLMs cite relevant medical references? An evaluation framework and analyses 3 Feb 2024 · 1 repository · arXiv:2402.02008
-
IMUSE: IMU-based Facial Expression Capture 3 Feb 2024 · 0 repositories · arXiv:2402.03944
-
ParZC: Parametric Zero-Cost Proxies for Efficient NAS 3 Feb 2024 · 0 repositories · arXiv:2402.02105
-
ScribFormer: Transformer Makes CNN Work Better for Scribble-based Medical Image Segmentation 3 Feb 2024 · 1 repository · arXiv:2402.02029
-
TCI-Former: Thermal Conduction-Inspired Transformer for Infrared Small Target Detection 3 Feb 2024 · 0 repositories · arXiv:2402.02046
-
Topology-Informed Graph Transformer 3 Feb 2024 · 2 repositories · arXiv:2402.02005Syntology official (archive's flag): 8 ran · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 2 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Hierarchical Structure Enhances the Convergence and Generalizability of Linear Molecular Representation 3 Feb 2024 · 1 repository · arXiv:2402.02164
-
A Data-Driven Analysis of Robust Automatic Piano Transcription 2 Feb 2024 · 0 repositories · arXiv:2402.01424
-
ALERT-Transformer: Bridging Asynchronous and Synchronous Machine Learning for Real-Time Event-based Spatio-Temporal Data 2 Feb 2024 · 0 repositories · arXiv:2402.01393