Methods › General › Regularization › Label Smoothing › Papers, page 55
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 55 of 144: papers 5,401 to 5,500 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
How Reliable Are Automatic Evaluation Methods for Instruction-Tuned LLMs? 16 Feb 2024 · 0 repositories · arXiv:2402.10770
-
In Search of Needles in a 11M Haystack: Recurrent Memory Finds What LLMs Miss 16 Feb 2024 · 2 repositories · arXiv:2402.10790
-
When "Competency" in Reasoning Opens the Door to Vulnerability: Jailbreaking LLMs via Novel Complex Ciphers 16 Feb 2024 · 1 repository · arXiv:2402.10601Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Large Language Models as Zero-shot Dialogue State Tracker through Function Calling 16 Feb 2024 · 1 repository · arXiv:2402.10466Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Large Language Models Fall Short: Understanding Complex Relationships in Detective Narratives 16 Feb 2024 · 0 repositories · arXiv:2402.11051
-
Linear Transformers with Learnable Kernel Functions are Better In-Context Models 16 Feb 2024 · 2 repositories · arXiv:2402.10644Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
LinkNER: Linking Local Named Entity Recognition Models to Large Language Models using Uncertainty 16 Feb 2024 · 1 repository · arXiv:2402.10573
-
ToolSword: Unveiling Safety Issues of Large Language Models in Tool Learning Across Three Stages 16 Feb 2024 · 1 repository · arXiv:2402.10753
-
Weak-Mamba-UNet: Visual Mamba Makes CNN and ViT Work Better for Scribble-based Medical Image Segmentation 16 Feb 2024 · 2 repositories · arXiv:2402.10887Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
A StrongREJECT for Empty Jailbreaks 15 Feb 2024 · 2 repositories · arXiv:2402.10260Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples)
-
An Analysis of Language Frequency and Error Correction for Esperanto 15 Feb 2024 · 0 repositories · arXiv:2402.09696
-
Camouflage is all you need: Evaluating and Enhancing Language Model Robustness Against Camouflage Adversarial Attacks 15 Feb 2024 · 0 repositories · arXiv:2402.09874
-
Data Engineering for Scaling Language Models to 128K Context 15 Feb 2024 · 1 repository · arXiv:2402.10171Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Fine-tuning Large Language Model (LLM) Artificial Intelligence Chatbots in Ophthalmology and LLM-based evaluation using GPT-4 15 Feb 2024 · 0 repositories · arXiv:2402.10083
-
GPT-4's assessment of its performance in a USMLE-based case study 15 Feb 2024 · 0 repositories · arXiv:2402.09654
-
Improving Non-autoregressive Machine Translation with Error Exposure and Consistency Regularization 15 Feb 2024 · 0 repositories · arXiv:2402.09725
-
Large Language Models for Forecasting and Anomaly Detection: A Systematic Literature Review 15 Feb 2024 · 0 repositories · arXiv:2402.10350
-
NYCTALE: Neuro-Evidence Transformer for Adaptive and Personalized Lung Nodule Invasiveness Prediction 15 Feb 2024 · 0 repositories · arXiv:2402.10066
-
OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset 15 Feb 2024 · 1 repository · arXiv:2402.10176
-
Prompt-Based Bias Calibration for Better Zero/Few-Shot Learning of Language Models 15 Feb 2024 · 0 repositories · arXiv:2402.10353
-
ProtChatGPT: Towards Understanding Proteins with Large Language Models 15 Feb 2024 · 0 repositories · arXiv:2402.09649
-
Spike-driven Transformer V2: Meta Spiking Neural Network Architecture Inspiring the Design of Next-generation Neuromorphic Chips 15 Feb 2024 · 1 repository · arXiv:2404.03663Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence 15 Feb 2024 · 1 repository · arXiv:2402.10175
-
X-lifecycle Learning for Cloud Incident Management using LLMs 15 Feb 2024 · 0 repositories · arXiv:2404.03662
-
API Pack: A Massive Multi-Programming Language Dataset for API Call Generation 14 Feb 2024 · 1 repository · arXiv:2402.09615
-
AQA-Bench: An Interactive Benchmark for Evaluating LLMs' Sequential Reasoning Ability 14 Feb 2024 · 1 repository · arXiv:2402.09404
-
Bidirectional Generative Pre-training for Improving Healthcare Time-series Representation Learning 14 Feb 2024 · 1 repository · arXiv:2402.09558
-
Changes by Butterflies: Farsighted Forecasting with Group Reservoir Transformer 14 Feb 2024 · 0 repositories · arXiv:2402.09573
-
Context Composing for Full Line Code Completion 14 Feb 2024 · 0 repositories · arXiv:2402.09230
-
FGeo-TP: A Language Model-Enhanced Solver for Geometry Problems 14 Feb 2024 · 0 repositories · arXiv:2402.09047
-
HEAL-ViT: Vision Transformers on a spherical mesh for medium-range weather forecasting 14 Feb 2024 · 0 repositories · arXiv:2403.17016
-
L3GO: Language Agents with Chain-of-3D-Thoughts for Generating Unconventional Objects 14 Feb 2024 · 0 repositories · arXiv:2402.09052
-
Leveraging Large Language Models for Enhanced NLP Task Performance through Knowledge Distillation and Optimized Training Strategies 14 Feb 2024 · 0 repositories · arXiv:2402.09282
-
LlaSMol: Advancing Large Language Models for Chemistry with a Large-Scale, Comprehensive, High-Quality Instruction Tuning Dataset 14 Feb 2024 · 1 repository · arXiv:2402.09391Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Pyramid Attention Network for Medical Image Registration 14 Feb 2024 · 1 repository · arXiv:2402.09016
-
Research and application of Transformer based anomaly detection model: A literature review 14 Feb 2024 · 0 repositories · arXiv:2402.08975
-
AutoTutor meets Large Language Models: A Language Model Tutor with Rich Pedagogy and Guardrails 14 Feb 2024 · 1 repository · arXiv:2402.09216Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Stochastic Spiking Attention: Accelerating Attention with Stochastic Computing in Spiking Networks 14 Feb 2024 · 0 repositories · arXiv:2402.09109
-
TDViT: Temporal Dilated Video Transformer for Dense Video Tasks 14 Feb 2024 · 1 repository · arXiv:2402.09257
-
Towards Next-Level Post-Training Quantization of Hyper-Scale Transformers 14 Feb 2024 · 0 repositories · arXiv:2402.08958Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
BBox-Adapter: Lightweight Adapting for Black-Box Large Language Models 13 Feb 2024 · 1 repository · arXiv:2402.08219Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
BEFUnet: A Hybrid CNN-Transformer Architecture for Precise Medical Image Segmentation 13 Feb 2024 · 1 repository · arXiv:2402.08793
-
Combining Insights From Multiple Large Language Models Improves Diagnostic Accuracy 13 Feb 2024 · 0 repositories · arXiv:2402.08806
-
eCeLLM: Generalizing Large Language Models for E-commerce from Large-scale, High-quality Instruction Data 13 Feb 2024 · 1 repository · arXiv:2402.08831Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
InstructGraph: Boosting Large Language Models via Graph-centric Instruction Tuning and Preference Alignment 13 Feb 2024 · 1 repository · arXiv:2402.08785
-
Large Language Models for the Automated Analysis of Optimization Algorithms 13 Feb 2024 · 1 repository · arXiv:2402.08472
-
LLaGA: Large Language and Graph Assistant 13 Feb 2024 · 2 repositories · arXiv:2402.08170Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
MetaTra: Meta-Learning for Generalized Trajectory Prediction in Unseen Domain 13 Feb 2024 · 0 repositories · arXiv:2402.08221
-
On Limitations of the Transformer Architecture 13 Feb 2024 · 0 repositories · arXiv:2402.08164
-
Optimized Information Flow for Transformer Tracking 13 Feb 2024 · 1 repository · arXiv:2402.08195
-
PRompt Optimization in Multi-Step Tasks (PROMST): Integrating Human Feedback and Heuristic-based Sampling 13 Feb 2024 · 1 repository · arXiv:2402.08702Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
The Last JITAI? Exploring Large Language Models for Issuing Just-in-Time Adaptive Interventions: Fostering Physical Activity in a Conceptual Cardiac Rehabilitation Setting 13 Feb 2024 · 0 repositories · arXiv:2402.08658
-
Transformer Mechanisms Mimic Frontostriatal Gating Operations When Trained on Human Working Memory Tasks 13 Feb 2024 · 0 repositories · arXiv:2402.08211
-
Translating Images to Road Network: A Sequence-to-Sequence Perspective 13 Feb 2024 · 3 repositories · arXiv:2402.08207
-
Addressing cognitive bias in medical language models 12 Feb 2024 · 1 repository · arXiv:2402.08113Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension 12 Feb 2024 · 1 repository · arXiv:2402.07729
-
AYDIV: Adaptable Yielding 3D Object Detection via Integrated Contextual Vision Transformer 12 Feb 2024 · 1 repository · arXiv:2402.07680
-
BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data 12 Feb 2024 · 0 repositories · arXiv:2402.08093
-
ClusterTabNet: Supervised clustering method for table detection and table structure recognition 12 Feb 2024 · 1 repository · arXiv:2402.07502Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 16 harvested samples) · 4 pointer-only (licence)
-
Dólares or Dollars? Unraveling the Bilingual Prowess of Financial LLMs Between Spanish and English 12 Feb 2024 · 1 repository · arXiv:2402.07405
-
Enhancing Multi-Criteria Decision Analysis with AI: Integrating Analytic Hierarchy Process and GPT-4 for Automated Decision Support 12 Feb 2024 · 0 repositories · arXiv:2402.07404
-
Enhancing Programming Error Messages in Real Time with Generative AI 12 Feb 2024 · 0 repositories · arXiv:2402.08072
-
Fourier Circuits in Neural Networks and Transformers: A Case Study of Modular Arithmetic with Multiple Inputs 12 Feb 2024 · 0 repositories · arXiv:2402.09469
-
Large Language Models "Ad Referendum": How Good Are They at Machine Translation in the Legal Domain? 12 Feb 2024 · 0 repositories · arXiv:2402.07681
-
Large Language Models are Few-shot Generators: Proposing Hybrid Prompt Algorithm To Generate Webshell Escape Samples 12 Feb 2024 · 1 repository · arXiv:2402.07408
-
Leveraging AI to Advance Science and Computing Education across Africa: Challenges, Progress and Opportunities 12 Feb 2024 · 0 repositories · arXiv:2402.07397
-
Message Detouring: A Simple Yet Effective Cycle Representation for Expressive Graph Learning 12 Feb 2024 · 0 repositories · arXiv:2402.08085
-
On the Self-Verification Limitations of Large Language Models on Reasoning and Planning Tasks 12 Feb 2024 · 0 repositories · arXiv:2402.08115
-
Only the Curve Shape Matters: Training Foundation Models for Zero-Shot Multivariate Time Series Forecasting through Next Curve Shape Prediction 12 Feb 2024 · 0 repositories · arXiv:2402.07570
-
Secret Collusion among Generative AI Agents: Multi-Agent Deception via Steganography 12 Feb 2024 · 0 repositories · arXiv:2402.07510
-
Sheet Music Transformer: End-To-End Optical Music Recognition Beyond Monophonic Transcription 12 Feb 2024 · 1 repository · arXiv:2402.07596
-
Suppressing Pink Elephants with Direct Principle Feedback 12 Feb 2024 · 0 repositories · arXiv:2402.07896
-
The I/O Complexity of Attention, or How Optimal is Flash Attention? 12 Feb 2024 · 0 repositories · arXiv:2402.07443
-
Towards an Understanding of Stepwise Inference in Transformers: A Synthetic Graph Navigation Model 12 Feb 2024 · 0 repositories · arXiv:2402.07757
-
TransAxx: Efficient Transformers with Approximate Computing 12 Feb 2024 · 0 repositories · arXiv:2402.07545
-
VCR: Video representation for Contextual Retrieval 12 Feb 2024 · 1 repository · arXiv:2402.07466
-
How do Large Language Models Navigate Conflicts between Honesty and Helpfulness? 11 Feb 2024 · 0 repositories · arXiv:2402.07282
-
Large-Language-Model Empowered Dose Volume Histogram Prediction for Intensity Modulated Radiotherapy 11 Feb 2024 · 0 repositories · arXiv:2402.07167
-
Multi-Modal Emotion Recognition by Text, Speech and Video Using Pretrained Transformers 11 Feb 2024 · 0 repositories · arXiv:2402.07327
-
Natural Language Reinforcement Learning 11 Feb 2024 · 0 repositories · arXiv:2402.07157
-
Semi-Mamba-UNet: Pixel-Level Contrastive and Pixel-Level Cross-Supervised Visual Mamba-based UNet for Semi-Supervised Medical Image Segmentation 11 Feb 2024 · 1 repository · arXiv:2402.07245Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
ChemLLM: A Chemical Large Language Model 10 Feb 2024 · 1 repository · arXiv:2402.06852
-
For Better or For Worse? Learning Minimum Variance Features With Label Augmentation 10 Feb 2024 · 0 repositories · arXiv:2402.06855
-
Gemini Goes to Med School: Exploring the Capabilities of Multimodal Large Language Models on Medical Challenge Problems & Hallucinations 10 Feb 2024 · 1 repository · arXiv:2402.07023
-
NLP for Knowledge Discovery and Information Extraction from Energetics Corpora 10 Feb 2024 · 0 repositories · arXiv:2402.06964
-
OpenFedLLM: Training Large Language Models on Decentralized Private Data via Federated Learning 10 Feb 2024 · 3 repositories · arXiv:2402.06954
-
UrbanKGent: A Unified Large Language Model Agent Framework for Urban Knowledge Graph Construction 10 Feb 2024 · 1 repository · arXiv:2402.06861Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
AI enhanced data assimilation and uncertainty quantification applied to Geological Carbon Storage 9 Feb 2024 · 0 repositories · arXiv:2402.06110
-
Bryndza at ClimateActivism 2024: Stance, Target and Hate Event Detection via Retrieval-Augmented GPT-4 and LLaMA 9 Feb 2024 · 2 repositories · arXiv:2402.06549
-
CultureLLM: Incorporating Cultural Differences into Large Language Models 9 Feb 2024 · 2 repositories · arXiv:2402.10946Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
CurveFormer++: 3D Lane Detection by Curve Propagation with Temporal Curve Queries and Attention 9 Feb 2024 · 0 repositories · arXiv:2402.06423
-
Inducing Systematicity in Transformers by Attending to Structurally Quantized Embeddings 9 Feb 2024 · 1 repository · arXiv:2402.06492
-
Jointly Learning Representations for Map Entities via Heterogeneous Graph Contrastive Learning 9 Feb 2024 · 0 repositories · arXiv:2402.06135
-
LLaVA-Docent: Instruction Tuning with Multimodal Large Language Model to Support Art Appreciation Education 9 Feb 2024 · 0 repositories · arXiv:2402.06264
-
Masked LoGoNet: Fast and Accurate 3D Image Analysis for Medical Domain 9 Feb 2024 · 0 repositories · arXiv:2402.06190
-
RareBench: Can LLMs Serve as Rare Diseases Specialists? 9 Feb 2024 · 1 repository · arXiv:2402.06341
-
ResumeFlow: An LLM-facilitated Pipeline for Personalized Resume Generation and Refinement 9 Feb 2024 · 1 repository · arXiv:2402.06221Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
A Prompt Response to the Demand for Automatic Gender-Neutral Translation 8 Feb 2024 · 1 repository · arXiv:2402.06041
-
Ai4Fapar: How artificial intelligence can help to forecast the seasonal earth observation signal 8 Feb 2024 · 0 repositories · arXiv:2402.06684
-
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers 8 Feb 2024 · 2 repositories · arXiv:2402.05602