Methods › General › Regularization › Label Smoothing › Papers, page 54
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 54 of 144: papers 5,301 to 5,400 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs 22 Feb 2024 · 1 repository · arXiv:2402.14903Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Uncertainty-Aware Evaluation for Vision-Language Models 22 Feb 2024 · 1 repository · arXiv:2402.14418Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 17 harvested samples)
-
Whose LLM is it Anyway? Linguistic Comparison and LLM Attribution for GPT-3.5, GPT-4 and Bard 22 Feb 2024 · 0 repositories · arXiv:2402.14533
-
An Explainable Transformer-based Model for Phishing Email Detection: A Large Language Model Approach 21 Feb 2024 · 0 repositories · arXiv:2402.13871
-
Are LLMs Effective Negotiators? Systematic Evaluation of the Multifaceted Capabilities of LLMs in Negotiation Dialogues 21 Feb 2024 · 1 repository · arXiv:2402.13550Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Beyond A*: Better Planning with Transformers via Search Dynamics Bootstrapping 21 Feb 2024 · 1 repository · arXiv:2402.14083Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Beyond Hate Speech: NLP's Challenges and Opportunities in Uncovering Dehumanizing Language 21 Feb 2024 · 0 repositories · arXiv:2402.13818
-
CriticEval: Evaluating Large Language Model as Critic 21 Feb 2024 · 2 repositories · arXiv:2402.13764
-
Data-driven Discovery with Large Generative Models 21 Feb 2024 · 0 repositories · arXiv:2402.13610
-
Do Efficient Transformers Really Save Computation? 21 Feb 2024 · 0 repositories · arXiv:2402.13934
-
Driving Generative Agents With Their Personality 21 Feb 2024 · 0 repositories · arXiv:2402.14879
-
EffLoc: Lightweight Vision Transformer for Efficient 6-DOF Camera Relocalization 21 Feb 2024 · 0 repositories · arXiv:2402.13537
-
Event-aware Video Corpus Moment Retrieval 21 Feb 2024 · 0 repositories · arXiv:2402.13566
-
Exploring ChatGPT and its Impact on Society 21 Feb 2024 · 0 repositories · arXiv:2403.14643
-
EyeTrans: Merging Human and Machine Attention for Neural Code Summarization 21 Feb 2024 · 1 repository · arXiv:2402.14096
-
FanOutQA: A Multi-Hop, Multi-Document Question Answering Benchmark for Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.14116Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Kuaiji: the First Chinese Accounting Large Language Model 21 Feb 2024 · 0 repositories · arXiv:2402.13866
-
Large Language Models for Data Annotation and Synthesis: A Survey 21 Feb 2024 · 1 repository · arXiv:2402.13446
-
MSTAR: Multi-Scale Backbone Architecture Search for Timeseries Classification 21 Feb 2024 · 0 repositories · arXiv:2402.13822
-
Multi-scale Spatio-temporal Transformer-based Imbalanced Longitudinal Learning for Glaucoma Forecasting from Irregular Time Series Images 21 Feb 2024 · 0 repositories · arXiv:2402.13475
-
OMGEval: An Open Multilingual Generative Evaluation Benchmark for Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.13524
-
AlgoFormer: An Efficient Transformer Framework with Algorithmic Structures 21 Feb 2024 · 0 repositories · arXiv:2402.13572
-
PCA-Bench: Evaluating Multimodal Large Language Models in Perception-Cognition-Action Chain 21 Feb 2024 · 1 repository · arXiv:2402.15527Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
SYNFAC-EDIT: Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization 21 Feb 2024 · 1 repository · arXiv:2402.13919
-
Test-Driven Development for Code Generation 21 Feb 2024 · 0 repositories · arXiv:2402.13521
-
Towards Building Multilingual Language Model for Medicine 21 Feb 2024 · 1 repository · arXiv:2402.13963Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TransGOP: Transformer-Based Gaze Object Prediction 21 Feb 2024 · 1 repository · arXiv:2402.13578Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 3 pointer-only (licence)
-
UniGraph: Learning a Unified Cross-Domain Foundation Model for Text-Attributed Graphs 21 Feb 2024 · 1 repository · arXiv:2402.13630Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
What's in a Name? Auditing Large Language Models for Race and Gender Bias 21 Feb 2024 · 1 repository · arXiv:2402.14875
-
WinoViz: Probing Visual Properties of Objects Under Different States 21 Feb 2024 · 0 repositories · arXiv:2402.13584
-
A Survey on Knowledge Distillation of Large Language Models 20 Feb 2024 · 1 repository · arXiv:2402.13116
-
Advancing GenAI Assisted Programming--A Comparative Study on Prompt Efficiency and Code Quality Between GPT-4 and GLM-4 20 Feb 2024 · 0 repositories · arXiv:2402.12782
-
AgentMD: Empowering Language Agents for Risk Prediction with Large-Scale Clinical Tool Learning 20 Feb 2024 · 0 repositories · arXiv:2402.13225
-
ASCEND: Accurate yet Efficient End-to-End Stochastic Computing Acceleration of Vision Transformer 20 Feb 2024 · 0 repositories · arXiv:2402.12820
-
Can Large Language Models be Used to Provide Psychological Counselling? An Analysis of GPT-4-Generated Responses Using Role-play Dialogues 20 Feb 2024 · 0 repositories · arXiv:2402.12738
-
Cell Graph Transformer for Nuclei Classification 20 Feb 2024 · 1 repository · arXiv:2402.12946Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Conditional Logical Message Passing Transformer for Complex Query Answering 20 Feb 2024 · 1 repository · arXiv:2402.12954
-
An Equivariant Pretrained Transformer for Unified 3D Molecular Representation Learning 20 Feb 2024 · 0 repositories · arXiv:2402.12714
-
HumanEval on Latest GPT Models -- 2024 20 Feb 2024 · 1 repository · arXiv:2402.14852Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
KetGPT -- Dataset Augmentation of Quantum Circuits using Transformers 20 Feb 2024 · 0 repositories · arXiv:2402.13352
-
Me LLaMA: Foundation Large Language Models for Medical Applications 20 Feb 2024 · 1 repository · arXiv:2402.12749Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
OLViT: Multi-Modal State Tracking via Attention-Based Embeddings for Video-Grounded Dialog 20 Feb 2024 · 0 repositories · arXiv:2402.13146
-
OPDAI at SemEval-2024 Task 6: Small LLMs can Accelerate Hallucination Detection with Weakly Supervised Data 20 Feb 2024 · 0 repositories · arXiv:2402.12913
-
PRECISE Framework: GPT-based Text For Improved Readability, Reliability, and Understandability of Radiology Reports For Patient-Centered Care 20 Feb 2024 · 0 repositories · arXiv:2403.00788
-
RhythmFormer: Extracting Patterned rPPG Signals based on Periodic Sparse Attention 20 Feb 2024 · 1 repository · arXiv:2402.12788
-
FinBen: A Holistic Financial Benchmark for Large Language Models 20 Feb 2024 · 2 repositories · arXiv:2402.12659Syntology community repositories only · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 7 pointer-only (licence)
-
The Impact of Demonstrations on Multilingual In-Context Learning: A Multidimensional Analysis 20 Feb 2024 · 1 repository · arXiv:2402.12976
-
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization 20 Feb 2024 · 1 repository · arXiv:2402.13249
-
Tree-Planted Transformers: Unidirectional Transformer Language Models with Implicit Syntactic Supervision 20 Feb 2024 · 0 repositories · arXiv:2402.12691
-
A Critical Evaluation of AI Feedback for Aligning Large Language Models 19 Feb 2024 · 1 repository · arXiv:2402.12366Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
A novel molecule generative model of VAE combined with Transformer for unseen structure generation 19 Feb 2024 · 0 repositories · arXiv:2402.11950
-
ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs 19 Feb 2024 · 1 repository · arXiv:2402.11753Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Asynchronous and Segmented Bidirectional Encoding for NMT 19 Feb 2024 · 0 repositories · arXiv:2402.14849
-
Creating a Fine Grained Entity Type Taxonomy Using LLMs 19 Feb 2024 · 0 repositories · arXiv:2402.12557
-
DeepCode AI Fix: Fixing Security Vulnerabilities with Large Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.13291
-
Surprising Efficacy of Fine-Tuned Transformers for Fact-Checking over Larger Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.12147
-
Evaluation of ChatGPT's Smart Contract Auditing Capabilities Based on Chain of Thought 19 Feb 2024 · 0 repositories · arXiv:2402.12023
-
FiT: Flexible Vision Transformer for Diffusion Model 19 Feb 2024 · 2 repositories · arXiv:2402.12376Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 17 harvested samples) · 3 pointer-only (licence)
-
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations 19 Feb 2024 · 2 repositories · arXiv:2402.12348
-
IMBUE: Improving Interpersonal Effectiveness through Simulation and Just-in-time Feedback with Human-Language Model Interaction 19 Feb 2024 · 0 repositories · arXiv:2402.12556
-
Is Open-Source There Yet? A Comparative Study on Commercial and Open-Source LLMs in Their Ability to Label Chest X-Ray Reports 19 Feb 2024 · 0 repositories · arXiv:2402.12298
-
Locality-Sensitive Hashing-Based Efficient Point Transformer with Applications in High-Energy Physics 19 Feb 2024 · 1 repository · arXiv:2402.12535Syntology official (archive's flag): 12 ran · 12 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 11 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples)
-
Enabling Weak LLMs to Judge Response Reliability via Meta Ranking 19 Feb 2024 · 0 repositories · arXiv:2402.12146
-
Cofca: A Step-Wise Counterfactual Multi-hop QA benchmark 19 Feb 2024 · 0 repositories · arXiv:2402.11924
-
Perceiving Longer Sequences With Bi-Directional Cross-Attention Transformers 19 Feb 2024 · 2 repositories · arXiv:2402.12138Syntology official (archive's flag): 16 ran · 17 ran (of which 11 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 6 where Syntology's instrument failed) · 11 unverified (of 28 harvested samples) · 27 pointer-only (licence)
-
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models 19 Feb 2024 · 1 repository · arXiv:2402.12336Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Shallow Synthesis of Knowledge in GPT-Generated Texts: A Case Study in Automatic Related Work Composition 19 Feb 2024 · 0 repositories · arXiv:2402.12255
-
SPML: A DSL for Defending Language Models Against Prompt Attacks 19 Feb 2024 · 0 repositories · arXiv:2402.11755
-
Standardize: Aligning Language Models with Expert-Defined Standards for Content Generation 19 Feb 2024 · 1 repository · arXiv:2402.12593
-
Stealing the Invisible: Unveiling Pre-Trained CNN Models through Adversarial Examples and Timing Side-Channels 19 Feb 2024 · 0 repositories · arXiv:2402.11953
-
CiMNet: Towards Joint Optimization for DNN Architecture and Configuration for Compute-In-Memory Hardware 19 Feb 2024 · 0 repositories · arXiv:2402.11780
-
Your Large Language Model is Secretly a Fairness Proponent and You Should Prompt it Like One 19 Feb 2024 · 0 repositories · arXiv:2402.12150
-
Can Deception Detection Go Deeper? Dataset, Evaluation, and Benchmark for Deception Reasoning 18 Feb 2024 · 0 repositories · arXiv:2402.11432
-
Decoding News Narratives: A Critical Analysis of Large Language Models in Framing Detection 18 Feb 2024 · 0 repositories · arXiv:2402.11621
-
DictLLM: Harnessing Key-Value Data Structures with Large Language Models for Enhanced Medical Diagnostics 18 Feb 2024 · 0 repositories · arXiv:2402.11481
-
EventRL: Enhancing Event Extraction with Outcome Supervision for Large Language Models 18 Feb 2024 · 1 repository · arXiv:2402.11430
-
FactPICO: Factuality Evaluation for Plain Language Summarization of Medical Evidence 18 Feb 2024 · 1 repository · arXiv:2402.11456Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
KMMLU: Measuring Massive Multitask Language Understanding in Korean 18 Feb 2024 · 0 repositories · arXiv:2402.11548
-
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents 18 Feb 2024 · 1 repository · arXiv:2402.11651Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
LongAgent: Scaling Language Models to 128k Context through Multi-Agent Collaboration 18 Feb 2024 · 1 repository · arXiv:2402.11550
-
MSynFD: Multi-hop Syntax aware Fake News Detection 18 Feb 2024 · 0 repositories · arXiv:2402.14834
-
Multi-dimensional Evaluation of Empathetic Dialog Responses 18 Feb 2024 · 0 repositories · arXiv:2402.11409
-
Multi-Task Inference: Can Large Language Models Follow Multiple Instructions at Once? 18 Feb 2024 · 1 repository · arXiv:2402.11597Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Ploutos: Towards interpretable stock movement prediction with financial large language model 18 Feb 2024 · 0 repositories · arXiv:2403.00782
-
Vision-Flan: Scaling Human-Labeled Tasks in Visual Instruction Tuning 18 Feb 2024 · 0 repositories · arXiv:2402.11690
-
Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models 17 Feb 2024 · 2 repositories · arXiv:2402.11140
-
Detecting a Proxy for Potential Comorbid ADHD in People Reporting Anxiety Symptoms from Social Media Data 17 Feb 2024 · 0 repositories · arXiv:2403.05561
-
Exploring ChatGPT for Next-generation Information Retrieval: Opportunities and Challenges 17 Feb 2024 · 0 repositories · arXiv:2402.11203
-
GenDec: A robust generative Question-decomposition method for Multi-hop reasoning 17 Feb 2024 · 0 repositories · arXiv:2402.11166
-
Reasoning before Comparison: LLM-Enhanced Semantic Similarity Metrics for Domain Specialized Text Analysis 17 Feb 2024 · 0 repositories · arXiv:2402.11398
-
ReViT: Enhancing Vision Transformers Feature Diversity with Attention Residual Connections 17 Feb 2024 · 1 repository · arXiv:2402.11301Syntology official (archive's flag): 22 ran · 22 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 24 harvested samples) · 1 pointer-only (licence)
-
ZeroG: Investigating Cross-dataset Zero-shot Transferability in Graphs 17 Feb 2024 · 1 repository · arXiv:2402.11235Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Assessing the Reasoning Abilities of ChatGPT in the Context of Claim Verification 16 Feb 2024 · 0 repositories · arXiv:2402.10735
-
Can LLMs Speak For Diverse People? Tuning LLMs via Debate to Generate Controllable Controversial Statements 16 Feb 2024 · 1 repository · arXiv:2402.10614Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can Separators Improve Chain-of-Thought Prompting? 16 Feb 2024 · 0 repositories · arXiv:2402.10645
-
ContiFormer: Continuous-Time Transformer for Irregular Time Series Modeling 16 Feb 2024 · 1 repository · arXiv:2402.10635Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples)
-
Dynamic Patch-aware Enrichment Transformer for Occluded Person Re-Identification 16 Feb 2024 · 0 repositories · arXiv:2402.10435
-
Emoji Driven Crypto Assets Market Reactions 16 Feb 2024 · 0 repositories · arXiv:2402.10481
-
FinTral: A Family of GPT-4 Level Multimodal Financial Large Language Models 16 Feb 2024 · 0 repositories · arXiv:2402.10986
-
German Text Simplification: Finetuning Large Language Models with Semi-Synthetic Data 16 Feb 2024 · 1 repository · arXiv:2402.10675