Methods › General › Regularization › Label Smoothing › Papers, page 28
Label Smoothing
Papers archive 2025-07-28
archive papers tagged: 14,327 · with a code link: 6,651 · where Syntology ran a sample: 2,259 (1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,259 of 14,327 tagged: 1,920 with a run with no instrument failure, 339 where every run was a failure of Syntology's instrument)
Page 28 of 144: papers 2,701 to 2,800 of 14,327, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Enhancing TinyBERT for Financial Sentiment Analysis Using GPT-Augmented FinBERT Distillation 19 Sep 2024 · 1 repository · arXiv:2409.18999
-
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering 19 Sep 2024 · 1 repository · arXiv:2409.12784
-
LMT-Net: Lane Model Transformer Network for Automated HD Mapping from Sparse Vehicle Observations 19 Sep 2024 · 0 repositories · arXiv:2409.12409
-
Prompts Are Programs Too! Understanding How Developers Build Software Containing Prompts 19 Sep 2024 · 0 repositories · arXiv:2409.12447
-
TACO-RL: Task Aware Prompt Compression Optimization with Reinforcement Learning 19 Sep 2024 · 0 repositories · arXiv:2409.13035
-
What Would You Ask When You First Saw a²+b²=c²? Evaluating LLM on Curiosity-Driven Questioning 19 Sep 2024 · 0 repositories · arXiv:2409.17172
-
Data Efficient Acoustic Scene Classification using Teacher-Informed Confusing Class Instruction 18 Sep 2024 · 0 repositories · arXiv:2409.11964
-
DPI-TTS: Directional Patch Interaction for Fast-Converging and Style Temporal Modeling in Text-to-Speech 18 Sep 2024 · 0 repositories · arXiv:2409.11835
-
DynaMo: In-Domain Dynamics Pretraining for Visuo-Motor Control 18 Sep 2024 · 0 repositories · arXiv:2409.12192
-
From Lists to Emojis: How Format Bias Affects Model Alignment 18 Sep 2024 · 0 repositories · arXiv:2409.11704
-
Harnessing LLMs for API Interactions: A Framework for Classification and Synthetic Data Generation 18 Sep 2024 · 0 repositories · arXiv:2409.11703
-
NT-ViT: Neural Transcoding Vision Transformers for EEG-to-fMRI Synthesis 18 Sep 2024 · 0 repositories · arXiv:2409.11836
-
Reinforcement Learning as an Improvement Heuristic for Real-World Production Scheduling 18 Sep 2024 · 0 repositories · arXiv:2409.11933
-
Unsupervised Feature Orthogonalization for Learning Distortion-Invariant Representations 18 Sep 2024 · 1 repository · arXiv:2409.12276
-
American Sign Language to Text Translation using Transformer and Seq2Seq with LSTM 17 Sep 2024 · 0 repositories · arXiv:2409.10874
-
Contrasformer: A Brain Network Contrastive Transformer for Neurodegenerative Condition Identification 17 Sep 2024 · 1 repository · arXiv:2409.10944
-
EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer 17 Sep 2024 · 0 repositories · arXiv:2409.10819
-
A Hybrid Multi-Factor Network with Dynamic Sequence Modeling for Early Warning of Intraoperative Hypotension 17 Sep 2024 · 1 repository · arXiv:2409.11064
-
Implicit Reasoning in Deep Time Series Forecasting 17 Sep 2024 · 0 repositories · arXiv:2409.10840
-
Investigating Context-Faithfulness in Large Language Models: The Roles of Memory Strength and Evidence Style 17 Sep 2024 · 0 repositories · arXiv:2409.10955
-
Linear Recency Bias During Training Improves Transformers' Fit to Reading Times 17 Sep 2024 · 0 repositories · arXiv:2409.11250
-
LOLA -- An Open-Source Massively Multilingual Large Language Model 17 Sep 2024 · 1 repository · arXiv:2409.11272
-
Multimodality Adaptive Transformer and Mutual Learning for Unsupervised Domain Adaptation Vehicle Re-Identification 17 Sep 2024 · 0 repositories
-
Norm of Mean Contextualized Embeddings Determines their Variance 17 Sep 2024 · 1 repository · arXiv:2409.11253
-
Semformer: Transformer Language Models with Semantic Planning 17 Sep 2024 · 0 repositories · arXiv:2409.11143
-
SkinMamba: A Precision Skin Lesion Segmentation Architecture with Cross-Scale Global State Modeling and Frequency Boundary Guidance 17 Sep 2024 · 1 repository · arXiv:2409.10890
-
Sparks of Artificial General Intelligence(AGI) in Semiconductor Material Science: Early Explorations into the Next Frontier of Generative AI-Assisted Electron Micrograph Analysis 17 Sep 2024 · 0 repositories · arXiv:2409.12244
-
Unleashing the Potential of Mamba: Boosting a LiDAR 3D Sparse Detector by Using Cross-Model Knowledge Distillation 17 Sep 2024 · 0 repositories · arXiv:2409.11018
-
Are Deep Learning Models Robust to Partial Object Occlusion in Visual Recognition Tasks? 16 Sep 2024 · 0 repositories · arXiv:2409.10775
-
beeFormer: Bridging the Gap Between Semantic and Interaction Similarity in Recommender Systems 16 Sep 2024 · 1 repository · arXiv:2409.10309
-
Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles 16 Sep 2024 · 1 repository · arXiv:2409.10502Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 11 harvested samples)
-
Flash STU: Fast Spectral Transform Units 16 Sep 2024 · 1 repository · arXiv:2409.10489
-
Garment Attribute Manipulation with Multi-level Attention 16 Sep 2024 · 0 repositories · arXiv:2409.10206
-
GPT takes the SAT: Tracing changes in Test Difficulty and Math Performance of Students 16 Sep 2024 · 0 repositories · arXiv:2409.10750
-
Kolmogorov-Arnold Transformer 16 Sep 2024 · 1 repository · arXiv:2409.10594Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
LLMs for clinical risk prediction 16 Sep 2024 · 0 repositories · arXiv:2409.10191
-
MindGuard: Towards Accessible and Sitgma-free Mental Health First Aid via Edge LLM 16 Sep 2024 · 0 repositories · arXiv:2409.10064
-
Mitigating Partial Observability in Adaptive Traffic Signal Control with Transformers 16 Sep 2024 · 0 repositories · arXiv:2409.10693
-
NARX Transformer: A Dynamic Model for Leveraging Multicycle Data in Long-Term Battery State of Health Estimation 16 Sep 2024 · 2 repositories
-
Recurrent Graph Transformer Network for Multiple Fault Localization in Naval Shipboard Systems 16 Sep 2024 · 0 repositories · arXiv:2409.10792
-
SelECT-SQL: Self-correcting ensemble Chain-of-Thought for Text-to-SQL 16 Sep 2024 · 1 repository · arXiv:2409.10007
-
Detection Made Easy: Potentials of Large Language Models for Solidity Vulnerabilities 15 Sep 2024 · 0 repositories · arXiv:2409.10574
-
Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion 15 Sep 2024 · 1 repository · arXiv:2409.09808Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
GP-GPT: Large Language Model for Gene-Phenotype Mapping 15 Sep 2024 · 0 repositories · arXiv:2409.09825
-
Latent Diffusion Models for Controllable RNA Sequence Generation 15 Sep 2024 · 0 repositories · arXiv:2409.09828
-
Leveraging Open-Source Large Language Models for Native Language Identification 15 Sep 2024 · 0 repositories · arXiv:2409.09659
-
Predicting building types and functions at transnational scale 15 Sep 2024 · 0 repositories · arXiv:2409.09692
-
SITSMamba for Crop Classification based on Satellite Image Time Series 15 Sep 2024 · 1 repository · arXiv:2409.09673
-
Underwater Image Enhancement via Dehazing and Color Restoration 15 Sep 2024 · 0 repositories · arXiv:2409.09779
-
Unveiling Gender Bias in Large Language Models: Using Teacher's Evaluation in Higher Education As an Example 15 Sep 2024 · 1 repository · arXiv:2409.09652
-
An empirical evaluation of using ChatGPT to summarize disputes for recommending similar labor and employment cases in Chinese 14 Sep 2024 · 0 repositories · arXiv:2409.09280
-
Autoregressive + Chain of Thought = Recurrent: Recurrence's Role in Language Models' Computability and a Revisit of Recurrent Transformer 14 Sep 2024 · 0 repositories · arXiv:2409.09239
-
Investigation of Hierarchical Spectral Vision Transformer Architecture for Classification of Hyperspectral Imagery 14 Sep 2024 · 0 repositories · arXiv:2409.09244
-
Keeping Humans in the Loop: Human-Centered Automated Annotation with Generative AI 14 Sep 2024 · 0 repositories · arXiv:2409.09467
-
Multi-Microphone and Multi-Modal Emotion Recognition in Reverberant Environment 14 Sep 2024 · 0 repositories · arXiv:2409.09545
-
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens 14 Sep 2024 · 0 repositories · arXiv:2409.09513
-
SEA-ViT: Sea Surface Currents Forecasting Using Vision Transformer and GRU-Based Spatio-Temporal Covariance Modeling 14 Sep 2024 · 1 repository · arXiv:2409.16313
-
Tran-GCN: A Transformer-Enhanced Graph Convolutional Network for Person Re-Identification in Monitoring Videos 14 Sep 2024 · 0 repositories · arXiv:2409.09391
-
VSFormer: Mining Correlations in Flexible View Set for Multi-view 3D Shape Understanding 14 Sep 2024 · 1 repository · arXiv:2409.09254
-
A RAG Approach for Generating Competency Questions in Ontology Engineering 13 Sep 2024 · 0 repositories · arXiv:2409.08820
-
ChangeChat: An Interactive Model for Remote Sensing Change Analysis via Multimodal Instruction Tuning 13 Sep 2024 · 1 repository · arXiv:2409.08582Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
HTR-VT: Handwritten Text Recognition with Vision Transformer 13 Sep 2024 · 2 repositories · arXiv:2409.08573
-
Integration of Mamba and Transformer -- MAT for Long-Short Range Time Series Forecasting with Application to Weather Dynamics 13 Sep 2024 · 0 repositories · arXiv:2409.08530
-
KodeXv0.1: A Family of State-of-the-Art Financial Large Language Models 13 Sep 2024 · 0 repositories · arXiv:2409.13749
-
PSTNet: Enhanced Polyp Segmentation with Multi-scale Alignment and Frequency Domain Integration 13 Sep 2024 · 0 repositories · arXiv:2409.08501
-
SkinFormer: Learning Statistical Texture Representation with Transformer for Skin Lesion Segmentation 13 Sep 2024 · 1 repository · arXiv:2409.08652
-
TabKANet: Tabular Data Modeling with Kolmogorov-Arnold Network and Transformer 13 Sep 2024 · 2 repositories · arXiv:2409.08806
-
Transformer with Controlled Attention for Synchronous Motion Captioning 13 Sep 2024 · 1 repository · arXiv:2409.09177
-
xTED: Cross-Domain Adaptation via Diffusion-Based Trajectory Editing 13 Sep 2024 · 1 repository · arXiv:2409.08687
-
AD-Lite Net: A Lightweight and Concatenated CNN Model for Alzheimer's Detection from MRI Images 12 Sep 2024 · 0 repositories · arXiv:2409.08170
-
Depth Matters: Exploring Deep Interactions of RGB-D for Semantic Segmentation in Traffic Scenes 12 Sep 2024 · 0 repositories · arXiv:2409.07995
-
Fine-tuning Large Language Models for Entity Matching 12 Sep 2024 · 1 repository · arXiv:2409.08185
-
Generated Data with Fake Privacy: Hidden Dangers of Fine-tuning Large Language Models on Generated Data 12 Sep 2024 · 0 repositories · arXiv:2409.11423
-
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers 12 Sep 2024 · 0 repositories · arXiv:2410.05273
-
Lagrange Duality and Compound Multi-Attention Transformer for Semi-Supervised Medical Image Segmentation 12 Sep 2024 · 1 repository · arXiv:2409.07793
-
Online vs Offline: A Comparative Study of First-Party and Third-Party Evaluations of Social Chatbots 12 Sep 2024 · 0 repositories · arXiv:2409.07823
-
Q-value Regularized Decision ConvFormer for Offline Reinforcement Learning 12 Sep 2024 · 0 repositories · arXiv:2409.08062
-
SDformer: Efficient End-to-End Transformer for Depth Completion 12 Sep 2024 · 1 repository · arXiv:2409.08159
-
SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer 12 Sep 2024 · 1 repository · arXiv:2409.08425
-
How Effectively Do LLMs Extract Feature-Sentiment Pairs from App Reviews? 11 Sep 2024 · 1 repository · arXiv:2409.07162
-
ART: Artifact Removal Transformer for Reconstructing Noise-Free Multichannel Electroencephalographic Signals 11 Sep 2024 · 0 repositories · arXiv:2409.07326
-
Attention Down-Sampling Transformer, Relative Ranking and Self-Consistency for Blind Image Quality Assessment 11 Sep 2024 · 1 repository · arXiv:2409.07115
-
Can We Count on LLMs? The Fixed-Effect Fallacy and Claims of GPT-4 Capabilities 11 Sep 2024 · 0 repositories · arXiv:2409.07638
-
CWT-Net: Super-resolution of Histopathology Images Using a Cross-scale Wavelet-based Transformer 11 Sep 2024 · 0 repositories · arXiv:2409.07092
-
Foundation Models Boost Low-Level Perceptual Similarity Metrics 11 Sep 2024 · 1 repository · arXiv:2409.07650
-
Intrapartum Ultrasound Image Segmentation of Pubic Symphysis and Fetal Head Using Dual Student-Teacher Framework with CNN-ViT Collaborative Learning 11 Sep 2024 · 1 repository · arXiv:2409.06928
-
Mamba for Scalable and Efficient Personalized Recommendations 11 Sep 2024 · 0 repositories · arXiv:2409.17165
-
Multimodal Emotion Recognition with Vision-language Prompting and Modality Dropout 11 Sep 2024 · 0 repositories · arXiv:2409.07078
-
SimulBench: Evaluating Language Models with Creative Simulation Tasks 11 Sep 2024 · 0 repositories · arXiv:2409.07641
-
SSR-Speech: Towards Stable, Safe and Robust Zero-shot Text-based Speech Editing and Synthesis 11 Sep 2024 · 1 repository · arXiv:2409.07556Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
Swin-LiteMedSAM: A Lightweight Box-Based Segment Anything Model for Large-Scale Medical Image Datasets 11 Sep 2024 · 1 repository · arXiv:2409.07172Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Token Turing Machines are Efficient Vision Models 11 Sep 2024 · 1 repository · arXiv:2409.07613Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
VMAS: Video-to-Music Generation via Semantic Alignment in Web Music Videos 11 Sep 2024 · 0 repositories · arXiv:2409.07450
-
Weather-Informed Probabilistic Forecasting and Scenario Generation in Power Systems 11 Sep 2024 · 0 repositories · arXiv:2409.07637
-
Mapping Biomedical Ontology Terms to IDs: Effect of Domain Prevalence on Prediction Accuracy 11 Sep 2024 · 0 repositories · arXiv:2409.13746
-
A Dataset for Evaluating LLM-based Evaluation Functions for Research Question Extraction Task 10 Sep 2024 · 0 repositories · arXiv:2409.06883
-
A Practical Gated Recurrent Transformer Network Incorporating Multiple Fusions for Video Denoising 10 Sep 2024 · 0 repositories · arXiv:2409.06603
-
Adaptive Transformer Modelling of Density Function for Nonparametric Survival Analysis 10 Sep 2024 · 1 repository · arXiv:2409.06209Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
AgileIR: Memory-Efficient Group Shifted Windows Attention for Agile Image Restoration 10 Sep 2024 · 0 repositories · arXiv:2409.06206
-
Can Large Language Models Unlock Novel Scientific Research Ideas? 10 Sep 2024 · 1 repository · arXiv:2409.06185Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)