Methods › General › Regularization › Weight Decay › Papers, page 34
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 34 of 108: papers 3,301 to 3,400 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
WIKIGENBENCH: Exploring Full-length Wikipedia Generation under Real-World Scenario 28 Feb 2024 · 1 repository · arXiv:2402.18264
-
Unsupervised Information Refinement Training of Large Language Models for Retrieval-Augmented Generation 28 Feb 2024 · 1 repository · arXiv:2402.18150Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
A Language Model based Framework for New Concept Placement in Ontologies 27 Feb 2024 · 1 repository · arXiv:2402.17897
-
COCOA: CBT-based Conversational Counseling Agent using Memory Specialized in Cognitive Distortions and Dynamic Prompt 27 Feb 2024 · 0 repositories · arXiv:2402.17546
-
Deep Learning Detection Method for Large Language Models-Generated Scientific Content 27 Feb 2024 · 0 repositories · arXiv:2403.00828
-
Emotional Voice Messages (EMOVOME) database: emotion recognition in spontaneous voice messages 27 Feb 2024 · 0 repositories · arXiv:2402.17496
-
Evaluating Very Long-Term Conversational Memory of LLM Agents 27 Feb 2024 · 1 repository · arXiv:2402.17753Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems 27 Feb 2024 · 1 repository · arXiv:2402.17840Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
JMLR: Joint Medical LLM and Retrieval Training for Enhancing Reasoning and Professional Question Answering Capability 27 Feb 2024 · 1 repository · arXiv:2402.17887
-
Measuring Vision-Language STEM Skills of Neural Models 27 Feb 2024 · 1 repository · arXiv:2402.17205Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Model Free Deep Deterministic Policy Gradient Controller for Setpoint Tracking of Non-minimum Phase Systems 27 Feb 2024 · 0 repositories · arXiv:2402.17703
-
MAGPIE: Multi-Task Media-Bias Analysis Generalization for Pre-Trained Identification of Expressions 27 Feb 2024 · 1 repository · arXiv:2403.07910
-
REAR: A Relevance-Aware Retrieval-Augmented Framework for Open-Domain Question Answering 27 Feb 2024 · 1 repository · arXiv:2402.17497Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Variational Learning is Effective for Large Deep Networks 27 Feb 2024 · 1 repository · arXiv:2402.17641Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Adaptation of Biomedical and Clinical Pretrained Models to French Long Documents: A Comparative Study 26 Feb 2024 · 1 repository · arXiv:2402.16689
-
An Integrated Data Processing Framework for Pretraining Foundation Models 26 Feb 2024 · 2 repositories · arXiv:2402.16358
-
Asymmetry in Low-Rank Adapters of Foundation Models 26 Feb 2024 · 1 repository · arXiv:2402.16842Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
From RAGs to riches: Using large language models to write documents for clinical trials 26 Feb 2024 · 0 repositories · arXiv:2402.16406
-
Retrieval Augmented Generation Systems: Automatic Dataset Creation, Evaluation and Boolean Agent Setup 26 Feb 2024 · 1 repository · arXiv:2403.00820
-
ChatMusician: Understanding and Generating Music Intrinsically with LLM 25 Feb 2024 · 1 repository · arXiv:2402.16153Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Deep Learning Approaches for Improving Question Answering Systems in Hepatocellular Carcinoma Research 25 Feb 2024 · 0 repositories · arXiv:2402.16038
-
Emotion Classification in Short English Texts using Deep Learning Techniques 25 Feb 2024 · 0 repositories · arXiv:2402.16034
-
Knowledge Fusion of Chat LLMs: A Preliminary Technical Report 25 Feb 2024 · 2 repositories · arXiv:2402.16107
-
Hitting "Probe"rty with Non-Linearity, and More 25 Feb 2024 · 0 repositories · arXiv:2402.16168
-
LLMs Can Defend Themselves Against Jailbreaking in a Practical Manner: A Vision Paper 24 Feb 2024 · 0 repositories · arXiv:2402.15727
-
Look Before You Leap: Problem Elaboration Prompting Improves Mathematical Reasoning in Large Language Models 24 Feb 2024 · 0 repositories · arXiv:2402.15764
-
MultiContrievers: Analysis of Dense Retrieval Representations 24 Feb 2024 · 1 repository · arXiv:2402.15925
-
PRP: Propagating Universal Perturbations to Attack Large Language Model Guard-Rails 24 Feb 2024 · 0 repositories · arXiv:2402.15911
-
SemEval-2024 Task 8: Weighted Layer Averaging RoBERTa for Black-Box Machine-Generated Text Detection 24 Feb 2024 · 1 repository · arXiv:2402.15873
-
A First Look at GPT Apps: Landscape and Vulnerability 23 Feb 2024 · 0 repositories · arXiv:2402.15105
-
Advancing Parameter Efficiency in Fine-tuning via Representation Editing 23 Feb 2024 · 2 repositories · arXiv:2402.15179Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
AttributionBench: How Hard is Automatic Attribution Evaluation? 23 Feb 2024 · 1 repository · arXiv:2402.15089
-
Dual Encoder: Exploiting the Potential of Syntactic and Semantic for Aspect Sentiment Triplet Extraction 23 Feb 2024 · 0 repositories · arXiv:2402.15370
-
Evaluating the Performance of ChatGPT for Spam Email Detection 23 Feb 2024 · 0 repositories · arXiv:2402.15537
-
Fine-tuning Large Language Models for Domain-specific Machine Translation 23 Feb 2024 · 0 repositories · arXiv:2402.15061
-
LLMs as Meta-Reviewers' Assistants: A Case Study 23 Feb 2024 · 1 repository · arXiv:2402.15589
-
The Good and The Bad: Exploring Privacy Issues in Retrieval-Augmented Generation (RAG) 23 Feb 2024 · 1 repository · arXiv:2402.16893Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Towards Efficient Active Learning in NLP via Pretrained Representations 23 Feb 2024 · 0 repositories · arXiv:2402.15613
-
2D Matryoshka Sentence Embeddings 22 Feb 2024 · 1 repository · arXiv:2402.14776
-
Assessing generalization capability of text ranking models in Polish 22 Feb 2024 · 0 repositories · arXiv:2402.14318
-
Can Large Language Models Detect Misinformation in Scientific News Reporting? 22 Feb 2024 · 0 repositories · arXiv:2402.14268
-
Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming 22 Feb 2024 · 0 repositories · arXiv:2402.14261
-
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge 22 Feb 2024 · 1 repository · arXiv:2402.14310
-
KoCoSa: Korean Context-aware Sarcasm Detection Dataset 22 Feb 2024 · 1 repository · arXiv:2402.14428
-
RoboScript: Code Generation for Free-Form Manipulation Tasks across Real and Simulation 22 Feb 2024 · 0 repositories · arXiv:2402.14623
-
Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs 22 Feb 2024 · 1 repository · arXiv:2402.14903Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Towards Understanding Counseling Conversations: Domain Knowledge and Large Language Models 22 Feb 2024 · 0 repositories · arXiv:2402.14200
-
Transferring BERT Capabilities from High-Resource to Low-Resource Languages Using Vocabulary Matching 22 Feb 2024 · 0 repositories · arXiv:2402.14408
-
ActiveRAG: Autonomously Knowledge Assimilation and Accommodation through Retrieval-Augmented Agents 21 Feb 2024 · 1 repository · arXiv:2402.13547
-
An Evaluation of Large Language Models in Bioinformatics Research 21 Feb 2024 · 0 repositories · arXiv:2402.13714
-
An Explainable Transformer-based Model for Phishing Email Detection: A Large Language Model Approach 21 Feb 2024 · 0 repositories · arXiv:2402.13871
-
Beyond Hate Speech: NLP's Challenges and Opportunities in Uncovering Dehumanizing Language 21 Feb 2024 · 0 repositories · arXiv:2402.13818
-
Do Efficient Transformers Really Save Computation? 21 Feb 2024 · 0 repositories · arXiv:2402.13934
-
Green AI: A Preliminary Empirical Study on Energy Consumption in DL Models Across Different Runtime Infrastructures 21 Feb 2024 · 0 repositories · arXiv:2402.13640
-
Hallucinations or Attention Misdirection? The Path to Strategic Value Extraction in Business Using Large Language Models 21 Feb 2024 · 0 repositories · arXiv:2402.14002
-
Improving Language Understanding from Screenshots 21 Feb 2024 · 1 repository · arXiv:2402.14073Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Knowledge Graph Enhanced Large Language Model Editing 21 Feb 2024 · 0 repositories · arXiv:2402.13593
-
A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.13457Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SYNFAC-EDIT: Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization 21 Feb 2024 · 1 repository · arXiv:2402.13919
-
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity 20 Feb 2024 · 1 repository · arXiv:2402.13130
-
Benchmarking Retrieval-Augmented Generation for Medicine 20 Feb 2024 · 2 repositories · arXiv:2402.13178Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Can GNN be Good Adapter for LLMs? 20 Feb 2024 · 2 repositories · arXiv:2402.12984Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
ChatEL: Entity Linking with Chatbots 20 Feb 2024 · 1 repository · arXiv:2402.14858
-
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries 20 Feb 2024 · 0 repositories · arXiv:2402.13372
-
Exploring the Impact of Table-to-Text Methods on Augmenting LLM-based Question Answering with Domain Hybrid Data 20 Feb 2024 · 0 repositories · arXiv:2402.12869
-
Is the System Message Really Important to Jailbreaks in Large Language Models? 20 Feb 2024 · 0 repositories · arXiv:2402.14857
-
MoELoRA: Contrastive Learning Guided Mixture of Experts on Parameter-Efficient Fine-Tuning for Large Language Models 20 Feb 2024 · 1 repository · arXiv:2402.12851
-
NL2Formula: Generating Spreadsheet Formulas from Natural Language Queries 20 Feb 2024 · 0 repositories · arXiv:2402.14853
-
PromptKD: Distilling Student-Friendly Knowledge for Generative Language Models via Prompt Tuning 20 Feb 2024 · 1 repository · arXiv:2402.12842Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Reflect-RL: Two-Player Online RL Fine-Tuning for LMs 20 Feb 2024 · 1 repository · arXiv:2402.12621Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
R³: "This is My SQL, Are You With Me?" A Consensus-Based Multi-Agent System for Text-to-SQL Tasks 20 Feb 2024 · 0 repositories · arXiv:2402.14851
-
The Impact of Demonstrations on Multilingual In-Context Learning: A Multidimensional Analysis 20 Feb 2024 · 1 repository · arXiv:2402.12976
-
A Critical Evaluation of AI Feedback for Aligning Large Language Models 19 Feb 2024 · 1 repository · arXiv:2402.12366Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Acquiring Clean Language Models from Backdoor Poisoned Datasets by Downscaling Frequency Space 19 Feb 2024 · 1 repository · arXiv:2402.12026Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AnaloBench: Benchmarking the Identification of Abstract and Long-context Analogies 19 Feb 2024 · 2 repositories · arXiv:2402.12370Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Ask Optimal Questions: Aligning Large Language Models with Retriever's Preference in Conversational Search 19 Feb 2024 · 0 repositories · arXiv:2402.11827
-
CodeArt: Better Code Models by Attention Regularization When Symbols Are Lacking 19 Feb 2024 · 1 repository · arXiv:2402.11842
-
DeepCode AI Fix: Fixing Security Vulnerabilities with Large Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.13291
-
Surprising Efficacy of Fine-Tuned Transformers for Fact-Checking over Larger Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.12147
-
FeB4RAG: Evaluating Federated Search in the Context of Retrieval Augmented Generation 19 Feb 2024 · 0 repositories · arXiv:2402.11891
-
Graph-Based Retriever Captures the Long Tail of Biomedical Knowledge 19 Feb 2024 · 0 repositories · arXiv:2402.12352
-
Head-wise Shareable Attention for Large Language Models 19 Feb 2024 · 2 repositories · arXiv:2402.11819
-
Is Open-Source There Yet? A Comparative Study on Commercial and Open-Source LLMs in Their Ability to Label Chest X-Ray Reports 19 Feb 2024 · 0 repositories · arXiv:2402.12298
-
KARL: Knowledge-Aware Retrieval and Representations aid Retention and Learning in Students 19 Feb 2024 · 0 repositories · arXiv:2402.12291
-
Language Model Adaptation to Specialized Domains through Selective Masking based on Genre and Topical Characteristics 19 Feb 2024 · 1 repository · arXiv:2402.12036
-
Mafin: Enhancing Black-Box Embeddings with Model Augmented Fine-Tuning 19 Feb 2024 · 0 repositories · arXiv:2402.12177
-
Enabling Weak LLMs to Judge Response Reliability via Meta Ranking 19 Feb 2024 · 0 repositories · arXiv:2402.12146
-
Ontology Enhanced Claim Detection 19 Feb 2024 · 0 repositories · arXiv:2402.12282
-
Query-Based Adversarial Prompt Generation 19 Feb 2024 · 2 repositories · arXiv:2402.12329
-
SPML: A DSL for Defending Language Models Against Prompt Attacks 19 Feb 2024 · 0 repositories · arXiv:2402.11755
-
Stick to your Role! Stability of Personal Values Expressed in Large Language Models 19 Feb 2024 · 0 repositories · arXiv:2402.14846
-
What Evidence Do Language Models Find Convincing? 19 Feb 2024 · 1 repository · arXiv:2402.11782Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 14 harvested samples)
-
Your Large Language Model is Secretly a Fairness Proponent and You Should Prompt it Like One 19 Feb 2024 · 0 repositories · arXiv:2402.12150
-
A Curious Case of Searching for the Correlation between Training Data and Adversarial Robustness of Transformer Textual Models 18 Feb 2024 · 1 repository · arXiv:2402.11469
-
Combinatorial Client-Master Multiagent Deep Reinforcement Learning for Task Offloading in Mobile Edge Computing 18 Feb 2024 · 2 repositories · arXiv:2402.11653
-
Decoding News Narratives: A Critical Analysis of Large Language Models in Framing Detection 18 Feb 2024 · 0 repositories · arXiv:2402.11621
-
GNNavi: Navigating the Information Flow in Large Language Models by Graph Neural Network 18 Feb 2024 · 1 repository · arXiv:2402.11709
-
Metric-Learning Encoding Models Identify Processing Profiles of Linguistic Features in BERT's Representations 18 Feb 2024 · 1 repository · arXiv:2402.11608
-
Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement 18 Feb 2024 · 1 repository · arXiv:2402.11436Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Utilizing BERT for Information Retrieval: Survey, Applications, Resources, and Challenges 18 Feb 2024 · 0 repositories · arXiv:2403.00784