Methods › General › Regularization › Weight Decay › Papers, page 28
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 28 of 108: papers 2,701 to 2,800 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Are PPO-ed Language Models Hackable? 28 May 2024 · 0 repositories · arXiv:2406.02577
-
ATM: Adversarial Tuning Multi-agent System Makes a Robust Retrieval-Augmented Generator 28 May 2024 · 1 repository · arXiv:2405.18111Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Attention-based sequential recommendation system using multimodal data 28 May 2024 · 0 repositories · arXiv:2405.17959
-
Don't Forget to Connect! Improving RAG with Graph-based Reranking 28 May 2024 · 0 repositories · arXiv:2405.18414
-
Edinburgh Clinical NLP at MEDIQA-CORR 2024: Guiding Large Language Models with Hints 28 May 2024 · 0 repositories · arXiv:2405.18028
-
LLMs and Memorization: On Quality and Specificity of Copyright Compliance 28 May 2024 · 1 repository · arXiv:2405.18492
-
Towards a Sampling Theory for Implicit Neural Representations 28 May 2024 · 0 repositories · arXiv:2405.18410
-
Understanding Intrinsic Socioeconomic Biases in Large Language Models 28 May 2024 · 0 repositories · arXiv:2405.18662
-
WIDIn: Wording Image for Domain-Invariant Representation in Single-Source Domain Generalization 28 May 2024 · 0 repositories · arXiv:2405.18405
-
Assessing LLMs Suitability for Knowledge Graph Completion 27 May 2024 · 1 repository · arXiv:2405.17249
-
Augmenting Textual Generation via Topology Aware Retrieval 27 May 2024 · 0 repositories · arXiv:2405.17602
-
DeeperImpact: Optimizing Sparse Learned Index Structures 27 May 2024 · 1 repository · arXiv:2405.17093
-
Detecting Deceptive Dark Patterns in E-commerce Platforms 27 May 2024 · 0 repositories · arXiv:2406.01608
-
Exploiting the Layered Intrinsic Dimensionality of Deep Models for Practical Adversarial Training 27 May 2024 · 0 repositories · arXiv:2405.17130
-
InversionView: A General-Purpose Method for Reading Information from Neural Activations 27 May 2024 · 1 repository · arXiv:2405.17653Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; the one sample that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
REVECA: Adaptive Planning and Trajectory-based Validation in Cooperative Language Agents using Information Relevance and Relative Proximity 27 May 2024 · 0 repositories · arXiv:2405.16751
-
NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models 27 May 2024 · 0 repositories · arXiv:2405.17428
-
PAE: LLM-based Product Attribute Extraction for E-Commerce Fashion Trends 27 May 2024 · 0 repositories · arXiv:2405.17533
-
Performance evaluation of Reddit Comments using Machine Learning and Natural Language Processing methods in Sentiment Analysis 27 May 2024 · 0 repositories · arXiv:2405.16810
-
QUB-Cirdan at "Discharge Me!": Zero shot discharge letter generation by open-source LLM 27 May 2024 · 0 repositories · arXiv:2406.00041
-
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation 27 May 2024 · 1 repository · arXiv:2405.17057Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
RTL-Repo: A Benchmark for Evaluating LLMs on Large-Scale RTL Design Projects 27 May 2024 · 1 repository · arXiv:2405.17378
-
The Scaling Law in Stellar Light Curves 27 May 2024 · 0 repositories · arXiv:2405.17156
-
THREAD: Thinking Deeper with Recursive Spawning 27 May 2024 · 1 repository · arXiv:2405.17402Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Video Enriched Retrieval Augmented Generation Using Aligned Video Captions 27 May 2024 · 1 repository · arXiv:2405.17706
-
AI-Generated Text Detection and Classification Based on BERT Deep Learning Algorithm 26 May 2024 · 0 repositories · arXiv:2405.16422
-
GRAG: Graph Retrieval-Augmented Generation 26 May 2024 · 1 repository · arXiv:2405.16506
-
M-RAG: Reinforcing Large Language Model Performance through Retrieval-Augmented Generation with Multiple Partitions 26 May 2024 · 0 repositories · arXiv:2405.16420
-
Accelerating Inference of Retrieval-Augmented Generation via Sparse Context Selection 25 May 2024 · 0 repositories · arXiv:2405.16178
-
Accelerating Transformers with Spectrum-Preserving Token Merging 25 May 2024 · 1 repository · arXiv:2405.16148Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
AutoManual: Constructing Instruction Manuals by LLM Agents via Interactive Environmental Learning 25 May 2024 · 1 repository · arXiv:2405.16247Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Incremental Comprehension of Garden-Path Sentences by Large Language Models: Semantic Interpretation, Syntactic Re-Analysis, and Attention 25 May 2024 · 0 repositories · arXiv:2405.16042
-
MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time 25 May 2024 · 0 repositories · arXiv:2405.16265
-
Towards Unlocking Insights from Logbooks Using AI 25 May 2024 · 0 repositories · arXiv:2406.12881
-
An Evaluation of Estimative Uncertainty in Large Language Models 24 May 2024 · 0 repositories · arXiv:2405.15185
-
Benchmarking the Performance of Pre-trained LLMs across Urdu NLP Tasks 24 May 2024 · 0 repositories · arXiv:2405.15453
-
CulturePark: Boosting Cross-cultural Understanding in Large Language Models 24 May 2024 · 1 repository · arXiv:2405.15145Syntology official: harvested, nothing ran · 0 ran · 7 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Enhancing Augmentative and Alternative Communication with Card Prediction and Colourful Semantics 24 May 2024 · 0 repositories · arXiv:2405.15896
-
Evaluating and Safeguarding the Adversarial Robustness of Retrieval-Based In-Context Learning 24 May 2024 · 1 repository · arXiv:2405.15984Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Generalizable and Scalable Multistage Biomedical Concept Normalization Leveraging Large Language Models 24 May 2024 · 1 repository · arXiv:2405.15122
-
GPT is Not an Annotator: The Necessity of Human Annotation in Fairness Benchmark Construction 24 May 2024 · 0 repositories · arXiv:2405.15760
-
Learning the Language of Protein Structure 24 May 2024 · 1 repository · arXiv:2405.15840Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Comet: A Communication-efficient and Performant Approximation for Private Transformer Inference 24 May 2024 · 0 repositories · arXiv:2405.17485
-
The Impact and Opportunities of Generative AI in Fact-Checking 24 May 2024 · 0 repositories · arXiv:2405.15985
-
The Buffer Mechanism for Multi-Step Information Reasoning in Language Models 24 May 2024 · 0 repositories · arXiv:2405.15302
-
A Structure-Aware Framework for Learning Device Placements on Computation Graphs 23 May 2024 · 1 repository · arXiv:2405.14185
-
CEEBERT: Cross-Domain Inference in Early Exit BERT 23 May 2024 · 1 repository · arXiv:2405.15039Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
EditWorld: Simulating World Dynamics for Instruction-Following Image Editing 23 May 2024 · 1 repository · arXiv:2405.14785Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Eliciting Informative Text Evaluations with Large Language Models 23 May 2024 · 1 repository · arXiv:2405.15077Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step 23 May 2024 · 1 repository · arXiv:2405.14838Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
HippoRAG: Neurobiologically Inspired Long-Term Memory for Large Language Models 23 May 2024 · 2 repositories · arXiv:2405.14831Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Large Language Models Can Self-Correct with Key Condition Verification 23 May 2024 · 0 repositories · arXiv:2405.14092
-
Not All Language Model Features Are Linear 23 May 2024 · 1 repository · arXiv:2405.14860Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
PhiNets: Brain-inspired Non-contrastive Learning Based on Temporal Prediction Hypothesis 23 May 2024 · 0 repositories · arXiv:2405.14650Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
RaFe: Ranking Feedback Improves Query Rewriting for RAG 23 May 2024 · 0 repositories · arXiv:2405.14431
-
ViHateT5: Enhancing Hate Speech Detection in Vietnamese With A Unified Text-to-Text Transformer Model 23 May 2024 · 1 repository · arXiv:2405.14141
-
WISE: Rethinking the Knowledge Memory for Lifelong Model Editing of Large Language Models 23 May 2024 · 1 repository · arXiv:2405.14768Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 9 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Automated Evaluation of Retrieval-Augmented Language Models with Task-Specific Exam Generation 22 May 2024 · 1 repository · arXiv:2405.13622
-
Automatically Identifying Local and Global Circuits with Linear Computation Graphs 22 May 2024 · 0 repositories · arXiv:2405.13868
-
Evaluating Large Language Models with Human Feedback: Establishing a Swedish Benchmark 22 May 2024 · 1 repository · arXiv:2405.14006
-
FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research 22 May 2024 · 1 repository · arXiv:2405.13576
-
How to set AdamW's weight decay as you scale model and dataset size 22 May 2024 · 0 repositories · arXiv:2405.13698
-
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR 22 May 2024 · 0 repositories · arXiv:2406.00014
-
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment 22 May 2024 · 1 repository · arXiv:2405.13911Syntology official (archive's flag): 3 ran · 6 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
TrojanRAG: Retrieval-Augmented Generation Can Be Backdoor Driver in Large Language Models 22 May 2024 · 1 repository · arXiv:2405.13401
-
Unleashing the Power of Unlabeled Data: A Self-supervised Learning Framework for Cyber Attack Detection in Smart Grids 22 May 2024 · 0 repositories · arXiv:2405.13965
-
FAdam: Adam is a natural gradient optimizer using diagonal empirical Fisher information 21 May 2024 · 1 repository · arXiv:2405.12807
-
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities 21 May 2024 · 0 repositories · arXiv:2405.12750
-
How Reliable AI Chatbots are for Disease Prediction from Patient Complaints? 21 May 2024 · 0 repositories · arXiv:2405.13219
-
Investigating Persuasion Techniques in Arabic: An Empirical Study Leveraging Large Language Models 21 May 2024 · 0 repositories · arXiv:2405.12884
-
Quantifying Semantic Emergence in Language Models 21 May 2024 · 1 repository · arXiv:2405.12617Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples)
-
The 2nd FutureDial Challenge: Dialog Systems with Retrieval Augmented Generation (FutureDial-RAG) 21 May 2024 · 1 repository · arXiv:2405.13084
-
A review on the use of large language models as virtual tutors 20 May 2024 · 0 repositories · arXiv:2405.11983
-
CReMa: Crisis Response through Computational Identification and Matching of Cross-Lingual Requests and Offers Shared on Social Media 20 May 2024 · 0 repositories · arXiv:2405.11897
-
Degree of Irrationality: Sentiment and Implied Volatility Surface 20 May 2024 · 0 repositories · arXiv:2405.11730
-
Evaluating and Modeling Social Intelligence: A Comparative Study of Human and AI Capabilities 20 May 2024 · 1 repository · arXiv:2405.11841
-
Question-Based Retrieval using Atomic Units for Enterprise RAG 20 May 2024 · 0 repositories · arXiv:2405.12363
-
DaVinci at SemEval-2024 Task 9: Few-shot prompting GPT-3.5 for Unconventional Reasoning 19 May 2024 · 0 repositories · arXiv:2405.11559
-
Human-Centered LLM-Agent User Interface: A Position Paper 19 May 2024 · 1 repository · arXiv:2405.13050
-
Your Transformer is Secretly Linear 19 May 2024 · 1 repository · arXiv:2405.12250Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Zero-Shot Stance Detection using Contextual Data Generation with LLMs 19 May 2024 · 1 repository · arXiv:2405.11637Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Cross-Language Assessment of Mathematical Capability of ChatGPT 18 May 2024 · 0 repositories · arXiv:2405.11264
-
Exploring speech style spaces with language models: Emotional TTS without emotion labels 18 May 2024 · 0 repositories · arXiv:2405.11413
-
Meta Reinforcement Learning for Resource Allocation in Multi-Antenna UAV Network with Rate Splitting Multiple Access 18 May 2024 · 0 repositories · arXiv:2405.11306
-
ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios 17 May 2024 · 0 repositories · arXiv:2405.10808
-
Empowering Prior to Court Legal Analysis: A Transparent and Accessible Dataset for Defensive Statement Classification and Interpretation 17 May 2024 · 0 repositories · arXiv:2405.10702
-
Evaluation of large language model performance on the Biomedical Language Understanding and Reasoning Benchmark 17 May 2024 · 0 repositories
-
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks 17 May 2024 · 1 repository · arXiv:2405.10548Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
NeuroAssist: Enhancing Cognitive-Computer Synergy with Adaptive AI and Advanced Neural Decoding for Efficient EEG Signal Classification 17 May 2024 · 0 repositories · arXiv:2406.01600
-
Tailoring Vaccine Messaging with Common-Ground Opinions 17 May 2024 · 1 repository · arXiv:2405.10861
-
FinTextQA: A Dataset for Long-form Financial Question Answering 16 May 2024 · 0 repositories · arXiv:2405.09980
-
GPT Store Mining and Analysis 16 May 2024 · 0 repositories · arXiv:2405.10210
-
HW-GPT-Bench: Hardware-Aware Architecture Benchmark for Language Models 16 May 2024 · 2 repositories · arXiv:2405.10299
-
Optimization Techniques for Sentiment Analysis Based on LLM (GPT-3) 16 May 2024 · 0 repositories · arXiv:2405.09770
-
Bridging the gap in online hate speech detection: a comparative analysis of BERT and traditional models for homophobic content identification on X/Twitter 15 May 2024 · 0 repositories · arXiv:2405.09221
-
IM-RAG: Multi-Round Retrieval-Augmented Generation Through Learning Inner Monologues 15 May 2024 · 0 repositories · arXiv:2405.13021
-
LoRA Learns Less and Forgets Less 15 May 2024 · 1 repository · arXiv:2405.09673
-
Matching domain experts by training from scratch on domain knowledge 15 May 2024 · 0 repositories · arXiv:2405.09395
-
Transfer Learning in Pre-Trained Large Language Models for Malware Detection Based on System Calls 15 May 2024 · 0 repositories · arXiv:2405.09318
-
Beyond Scaling Laws: Understanding Transformer Performance with Associative Memory 14 May 2024 · 0 repositories · arXiv:2405.08707