Methods › General › Regularization › Weight Decay › Papers, page 21
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 21 of 108: papers 2,001 to 2,100 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
HyPA-RAG: A Hybrid Parameter Adaptive Retrieval-Augmented Generation System for AI Legal and Policy Applications 29 Aug 2024 · 0 repositories · arXiv:2409.09046
-
LLaVA-Chef: A Multi-modal Generative Model for Food Recipes 29 Aug 2024 · 1 repository · arXiv:2408.16889
-
A Simple Baseline with Single-encoder for Referring Image Segmentation 28 Aug 2024 · 0 repositories · arXiv:2408.15521
-
An Extremely Data-efficient and Generative LLM-based Reinforcement Learning Agent for Recommenders 28 Aug 2024 · 0 repositories · arXiv:2408.16032
-
CBF-LLM: Safe Control for LLM Alignment 28 Aug 2024 · 1 repository · arXiv:2408.15625
-
Conan-embedding: General Text Embedding with More and Better Negative Samples 28 Aug 2024 · 0 repositories · arXiv:2408.15710
-
FRACTURED-SORRY-Bench: Framework for Revealing Attacks in Conversational Turns Undermining Refusal Efficacy and Defenses over SORRY-Bench (Automated Multi-shot Jailbreaks) 28 Aug 2024 · 0 repositories · arXiv:2408.16163
-
Is Personality Prediction Possible Based on Reddit Comments? 28 Aug 2024 · 0 repositories · arXiv:2408.16089
-
Leveraging Large Language Models for Wireless Symbol Detection via In-Context Learning 28 Aug 2024 · 0 repositories · arXiv:2409.00124
-
LRP4RAG: Detecting Hallucinations in Retrieval-Augmented Generation via Layer-wise Relevance Propagation 28 Aug 2024 · 3 repositories · arXiv:2408.15533
-
Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation 28 Aug 2024 · 1 repository · arXiv:2408.15876
-
A Survey of Large Language Models for European Languages 27 Aug 2024 · 0 repositories · arXiv:2408.15040
-
Into the Unknown Unknowns: Engaged Human Learning through Participation in Language Model Agent Conversations 27 Aug 2024 · 1 repository · arXiv:2408.15232
-
Strategic Optimization and Challenges of Large Language Models in Object-Oriented Programming 27 Aug 2024 · 0 repositories · arXiv:2408.14834
-
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis 27 Aug 2024 · 0 repositories · arXiv:2409.00106
-
CHARTOM: A Visual Theory-of-Mind Benchmark for Multimodal Large Language Models 26 Aug 2024 · 1 repository · arXiv:2408.14419
-
Question answering system of bridge design specification based on large language model 26 Aug 2024 · 1 repository · arXiv:2408.13282
-
Bidirectional Awareness Induction in Autoregressive Seq2Seq Models 25 Aug 2024 · 0 repositories · arXiv:2408.13959
-
CodeGraph: Enhancing Graph Reasoning of LLMs with Code 25 Aug 2024 · 1 repository · arXiv:2408.13863
-
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task 25 Aug 2024 · 0 repositories · arXiv:2408.13909
-
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models 25 Aug 2024 · 1 repository · arXiv:2409.00084
-
Pandora's Box or Aladdin's Lamp: A Comprehensive Analysis Revealing the Role of RAG Noise in Large Language Models 24 Aug 2024 · 1 repository · arXiv:2408.13533Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
An In-Depth Investigation of Data Collection in LLM App Ecosystems 23 Aug 2024 · 0 repositories · arXiv:2408.13247
-
Enhancing Automated Program Repair with Solution Design 22 Aug 2024 · 0 repositories · arXiv:2408.12056
-
Enhancing Multi-hop Reasoning through Knowledge Erasure in Large Language Model Editing 22 Aug 2024 · 0 repositories · arXiv:2408.12456
-
GRATR: Zero-Shot Evidence Graph Retrieval-Augmented Trustworthiness Reasoning 22 Aug 2024 · 1 repository · arXiv:2408.12333
-
Large Language Models as Foundations for Next-Gen Dense Retrieval: A Comprehensive Empirical Assessment 22 Aug 2024 · 0 repositories · arXiv:2408.12194
-
LLMs are not Zero-Shot Reasoners for Biomedical Information Extraction 22 Aug 2024 · 0 repositories · arXiv:2408.12249
-
Optimizing Performance: How Compact Models Match or Exceed GPT's Classification Capabilities through Fine-Tuning 22 Aug 2024 · 0 repositories · arXiv:2409.11408
-
Unlearning Trojans in Large Language Models: A Comparison Between Natural Language and Source Code 22 Aug 2024 · 0 repositories · arXiv:2408.12416
-
A Quick, trustworthy spectral knowledge Q&A system leveraging retrieval-augmented generation on LLM 21 Aug 2024 · 1 repository · arXiv:2408.11557
-
Ancient Wisdom, Modern Tools: Exploring Retrieval-Augmented LLMs for Ancient Indian Philosophy 21 Aug 2024 · 1 repository · arXiv:2408.11903
-
Applying and Evaluating Large Language Models in Mental Health Care: A Scoping Review of Human-Assessed Generative Tasks 21 Aug 2024 · 0 repositories · arXiv:2408.11288
-
Approaching Deep Learning through the Spectral Dynamics of Weights 21 Aug 2024 · 1 repository · arXiv:2408.11804Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples)
-
D-RMGPT: Robot-assisted collaborative tasks driven by large multimodal models 21 Aug 2024 · 0 repositories · arXiv:2408.11761
-
Mixed Sparsity Training: Achieving 4× FLOP Reduction for Transformer Pretraining 21 Aug 2024 · 0 repositories · arXiv:2408.11746
-
WeQA: A Benchmark for Retrieval Augmented Generation in Wind Energy Domain 21 Aug 2024 · 0 repositories · arXiv:2408.11800
-
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization 21 Aug 2024 · 0 repositories · arXiv:2408.12003
-
RAGLAB: A Modular and Research-Oriented Unified Framework for Retrieval-Augmented Generation 21 Aug 2024 · 1 repository · arXiv:2408.11381
-
The Self-Contained Negation Test Set 21 Aug 2024 · 0 repositories · arXiv:2408.11469
-
Unlocking Adversarial Suffix Optimization Without Affirmative Phrases: Efficient Black-box Jailbreaking via LLM as Optimizer 21 Aug 2024 · 1 repository · arXiv:2408.11313Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CTP-LLM: Clinical Trial Phase Transition Prediction Using Large Language Models 20 Aug 2024 · 0 repositories · arXiv:2408.10995
-
How Well Do Large Language Models Serve as End-to-End Secure Code Agents for Python? 20 Aug 2024 · 0 repositories · arXiv:2408.10495
-
Language Modeling on Tabular Data: A Survey of Foundations, Techniques and Evolution 20 Aug 2024 · 1 repository · arXiv:2408.10548
-
Reading with Intent 20 Aug 2024 · 0 repositories · arXiv:2408.11189
-
Reconciling Methodological Paradigms: Employing Large Language Models as Novice Qualitative Research Assistants in Talent Management Research 20 Aug 2024 · 0 repositories · arXiv:2408.11043
-
Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs 20 Aug 2024 · 1 repository · arXiv:2408.10902
-
Towardseffective teaching assistants: From intent-based chatbots to LLM-poweredteachingassistants 20 Aug 2024 · 0 repositories
-
Tracing Privacy Leakage of Language Models to Training Data via Adjusted Influence Functions 20 Aug 2024 · 0 repositories · arXiv:2408.10468
-
Security Attacks on LLM-based Code Completion Tools 20 Aug 2024 · 1 repository · arXiv:2408.11006
-
A Strategy to Combine 1stGen Transformers and Open LLMs for Automatic Text Classification 19 Aug 2024 · 0 repositories · arXiv:2408.09629
-
Acquiring Bidirectionality via Large and Small Language Models 19 Aug 2024 · 1 repository · arXiv:2408.09640
-
Active Learning for Identifying Disaster-Related Tweets: A Comparison with Keyword Filtering and Generic Fine-Tuning 19 Aug 2024 · 0 repositories · arXiv:2408.09914
-
Carbon Footprint Accounting Driven by Large Language Models and Retrieval-augmented Generation 19 Aug 2024 · 0 repositories · arXiv:2408.09713
-
ELDER: Enhancing Lifelong Model Editing with Mixture-of-LoRA 19 Aug 2024 · 1 repository · arXiv:2408.11869
-
Enhanced document retrieval with topic embeddings 19 Aug 2024 · 0 repositories · arXiv:2408.10435
-
GARLIC: GPT-Augmented Reinforcement Learning with Intelligent Control for Vehicle Dispatching 19 Aug 2024 · 0 repositories · arXiv:2408.10286
-
LegalBench-RAG: A Benchmark for Retrieval-Augmented Generation in the Legal Domain 19 Aug 2024 · 1 repository · arXiv:2408.10343Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Rhyme-aware Chinese lyric generator based on GPT 19 Aug 2024 · 0 repositories · arXiv:2408.10130
-
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training 19 Aug 2024 · 0 repositories · arXiv:2408.10013
-
Agentic Retrieval-Augmented Generation for Time Series Analysis 18 Aug 2024 · 0 repositories · arXiv:2408.14484
-
Clustering and Alignment: Understanding the Training Dynamics in Modular Addition 18 Aug 2024 · 0 repositories · arXiv:2408.09414
-
ConVerSum: A Contrastive Learning-based Approach for Data-Scarce Solution of Cross-Lingual Summarization Beyond Direct Equivalents 17 Aug 2024 · 0 repositories · arXiv:2408.09273
-
Sentiment analysis of preservice teachers' reflections using a large language model 17 Aug 2024 · 0 repositories · arXiv:2408.11862
-
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering 17 Aug 2024 · 0 repositories · arXiv:2408.09174
-
TC-RAG:Turing-Complete RAG's Case study on Medical LLM Systems 17 Aug 2024 · 2 repositories · arXiv:2408.09199
-
A Mean Field Ansatz for Zero-Shot Weight Transfer 16 Aug 2024 · 0 repositories · arXiv:2408.08681
-
CIKMar: A Dual-Encoder Approach to Prompt-Based Reranking in Educational Dialogue Systems 16 Aug 2024 · 0 repositories · arXiv:2408.08805
-
CommunityKG-RAG: Leveraging Community Structures in Knowledge Graphs for Advanced Retrieval-Augmented Generation in Fact-Checking 16 Aug 2024 · 1 repository · arXiv:2408.08535
-
Fine-tuning LLMs for Autonomous Spacecraft Control: A Case Study Using Kerbal Space Program 16 Aug 2024 · 1 repository · arXiv:2408.08676
-
Improving VTE Identification through Language Models from Radiology Reports: A Comparative Study of Mamba, Phi-3 Mini, and BERT 16 Aug 2024 · 0 repositories · arXiv:2408.09043
-
Information-Theoretic Progress Measures reveal Grokking is an Emergent Phase Transition 16 Aug 2024 · 0 repositories · arXiv:2408.08944
-
Meta Knowledge for Retrieval Augmented Large Language Models 16 Aug 2024 · 0 repositories · arXiv:2408.09017
-
Quantifying the Effectiveness of Student Organization Activities using Natural Language Processing 16 Aug 2024 · 0 repositories · arXiv:2408.08694
-
The Fellowship of the LLMs: Multi-Agent Workflows for Synthetic Preference Optimization Dataset Generation 16 Aug 2024 · 1 repository · arXiv:2408.08688
-
VERA: Validation and Evaluation of Retrieval-Augmented Systems 16 Aug 2024 · 0 repositories · arXiv:2409.03759
-
Analytical Uncertainty-Based Loss Weighting in Multi-Task Learning 15 Aug 2024 · 0 repositories · arXiv:2408.07985
-
FuseChat: Knowledge Fusion of Chat Models 15 Aug 2024 · 3 repositories · arXiv:2408.07990
-
Graph Retrieval-Augmented Generation: A Survey 15 Aug 2024 · 1 repository · arXiv:2408.08921
-
Leveraging Web-Crawled Data for High-Quality Fine-Tuning 15 Aug 2024 · 1 repository · arXiv:2408.08003
-
Plan with Code: Comparing approaches for robust NL to DSL generation 15 Aug 2024 · 0 repositories · arXiv:2408.08335
-
Predicting Lung Cancer Patient Prognosis with Large Language Models 15 Aug 2024 · 0 repositories · arXiv:2408.07971
-
RAGChecker: A Fine-grained Framework for Diagnosing Retrieval-Augmented Generation 15 Aug 2024 · 1 repository · arXiv:2408.08067Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
CodeMirage: Hallucinations in Code Generated by Large Language Models 14 Aug 2024 · 0 repositories · arXiv:2408.08333
-
DataVisT5: A Pre-trained Language Model for Jointly Understanding Text and Data Visualization 14 Aug 2024 · 1 repository · arXiv:2408.07401
-
Enhancing Visual Question Answering through Ranking-Based Hybrid Training and Multimodal Fusion 14 Aug 2024 · 0 repositories · arXiv:2408.07303
-
Exploring Retrieval Augmented Generation in Arabic 14 Aug 2024 · 1 repository · arXiv:2408.07425
-
LiPCoT: Linear Predictive Coding based Tokenizer for Self-supervised Learning of Time Series Data via Language Models 14 Aug 2024 · 1 repository · arXiv:2408.07292
-
SAGE-RT: Synthetic Alignment data Generation for Safety Evaluation and Red Teaming 14 Aug 2024 · 0 repositories · arXiv:2408.11851
-
Transformers and Large Language Models for Efficient Intrusion Detection Systems: A Comprehensive Survey 14 Aug 2024 · 0 repositories · arXiv:2408.07583
-
BERT's Conceptual Cartography: Mapping the Landscapes of Meaning 13 Aug 2024 · 0 repositories · arXiv:2408.07190
-
Evaluating Cultural Adaptability of a Large Language Model via Simulation of Synthetic Personas 13 Aug 2024 · 1 repository · arXiv:2408.06929
-
Generative AI for automatic topic labelling 13 Aug 2024 · 0 repositories · arXiv:2408.07003
-
Pragmatic inference of scalar implicature by LLMs 13 Aug 2024 · 0 repositories · arXiv:2408.06673
-
TableGuard -- Securing Structured & Unstructured Data 13 Aug 2024 · 0 repositories · arXiv:2408.07045
-
Bayesian inference to improve quality of Retrieval Augmented Generation 12 Aug 2024 · 0 repositories · arXiv:2408.08901
-
Improving Structural Diversity of Blackbox LLMs via Chain-of-Specification Prompting 12 Aug 2024 · 0 repositories · arXiv:2408.06186
-
LOLgorithm: Integrating Semantic,Syntactic and Contextual Elements for Humor Classification 12 Aug 2024 · 0 repositories · arXiv:2408.06335
-
Optimizing RAG Techniques for Automotive Industry PDF Chatbots: A Case Study with Locally Deployed Ollama Models 12 Aug 2024 · 0 repositories · arXiv:2408.05933
-
The Language of Trauma: Modeling Traumatic Event Descriptions Across Domains with Explainable AI 12 Aug 2024 · 0 repositories · arXiv:2408.05977