Methods › General › Regularization › Weight Decay › Papers, page 17
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 17 of 108: papers 1,601 to 1,700 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Integrating Temporal Representations for Dynamic Memory Retrieval and Management in Large Language Models 17 Oct 2024 · 0 repositories · arXiv:2410.13553
-
Jailbreaking LLM-Controlled Robots 17 Oct 2024 · 0 repositories · arXiv:2410.13691
-
Linguistically Grounded Analysis of Language Models using Shapley Head Values 17 Oct 2024 · 0 repositories · arXiv:2410.13396
-
LoLDU: Low-Rank Adaptation via Lower-Diag-Upper Decomposition for Parameter-Efficient Fine-Tuning 17 Oct 2024 · 1 repository · arXiv:2410.13618Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Judgment of Learning: A Human Ability Beyond Generative Artificial Intelligence 17 Oct 2024 · 0 repositories · arXiv:2410.13392
-
MIRAGE-Bench: Automatic Multilingual Benchmark Arena for Retrieval-Augmented Generation Systems 17 Oct 2024 · 1 repository · arXiv:2410.13716
-
RAG-DDR: Optimizing Retrieval-Augmented Generation Using Differentiable Data Rewards 17 Oct 2024 · 1 repository · arXiv:2410.13509Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
SBI-RAG: Enhancing Math Word Problem Solving for Students through Schema-Based Instruction and Retrieval-Augmented Generation 17 Oct 2024 · 1 repository · arXiv:2410.13293
-
SouLLMate: An Application Enhancing Diverse Mental Health Support with Adaptive LLMs, Prompt Engineering, and RAG Techniques 17 Oct 2024 · 0 repositories · arXiv:2410.16322
-
Training Compute-Optimal Vision Transformers for Brain Encoding 17 Oct 2024 · 0 repositories · arXiv:2410.19810
-
Agent Skill Acquisition for Large Language Models via CycleQD 16 Oct 2024 · 1 repository · arXiv:2410.14735Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
AT-RAG: An Adaptive RAG Model Enhancing Query Efficiency with Topic Filtering and Iterative Reasoning 16 Oct 2024 · 1 repository · arXiv:2410.12886
-
CoFE-RAG: A Comprehensive Full-chain Evaluation Framework for Retrieval-Augmented Generation with Enhanced Data Diversity 16 Oct 2024 · 1 repository · arXiv:2410.12248
-
Communication-Efficient and Tensorized Federated Fine-Tuning of Large Language Models 16 Oct 2024 · 0 repositories · arXiv:2410.13097
-
Context-Scaling versus Task-Scaling in In-Context Learning 16 Oct 2024 · 0 repositories · arXiv:2410.12783
-
Evaluation of Attribution Bias in Retrieval-Augmented Large Language Models 16 Oct 2024 · 0 repositories · arXiv:2410.12380
-
Exploring Large Language Models for Hate Speech Detection in Rioplatense Spanish 16 Oct 2024 · 0 repositories · arXiv:2410.12174
-
FusionLLM: A Decentralized LLM Training System on Geo-distributed GPUs with Adaptive Compression 16 Oct 2024 · 0 repositories · arXiv:2410.12707
-
Is Semantic Chunking Worth the Computational Cost? 16 Oct 2024 · 0 repositories · arXiv:2410.13070
-
Kallini et al. (2024) do not compare impossible languages with constituency-based ones 16 Oct 2024 · 0 repositories · arXiv:2410.12271
-
Meta-Chunking: Learning Text Segmentation and Semantic Completion via Logical Perception 16 Oct 2024 · 1 repository · arXiv:2410.12788Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models 16 Oct 2024 · 1 repository · arXiv:2410.13085Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
ShapefileGPT: A Multi-Agent Large Language Model Framework for Automated Shapefile Processing 16 Oct 2024 · 0 repositories · arXiv:2410.12376
-
Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective 16 Oct 2024 · 1 repository · arXiv:2410.12490Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Table-LLM-Specialist: Language Model Specialists for Tables using Iterative Generator-Validator Fine-tuning 16 Oct 2024 · 0 repositories · arXiv:2410.12164
-
Unitary Multi-Margin BERT for Robust Natural Language Processing 16 Oct 2024 · 1 repository · arXiv:2410.12759
-
When Not to Answer: Evaluating Prompts on GPT Models for Effective Abstention in Unanswerable Math Word Problems 16 Oct 2024 · 0 repositories · arXiv:2410.13029
-
Athena: Retrieval-augmented Legal Judgment Prediction with Large Language Models 15 Oct 2024 · 0 repositories · arXiv:2410.11195
-
Deciphering the Chaos: Enhancing Jailbreak Attacks via Adversarial Prompt Translation 15 Oct 2024 · 1 repository · arXiv:2410.11317Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
DynamicER: Resolving Emerging Mentions to Dynamic Entities for RAG 15 Oct 2024 · 1 repository · arXiv:2410.11494
-
Evidence of Cognitive Deficits andDevelopmental Advances in Generative AI: A Clock Drawing Test Analysis 15 Oct 2024 · 0 repositories · arXiv:2410.11756
-
Holistic Reasoning with Long-Context LMs: A Benchmark for Database Operations on Massive Textual Data 15 Oct 2024 · 0 repositories · arXiv:2410.11996
-
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection 15 Oct 2024 · 0 repositories · arXiv:2410.11230
-
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions 15 Oct 2024 · 0 repositories · arXiv:2410.11833
-
MTU-Bench: A Multi-granularity Tool-Use Benchmark for Large Language Models 15 Oct 2024 · 1 repository · arXiv:2410.11710
-
Nonlinear Gaussian process tomography with imposed non-negativity constraints on physical quantities for plasma diagnostics 15 Oct 2024 · 0 repositories · arXiv:2410.11454
-
On the Capacity of Citation Generation by Large Language Models 15 Oct 2024 · 0 repositories · arXiv:2410.11217
-
Pixology: Probing the Linguistic and Visual Capabilities of Pixel-based Language Models 15 Oct 2024 · 1 repository · arXiv:2410.12011Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
ReDeEP: Detecting Hallucination in Retrieval-Augmented Generation via Mechanistic Interpretability 15 Oct 2024 · 0 repositories · arXiv:2410.11414
-
Retrieval Augmented Spelling Correction for E-Commerce Applications 15 Oct 2024 · 0 repositories · arXiv:2410.11655
-
RuleRAG: Rule-guided retrieval-augmented generation with language models for question answering 15 Oct 2024 · 1 repository · arXiv:2410.22353
-
SEER: Self-Aligned Evidence Extraction for Retrieval-Augmented Generation 15 Oct 2024 · 0 repositories · arXiv:2410.11315
-
Self-adaptive Multimodal Retrieval-Augmented Generation 15 Oct 2024 · 1 repository · arXiv:2410.11321
-
Sorted Weight Sectioning for Energy-Efficient Unstructured Sparse DNNs on Compute-in-Memory Crossbars 15 Oct 2024 · 0 repositories · arXiv:2410.11298
-
Synthetic Interlocutors. Experiments with Generative AI to Prolong Ethnographic Encounters 15 Oct 2024 · 0 repositories · arXiv:2410.11395
-
Telco-DPR: A Hybrid Dataset for Evaluating Retrieval Models of 3GPP Technical Specifications 15 Oct 2024 · 0 repositories · arXiv:2410.19790
-
The Fair Language Model Paradox 15 Oct 2024 · 0 repositories · arXiv:2410.11985
-
An Annotated Dataset of Errors in Premodern Greek and Baselines for Detecting Them 14 Oct 2024 · 1 repository · arXiv:2410.11071
-
Enhancing Retrieval-Augmented Audio Captioning with Generation-Assisted Multimodal Querying and Progressive Learning 14 Oct 2024 · 0 repositories · arXiv:2410.10913
-
Beyond-RAG: Question Identification and Answer Generation in Real-Time Conversations 14 Oct 2024 · 0 repositories · arXiv:2410.10136
-
Code-Mixer Ya Nahi: Novel Approaches to Measuring Multilingual LLMs' Code-Mixing Capabilities 14 Oct 2024 · 0 repositories · arXiv:2410.11079
-
Dissecting embedding method: learning higher-order structures from data 14 Oct 2024 · 0 repositories · arXiv:2410.10917
-
Double Jeopardy and Climate Impact in the Use of Large Language Models: Socio-economic Disparities and Reduced Utility for Non-English Speakers 14 Oct 2024 · 1 repository · arXiv:2410.10665
-
EasyRAG: Efficient Retrieval-Augmented Generation Framework for Automated Network Operations 14 Oct 2024 · 1 repository · arXiv:2410.10315
-
FunnelRAG: A Coarse-to-Fine Progressive Retrieval Paradigm for RAG 14 Oct 2024 · 0 repositories · arXiv:2410.10293
-
Gender Bias of LLM in Economics: An Existentialism Perspective 14 Oct 2024 · 0 repositories · arXiv:2410.19775
-
Graph of Records: Boosting Retrieval Augmented Generation for Long-context Summarization with Graphs 14 Oct 2024 · 1 repository · arXiv:2410.11001
-
One Language, Many Gaps: Evaluating Dialect Fairness and Robustness of Large Language Models in Reasoning Tasks 14 Oct 2024 · 1 repository · arXiv:2410.11005Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
Performance in a dialectal profiling task of LLMs for varieties of Brazilian Portuguese 14 Oct 2024 · 0 repositories · arXiv:2410.10991
-
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models 14 Oct 2024 · 1 repository · arXiv:2410.10542
-
RoCoFT: Efficient Finetuning of Large Language Models with Row-Column Updates 14 Oct 2024 · 1 repository · arXiv:2410.10075
-
STACKFEED: Structured Textual Actor-Critic Knowledge Base Editing with FeedBack 14 Oct 2024 · 0 repositories · arXiv:2410.10584
-
Towards Better Multi-head Attention via Channel-wise Sample Permutation 14 Oct 2024 · 1 repository · arXiv:2410.10914Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents 14 Oct 2024 · 1 repository · arXiv:2410.10594Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Will LLMs Replace the Encoder-Only Models in Temporal Relation Classification? 14 Oct 2024 · 1 repository · arXiv:2410.10476Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
A Comparative Study of PDF Parsing Tools Across Diverse Document Categories 13 Oct 2024 · 0 repositories · arXiv:2410.09871
-
Can In-context Learning Really Generalize to Out-of-distribution Tasks? 13 Oct 2024 · 0 repositories · arXiv:2410.09695
-
Evaluating Gender Bias of LLMs in Making Morality Judgements 13 Oct 2024 · 0 repositories · arXiv:2410.09992
-
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG 13 Oct 2024 · 0 repositories · arXiv:2410.09699
-
Investigating Implicit Bias in Large Language Models: A Large-Scale Study of Over 50 LLMs 13 Oct 2024 · 0 repositories · arXiv:2410.12864
-
Learning to Rank for Multiple Retrieval-Augmented Models through Iterative Utility Maximization 13 Oct 2024 · 0 repositories · arXiv:2410.09942
-
M2M-Gen: A Multimodal Framework for Automated Background Music Generation in Japanese Manga Using Large Language Models 13 Oct 2024 · 0 repositories · arXiv:2410.09928
-
Automatic Speech Recognition with BERT and CTC Transformers: A Review 12 Oct 2024 · 0 repositories · arXiv:2410.09456
-
\llinstruct: An Instruction-tuned model for English Language Proficiency Assessments 12 Oct 2024 · 0 repositories · arXiv:2410.09314
-
Toward General Instruction-Following Alignment for Retrieval-Augmented Generation 12 Oct 2024 · 1 repository · arXiv:2410.09584Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
A Methodology for Evaluating RAG Systems: A Case Study On Configuration Dependency Validation 11 Oct 2024 · 1 repository · arXiv:2410.08801
-
AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation 11 Oct 2024 · 1 repository · arXiv:2410.09040
-
Enhancing Long Context Performance in LLMs Through Inner Loop Query Mechanism 11 Oct 2024 · 0 repositories · arXiv:2410.12859
-
Extra Global Attention Designation Using Keyword Detection in Sparse Transformer Architectures 11 Oct 2024 · 0 repositories · arXiv:2410.08971
-
Fine-Tuning In-House Large Language Models to Infer Differential Diagnosis from Radiology Reports 11 Oct 2024 · 0 repositories · arXiv:2410.09234
-
Humanity in AI: Detecting the Personality of Large Language Models 11 Oct 2024 · 0 repositories · arXiv:2410.08545
-
Long Range Named Entity Recognition for Marathi Documents 11 Oct 2024 · 0 repositories · arXiv:2410.09192
-
Observing the Southern US Culture of Honor Using Large-Scale Social Media Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.13887
-
Optimized Biomedical Question-Answering Services with LLM and Multi-BERT Integration 11 Oct 2024 · 0 repositories · arXiv:2410.12856
-
oRetrieval Augmented Generation for 10 Large Language Models and its Generalizability in Assessing Medical Fitness 11 Oct 2024 · 0 repositories · arXiv:2410.08431
-
Retriever-and-Memory: Towards Adaptive Note-Enhanced Retrieval-Augmented Generation 11 Oct 2024 · 1 repository · arXiv:2410.08821Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08698
-
StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization 11 Oct 2024 · 1 repository · arXiv:2410.08815
-
Synth-SONAR: Sonar Image Synthesis with Enhanced Diversity and Realism via Dual Diffusion Models and GPT Prompting 11 Oct 2024 · 1 repository · arXiv:2410.08612
-
Adam Exploits ℓ_∞-geometry of Loss Landscape via Coordinate-wise Adaptivity 10 Oct 2024 · 1 repository · arXiv:2410.08198Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models 10 Oct 2024 · 0 repositories · arXiv:2410.08207
-
Do You Know What You Are Talking About? Characterizing Query-Knowledge Relevance For Reliable Retrieval Augmented Generation 10 Oct 2024 · 0 repositories · arXiv:2410.08320
-
Fine-Tuning Language Models for Ethical Ambiguity: A Comparative Study of Alignment with Human Responses 10 Oct 2024 · 0 repositories · arXiv:2410.07826
-
FLIER: Few-shot Language Image Models Embedded with Latent Representations 10 Oct 2024 · 0 repositories · arXiv:2410.07648
-
News Reporter: A Multi-lingual LLM Framework for Broadcast T.V News 10 Oct 2024 · 0 repositories · arXiv:2410.07520
-
No Free Lunch: Retrieval-Augmented Generation Undermines Fairness in LLMs, Even for Vigilant Users 10 Oct 2024 · 0 repositories · arXiv:2410.07589
-
Privately Learning from Graphs with Applications in Fine-tuning Large Language Models 10 Oct 2024 · 1 repository · arXiv:2410.08299
-
Robust AI-Generated Text Detection by Restricted Embeddings 10 Oct 2024 · 1 repository · arXiv:2410.08113Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
The Rise of AI-Generated Content in Wikipedia 10 Oct 2024 · 1 repository · arXiv:2410.08044Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TurboRAG: Accelerating Retrieval-Augmented Generation with Precomputed KV Caches for Chunked Text 10 Oct 2024 · 1 repository · arXiv:2410.07590