Methods › General › Regularization › Weight Decay › Papers, page 15
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 15 of 108: papers 1,401 to 1,500 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Prompt-Efficient Fine-Tuning for GPT-like Deep Models to Reduce Hallucination and to Improve Reproducibility in Scientific Text Generation Using Stochastic Optimisation Techniques 10 Nov 2024 · 0 repositories · arXiv:2411.06445
-
Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement 10 Nov 2024 · 1 repository · arXiv:2411.06558Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Clustering Algorithms and RAG Enhancing Semi-Supervised Text Classification with Large LLMs 9 Nov 2024 · 0 repositories · arXiv:2411.06175
-
Detecting Reference Errors in Scientific Literature with Large Language Models 9 Nov 2024 · 1 repository · arXiv:2411.06101
-
Exploring Knowledge Boundaries in Large Language Models for Retrieval Judgment 9 Nov 2024 · 0 repositories · arXiv:2411.06207
-
Improved intent classification based on context information using a windows-based approach 9 Nov 2024 · 0 repositories · arXiv:2411.06022
-
Leveraging Retrieval-Augmented Generation for Persian University Knowledge Retrieval 9 Nov 2024 · 0 repositories · arXiv:2411.06237
-
Robust Detection of LLM-Generated Text: A Comparative Analysis 9 Nov 2024 · 0 repositories · arXiv:2411.06248
-
Sufficient Context: A New Lens on Retrieval Augmented Generation Systems 9 Nov 2024 · 0 repositories · arXiv:2411.06037
-
AgentOps: Enabling Observability of LLM Agents 8 Nov 2024 · 1 repository · arXiv:2411.05285
-
Enhancing Visual Classification using Comparative Descriptors 8 Nov 2024 · 1 repository · arXiv:2411.05357
-
FinDVer: Explainable Claim Verification over Long and Hybrid-Content Financial Documents 8 Nov 2024 · 1 repository · arXiv:2411.05764
-
GPT Semantic Cache: Reducing LLM Costs and Latency via Semantic Embedding Caching 8 Nov 2024 · 0 repositories · arXiv:2411.05276
-
HeartBERT: A Self-Supervised ECG Embedding Model for Efficient and Effective Medical Signal Analysis 8 Nov 2024 · 1 repository · arXiv:2411.11896
-
IntellBot: Retrieval Augmented LLM Chatbot for Cyber Threat Knowledge Delivery 8 Nov 2024 · 1 repository · arXiv:2411.05442
-
Learning the rules of peptide self-assembly through data mining with large language models 8 Nov 2024 · 1 repository · arXiv:2411.05421
-
Multi-Document Financial Question Answering using LLMs 8 Nov 2024 · 0 repositories · arXiv:2411.07264
-
NeKo: Toward Post Recognition Generative Correction Large Language Models with Task-Oriented Experts 8 Nov 2024 · 0 repositories · arXiv:2411.05945
-
Qwen2.5-32B: Leveraging Self-Consistent Tool-Integrated Reasoning for Bengali Mathematical Olympiad Problem Solving 8 Nov 2024 · 0 repositories · arXiv:2411.05934
-
Sentiment Analysis of Cyberbullying Data in Social Media 8 Nov 2024 · 1 repository · arXiv:2411.05958
-
Adversarial Robustness of In-Context Learning in Transformers for Linear Regression 7 Nov 2024 · 0 repositories · arXiv:2411.05189
-
Best Practices for Distilling Large Language Models into BERT for Web Search Ranking 7 Nov 2024 · 0 repositories · arXiv:2411.04539
-
Deploying Large Language Models With Retrieval Augmented Generation 7 Nov 2024 · 1 repository · arXiv:2411.11895
-
Enhancing classroom teaching with LLMs and RAG 7 Nov 2024 · 0 repositories · arXiv:2411.04341
-
FineTuneBench: How well do commercial fine-tuning APIs infuse knowledge into LLMs? 7 Nov 2024 · 1 repository · arXiv:2411.05059
-
GPT-Guided Monte Carlo Tree Search for Symbolic Regression in Financial Fraud Detection 7 Nov 2024 · 0 repositories · arXiv:2411.04459
-
M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding 7 Nov 2024 · 0 repositories · arXiv:2411.04952
-
RetrieveGPT: Merging Prompts and Mathematical Models for Enhanced Code-Mixed Information Retrieval 7 Nov 2024 · 0 repositories · arXiv:2411.04752
-
Selecting Between BERT and GPT for Text Classification in Political Science Research 7 Nov 2024 · 0 repositories · arXiv:2411.05050
-
STAND-Guard: A Small Task-Adaptive Content Moderation Model 7 Nov 2024 · 0 repositories · arXiv:2411.05214
-
Words that Move Markets- Quantifying the Impact of RBI's Monetary Policy Communications on Indian Financial Market 7 Nov 2024 · 0 repositories · arXiv:2411.04808
-
A Comparative Study of Recent Large Language Models on Generating Hospital Discharge Summaries for Lung Cancer Patients 6 Nov 2024 · 0 repositories · arXiv:2411.03805
-
A Multilingual Sentiment Lexicon for Low-Resource Language Translation using Large Languages Models and Explainable AI 6 Nov 2024 · 0 repositories · arXiv:2411.04316
-
Advanced RAG Models with Graph Structures: Optimizing Complex Knowledge Reasoning and Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03572
-
Can Custom Models Learn In-Context? An Exploration of Hybrid Architecture Performance on In-Context Learning Tasks 6 Nov 2024 · 1 repository · arXiv:2411.03945
-
Fine-Grained Guidance for Retrievers: Leveraging LLMs' Feedback in Retrieval-Augmented Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03957
-
From Word Vectors to Multimodal Embeddings: Techniques, Applications, and Future Directions For Large Language Models 6 Nov 2024 · 0 repositories · arXiv:2411.05036
-
On-Device Emoji Classifier Trained with GPT-based Data Augmentation for a Mobile Keyboard 6 Nov 2024 · 0 repositories · arXiv:2411.05031
-
PhDGPT: Introducing a psychometric and linguistic dataset about how large language models perceive graduate students and professors in psychology 6 Nov 2024 · 0 repositories · arXiv:2411.10473
-
Prompt Engineering Using GPT for Word-Level Code-Mixed Language Identification in Low-Resource Dravidian Languages 6 Nov 2024 · 0 repositories · arXiv:2411.04025
-
RAGulator: Lightweight Out-of-Context Detectors for Grounded Text Generation 6 Nov 2024 · 0 repositories · arXiv:2411.03920
-
Towards Interpreting Language Models: A Case Study in Multi-Hop Reasoning 6 Nov 2024 · 1 repository · arXiv:2411.05037Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Understanding the Effects of Human-written Paraphrases in LLM-generated Text Detection 6 Nov 2024 · 1 repository · arXiv:2411.03806
-
YouTube Comments Decoded: Leveraging LLMs for Low Resource Language Classification 6 Nov 2024 · 0 repositories · arXiv:2411.05039
-
Automatic Generation of Question Hints for Mathematics Problems using Large Language Models in Educational Technology 5 Nov 2024 · 0 repositories · arXiv:2411.03495
-
Enhancing Transformer Training Efficiency with Dynamic Dropout 5 Nov 2024 · 0 repositories · arXiv:2411.03236
-
Exploring the Benefits of Domain-Pretraining of Generative Large Language Models for Chemistry 5 Nov 2024 · 0 repositories · arXiv:2411.03542
-
HtmlRAG: HTML is Better Than Plain Text for Modeling Retrieved Knowledge in RAG Systems 5 Nov 2024 · 1 repository · arXiv:2411.02959
-
LASER: Attention with Exponential Transformation 5 Nov 2024 · 0 repositories · arXiv:2411.03493
-
Long Context RAG Performance of Large Language Models 5 Nov 2024 · 0 repositories · arXiv:2411.03538
-
PersianRAG: A Retrieval-Augmented Generation System for Persian Language 5 Nov 2024 · 0 repositories · arXiv:2411.02832
-
A Comparative Analysis of Counterfactual Explanation Methods for Text Classifiers 4 Nov 2024 · 0 repositories · arXiv:2411.02643
-
Advancements and limitations of LLMs in replicating human color-word associations 4 Nov 2024 · 0 repositories · arXiv:2411.02116
-
Ask, and it shall be given: On the Turing completeness of prompting 4 Nov 2024 · 1 repository · arXiv:2411.01992
-
Can Language Models Enable In-Context Database? 4 Nov 2024 · 0 repositories · arXiv:2411.01807
-
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast 4 Nov 2024 · 0 repositories · arXiv:2411.02318
-
Grounding Emotional Descriptions to Electrovibration Haptic Signals 4 Nov 2024 · 0 repositories · arXiv:2411.02118
-
MdEval: Massively Multilingual Code Debugging 4 Nov 2024 · 0 repositories · arXiv:2411.02310
-
RAGViz: Diagnose and Visualize Retrieval-Augmented Generation 4 Nov 2024 · 1 repository · arXiv:2411.01751
-
TeleOracle: Fine-Tuned Retrieval-Augmented Generation with Long-Context Support for Network 4 Nov 2024 · 1 repository · arXiv:2411.02617
-
Wave Network: An Ultra-Small Language Model 4 Nov 2024 · 0 repositories · arXiv:2411.02674
-
Data Extraction Attacks in Retrieval-Augmented Generation via Backdoors 3 Nov 2024 · 0 repositories · arXiv:2411.01705
-
Enriching Tabular Data with Contextual LLM Embeddings: A Comprehensive Ablation Study for Ensemble Classifiers 3 Nov 2024 · 0 repositories · arXiv:2411.01645
-
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models 3 Nov 2024 · 2 repositories · arXiv:2411.01713Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Can Large Language Model Predict Employee Attrition? 2 Nov 2024 · 0 repositories · arXiv:2411.01353
-
Enhancing Neural Network Interpretability with Feature-Aligned Sparse Autoencoders 2 Nov 2024 · 1 repository · arXiv:2411.01220Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AttackQA: Development and Adoption of a Dataset for Assisting Cybersecurity Operations using Fine-tuned and Open-Source LLMs 1 Nov 2024 · 0 repositories · arXiv:2411.01073
-
CORAG: A Cost-Constrained Retrieval Optimization System for Retrieval-Augmented Generation 1 Nov 2024 · 0 repositories · arXiv:2411.00744
-
Evaluating the Impact of Lab Test Results on Large Language Models Generated Differential Diagnoses from Clinical Case Vignettes 1 Nov 2024 · 0 repositories · arXiv:2411.02523
-
LLM-Ref: Enhancing Reference Handling in Technical Writing with Large Language Models 1 Nov 2024 · 0 repositories · arXiv:2411.00294
-
LLMs: A Game-Changer for Software Engineers? 1 Nov 2024 · 0 repositories · arXiv:2411.00932
-
Provenance: A Light-weight Fact-checker for Retrieval Augmented LLM Generation Output 1 Nov 2024 · 0 repositories · arXiv:2411.01022
-
Rationale-Guided Retrieval Augmented Generation for Medical Question Answering 1 Nov 2024 · 1 repository · arXiv:2411.00300Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Towards Multi-Source Retrieval-Augmented Generation via Synergizing Reasoning and Preference-Driven Retrieval 1 Nov 2024 · 0 repositories · arXiv:2411.00689
-
Analyzing & Reducing the Need for Learning Rate Warmup in GPT Training 31 Oct 2024 · 0 repositories · arXiv:2410.23922
-
Automating Quantum Software Maintenance: Flakiness Detection and Root Cause Analysis 31 Oct 2024 · 0 repositories · arXiv:2410.23578
-
Global Convergence in Training Large-Scale Transformers 31 Oct 2024 · 0 repositories · arXiv:2410.23610
-
JudgeRank: Leveraging Large Language Models for Reasoning-Intensive Reranking 31 Oct 2024 · 0 repositories · arXiv:2411.00142
-
LEAF: Learning and Evaluation Augmented by Fact-Checking to Improve Factualness in Large Language Models 31 Oct 2024 · 0 repositories · arXiv:2410.23526
-
Responsible Retrieval Augmented Generation for Climate Decision Making from Documents 31 Oct 2024 · 0 repositories · arXiv:2410.23902
-
SelfCodeAlign: Self-Alignment for Code Generation 31 Oct 2024 · 2 repositories · arXiv:2410.24198Syntology official (archive's flag): 9 ran · 30 ran (of which 3 constructed an object rather than computing a result; 22 with no instrument failure: 1 honoured, 0 violated, 21 with no contract checked; 8 where Syntology's instrument failed) · 7 unverified (of 37 harvested samples)
-
Weight decay induces low-rank attention layers 31 Oct 2024 · 0 repositories · arXiv:2410.23819
-
A Comprehensive Study on Quantization Techniques for Large Language Models 30 Oct 2024 · 0 repositories · arXiv:2411.02530
-
CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation 30 Oct 2024 · 1 repository · arXiv:2410.23090
-
Eliciting Critical Reasoning in Retrieval-Augmented Language Models via Contrastive Explanations 30 Oct 2024 · 0 repositories · arXiv:2410.22874
-
Emotional RAG: Enhancing Role-Playing Agents through Emotional Retrieval 30 Oct 2024 · 1 repository · arXiv:2410.23041
-
HijackRAG: Hijacking Attacks against Retrieval-Augmented Large Language Models 30 Oct 2024 · 0 repositories · arXiv:2410.22832
-
ProTransformer: Robustify Transformers via Plug-and-Play Paradigm 30 Oct 2024 · 1 repository · arXiv:2410.23182
-
Retrieval-Augmented Generation with Estimation of Source Reliability 30 Oct 2024 · 0 repositories · arXiv:2410.22954
-
Semantic Enrichment of the Quantum Cascade Laser Properties in Text- A Knowledge Graph Generation Approach 30 Oct 2024 · 1 repository · arXiv:2410.22996
-
Long²RAG: Evaluating Long-Context & Long-Form Retrieval-Augmented Generation with Key Point Recall 30 Oct 2024 · 0 repositories · arXiv:2410.23000
-
Abrupt Learning in Transformers: A Case Study on Matrix Completion 29 Oct 2024 · 0 repositories · arXiv:2410.22244Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Beyond Text: Optimizing RAG with Multimodal Inputs for Industrial Applications 29 Oct 2024 · 1 repository · arXiv:2410.21943Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
CFSafety: Comprehensive Fine-grained Safety Assessment for LLMs 29 Oct 2024 · 0 repositories · arXiv:2410.21695
-
Coupling quantum-like cognition with the neuronal networks within generalized probability theory 29 Oct 2024 · 0 repositories · arXiv:2411.00036
-
FactBench: A Dynamic Benchmark for In-the-Wild Language Model Factuality Evaluation 29 Oct 2024 · 0 repositories · arXiv:2410.22257
-
Meta-Learning Adaptable Foundation Models 29 Oct 2024 · 0 repositories · arXiv:2410.22264
-
Sequential choice in ordered bundles 29 Oct 2024 · 0 repositories · arXiv:2410.21670
-
A Simple Yet Effective Corpus Construction Framework for Indonesian Grammatical Error Correction 28 Oct 2024 · 1 repository · arXiv:2410.20838
-
AutoRAG: Automated Framework for optimization of Retrieval Augmented Generation Pipeline 28 Oct 2024 · 2 repositories · arXiv:2410.20878