Methods › General › Regularization › Weight Decay › Papers, page 26
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 26 of 108: papers 2,501 to 2,600 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Intermediate Distillation: Data-Efficient Distillation from Black-Box LLMs for Information Retrieval 18 Jun 2024 · 0 repositories · arXiv:2406.12169
-
IPEval: A Bilingual Intellectual Property Agency Consultation Evaluation Benchmark for Large Language Models 18 Jun 2024 · 1 repository · arXiv:2406.12386
-
PlanRAG: A Plan-then-Retrieval Augmented Generation for Generative Large Language Models as Decision Makers 18 Jun 2024 · 1 repository · arXiv:2406.12430Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Retrieval-Augmented Generation for Generative Artificial Intelligence in Medicine 18 Jun 2024 · 0 repositories · arXiv:2406.12449
-
RichRAG: Crafting Rich Responses for Multi-faceted Queries in Retrieval-Augmented Generation 18 Jun 2024 · 0 repositories · arXiv:2406.12566
-
Towards a Client-Centered Assessment of LLM Therapists by Client Simulation 18 Jun 2024 · 1 repository · arXiv:2406.12266Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Unified Active Retrieval for Retrieval Augmented Generation 18 Jun 2024 · 1 repository · arXiv:2406.12534
-
UrbanLLM: Autonomous Urban Activity Planning and Management with Large Language Models 18 Jun 2024 · 0 repositories · arXiv:2406.12360
-
Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping 18 Jun 2024 · 0 repositories · arXiv:2406.12679
-
What Makes Two Language Models Think Alike? 18 Jun 2024 · 0 repositories · arXiv:2406.12620
-
"You Gotta be a Doctor, Lin": An Investigation of Name-Based Bias of Large Language Models in Employment Recommendations 18 Jun 2024 · 0 repositories · arXiv:2406.12232
-
Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance 17 Jun 2024 · 0 repositories · arXiv:2406.11139
-
Building another Spanish dictionary, this time with GPT-4 17 Jun 2024 · 1 repository · arXiv:2406.11218
-
CrAM: Credibility-Aware Attention Modification in LLMs for Combating Misinformation in RAG 17 Jun 2024 · 1 repository · arXiv:2406.11497
-
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting 17 Jun 2024 · 0 repositories · arXiv:2406.11661
-
SeRTS: Self-Rewarding Tree Search for Biomedical Retrieval-Augmented Generation 17 Jun 2024 · 0 repositories · arXiv:2406.11258
-
Enhancing Text Classification through LLM-Driven Active Learning and Human Annotation 17 Jun 2024 · 1 repository · arXiv:2406.12114
-
Estimating the Increase in Emissions caused by AI-augmented Search 17 Jun 2024 · 0 repositories · arXiv:2407.16894
-
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications? 17 Jun 2024 · 1 repository · arXiv:2406.11402
-
Evaluating the Efficacy of Open-Source LLMs in Enterprise-Specific RAG Systems: A Comparative Study of Performance and Scalability 17 Jun 2024 · 1 repository · arXiv:2406.11424
-
Exploring Safety-Utility Trade-Offs in Personalized Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11107
-
Fine-Tuning or Fine-Failing? Debunking Performance Myths in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11201
-
GPT-Powered Elicitation Interview Script Generator for Requirements Engineering Training 17 Jun 2024 · 0 repositories · arXiv:2406.11439
-
Improving Multi-Agent Debate with Sparse Communication Topology 17 Jun 2024 · 0 repositories · arXiv:2406.11776
-
Investigating Annotator Bias in Large Language Models for Hate Speech Detection 17 Jun 2024 · 3 repositories · arXiv:2406.11109
-
Iterative Utility Judgment Framework via LLMs Inspired by Relevance in Philosophy 17 Jun 2024 · 0 repositories · arXiv:2406.11290
-
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.15484
-
Promises, Outlooks and Challenges of Diffusion Language Modeling 17 Jun 2024 · 0 repositories · arXiv:2406.11473
-
R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models 17 Jun 2024 · 1 repository · arXiv:2406.11681
-
Satyrn: A Platform for Analytics Augmented Generation 17 Jun 2024 · 1 repository · arXiv:2406.12069
-
Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99% 17 Jun 2024 · 1 repository · arXiv:2406.11837Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment 17 Jun 2024 · 0 repositories · arXiv:2406.11285
-
Refiner: Restructure Retrieval Content Efficiently to Advance Question-Answering Capabilities 17 Jun 2024 · 1 repository · arXiv:2406.11357Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TRACE the Evidence: Constructing Knowledge-Grounded Reasoning Chains for Retrieval-Augmented Generation 17 Jun 2024 · 2 repositories · arXiv:2406.11460Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions 17 Jun 2024 · 1 repository · arXiv:2406.12058Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing Supermarket Robot Interaction: A Multi-Level LLM Conversational Interface for Handling Diverse Customer Intents 16 Jun 2024 · 0 repositories · arXiv:2406.11047
-
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning 16 Jun 2024 · 0 repositories · arXiv:2406.10834
-
Generating Tables from the Parametric Knowledge of Language Models 16 Jun 2024 · 1 repository · arXiv:2406.10922
-
Grading Massive Open Online Courses Using Large Language Models 16 Jun 2024 · 0 repositories · arXiv:2406.11102
-
KGPA: Robustness Evaluation for Large Language Models via Cross-Domain Knowledge Graphs 16 Jun 2024 · 1 repository · arXiv:2406.10802
-
Large Language Models for Automatic Milestone Detection in Group Discussions 16 Jun 2024 · 0 repositories · arXiv:2406.10842
-
Predicting the Understandability of Computational Notebooks through Code Metrics Analysis 16 Jun 2024 · 1 repository · arXiv:2406.10989
-
ShareLoRA: Parameter Efficient and Robust Large Language Model Fine-tuning via Shared Low-Rank Adaptation 16 Jun 2024 · 1 repository · arXiv:2406.10785Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
ViD-GPT: Introducing GPT-style Autoregressive Generation in Video Diffusion Models 16 Jun 2024 · 1 repository · arXiv:2406.10981
-
A Comprehensive Survey of Foundation Models in Medicine 15 Jun 2024 · 0 repositories · arXiv:2406.10729
-
Beyond Raw Videos: Understanding Edited Videos with Large Multimodal Model 15 Jun 2024 · 1 repository · arXiv:2406.10484
-
MINT: a Multi-modal Image and Narrative Text Dubbing Dataset for Foley Audio Content Planning and Generation 15 Jun 2024 · 1 repository · arXiv:2406.10591
-
We Care: Multimodal Depression Detection and Knowledge Infused Mental Health Therapeutic Response Generation 15 Jun 2024 · 0 repositories · arXiv:2406.10561
-
Bag of Lies: Robustness in Continuous Pre-training BERT 14 Jun 2024 · 0 repositories · arXiv:2406.09967
-
Exploring the Correlation between Human and Machine Evaluation of Simultaneous Speech Translation 14 Jun 2024 · 0 repositories · arXiv:2406.10091
-
HIRO: Hierarchical Information Retrieval Optimization 14 Jun 2024 · 1 repository · arXiv:2406.09979
-
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models 14 Jun 2024 · 1 repository · arXiv:2406.10130
-
Towards Efficient Pareto Set Approximation via Mixture of Experts Based Model Fusion 14 Jun 2024 · 1 repository · arXiv:2406.09770
-
A More Practical Approach to Machine Unlearning 13 Jun 2024 · 0 repositories · arXiv:2406.09391
-
Analyzing Gender Polarity in Short Social Media Texts with BERT: The Role of Emojis and Emoticons 13 Jun 2024 · 0 repositories · arXiv:2406.09573
-
BPE-knockout: Pruning Pre-existing BPE Tokenisers with Backwards-compatible Morphological Semi-supervision 13 Jun 2024 · 1 repository
-
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination 13 Jun 2024 · 0 repositories · arXiv:2406.08818
-
Optimizing Large Model Training through Overlapped Activation Recomputation 13 Jun 2024 · 0 repositories · arXiv:2406.08756
-
PC-LoRA: Low-Rank Adaptation for Progressive Model Compression with Knowledge Distillation 13 Jun 2024 · 0 repositories · arXiv:2406.09117
-
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models 13 Jun 2024 · 0 repositories · arXiv:2406.09519
-
Ad Auctions for LLMs via Retrieval Augmented Generation 12 Jun 2024 · 0 repositories · arXiv:2406.09459
-
Exploring Fact Memorization and Style Imitation in LLMs Using QLoRA: An Experimental Study and Quality Assessment Methods 12 Jun 2024 · 0 repositories · arXiv:2406.08582
-
FaithFill: Faithful Inpainting for Object Completion Using a Single Reference Image 12 Jun 2024 · 0 repositories · arXiv:2406.07865
-
Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification 12 Jun 2024 · 1 repository · arXiv:2406.08660Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
How well it works: Benchmarking performance of GPT models on medical natural language processing tasks 12 Jun 2024 · 0 repositories
-
Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests 12 Jun 2024 · 0 repositories · arXiv:2406.07794
-
Label-aware Hard Negative Sampling Strategies with Momentum Contrastive Learning for Implicit Hate Speech Detection 12 Jun 2024 · 1 repository · arXiv:2406.07886
-
Leveraging Large Language Models for Web Scraping 12 Jun 2024 · 0 repositories · arXiv:2406.08246
-
Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation 12 Jun 2024 · 0 repositories · arXiv:2406.08328
-
Tailoring Generative AI Chatbots for Multiethnic Communities in Disaster Preparedness Communication: Extending the CASA Paradigm 12 Jun 2024 · 1 repository · arXiv:2406.08411
-
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis 11 Jun 2024 · 0 repositories · arXiv:2406.10273
-
Bilingual Sexism Classification: Fine-Tuned XLM-RoBERTa and GPT-3.5 Few-Shot Learning 11 Jun 2024 · 0 repositories · arXiv:2406.07287
-
COVID-19 Twitter Sentiment Classification Using Hybrid Deep Learning Model Based on Grid Search Methodology 11 Jun 2024 · 0 repositories · arXiv:2406.10266
-
DR-RAG: Applying Dynamic Document Relevance to Retrieval-Augmented Generation for Question-Answering 11 Jun 2024 · 0 repositories · arXiv:2406.07348
-
Flextron: Many-in-One Flexible Large Language Model 11 Jun 2024 · 0 repositories · arXiv:2406.10260
-
Multi-objective Reinforcement learning from AI Feedback 11 Jun 2024 · 1 repository · arXiv:2406.07295
-
Multimodal Belief Prediction 11 Jun 2024 · 1 repository · arXiv:2406.07466
-
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language 11 Jun 2024 · 0 repositories · arXiv:2406.08519
-
Unused information in token probability distribution of generative LLM: improving LLM reading comprehension through calculation of expected values 11 Jun 2024 · 1 repository · arXiv:2406.10267
-
AGB-DE: A Corpus for the Automated Legal Assessment of Clauses in German Consumer Contracts 10 Jun 2024 · 1 repository · arXiv:2406.06809
-
Compute Better Spent: Replacing Dense Layers with Structured Matrices 10 Jun 2024 · 1 repository · arXiv:2406.06248Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
In-Context Learning and Fine-Tuning GPT for Argument Mining 10 Jun 2024 · 1 repository · arXiv:2406.06699
-
Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing 10 Jun 2024 · 0 repositories · arXiv:2406.06723
-
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching 10 Jun 2024 · 0 repositories · arXiv:2406.06799
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection 10 Jun 2024 · 0 repositories · arXiv:2406.06663
-
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs 10 Jun 2024 · 0 repositories · arXiv:2406.10251
-
UMBRELA: UMbrela is the (Open-Source Reproduction of the) Bing RELevance Assessor 10 Jun 2024 · 1 repository · arXiv:2406.06519
-
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation 9 Jun 2024 · 2 repositories · arXiv:2406.05654Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Hidden Holes: topological aspects of language models 9 Jun 2024 · 0 repositories · arXiv:2406.05798
-
Machine Against the RAG: Jamming Retrieval-Augmented Generation with Blocker Documents 9 Jun 2024 · 0 repositories · arXiv:2406.05870
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering 9 Jun 2024 · 0 repositories · arXiv:2406.05845
-
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation 9 Jun 2024 · 1 repository · arXiv:2406.05794
-
Text2VP: Generative AI for Visual Programming and Parametric Modeling 9 Jun 2024 · 0 repositories · arXiv:2407.07732
-
Advancing Semantic Textual Similarity Modeling: A Regression Framework with Translated ReLU and Smooth K2 Loss 8 Jun 2024 · 2 repositories · arXiv:2406.05326
-
Critical Phase Transition in Large Language Models 8 Jun 2024 · 0 repositories · arXiv:2406.05335
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts 8 Jun 2024 · 0 repositories · arXiv:2406.05569
-
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature 8 Jun 2024 · 0 repositories · arXiv:2406.05431
-
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner 8 Jun 2024 · 0 repositories · arXiv:2406.05498
-
VP-LLM: Text-Driven 3D Volume Completion with Large Language Models through Patchification 8 Jun 2024 · 0 repositories · arXiv:2406.05543
-
BAMO at SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense 7 Jun 2024 · 1 repository · arXiv:2406.04947