Methods › General › Regularization › Weight Decay › Papers, page 27
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 27 of 108: papers 2,601 to 2,700 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BERTs are Generative In-Context Learners 7 Jun 2024 · 1 repository · arXiv:2406.04823Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 26 harvested samples)
-
Corpus Poisoning via Approximate Greedy Gradient Descent 7 Jun 2024 · 1 repository · arXiv:2406.05087
-
CRAG -- Comprehensive RAG Benchmark 7 Jun 2024 · 2 repositories · arXiv:2406.04744Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents 7 Jun 2024 · 2 repositories · arXiv:2406.06613Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Large Generative Graph Models 7 Jun 2024 · 0 repositories · arXiv:2406.05109
-
Low-Resource Cross-Lingual Summarization through Few-Shot Learning with Large Language Models 7 Jun 2024 · 0 repositories · arXiv:2406.04630
-
Multi-Head RAG: Solving Multi-Aspect Problems with LLMs 7 Jun 2024 · 2 repositories · arXiv:2406.05085
-
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation 7 Jun 2024 · 1 repository · arXiv:2406.05213Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
VTrans: Accelerating Transformer Compression with Variational Information Bottleneck based Pruning 7 Jun 2024 · 0 repositories · arXiv:2406.05276
-
A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions 6 Jun 2024 · 0 repositories · arXiv:2406.03712
-
Do Language Models Understand Morality? Towards a Robust Detection of Moral Content 6 Jun 2024 · 1 repository · arXiv:2406.04143
-
Empirical Guidelines for Deploying LLMs onto Resource-constrained Edge Devices 6 Jun 2024 · 0 repositories · arXiv:2406.03777
-
HORAE: A Domain-Agnostic Language for Automated Service Regulation 6 Jun 2024 · 1 repository · arXiv:2406.06600
-
LLMEmbed: Rethinking Lightweight LLM's Genuine Function in Text Classification 6 Jun 2024 · 1 repository · arXiv:2406.03725
-
Simplified and Generalized Masked Diffusion for Discrete Data 6 Jun 2024 · 1 repository · arXiv:2406.04329Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 15 unverified (of 21 harvested samples)
-
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech 6 Jun 2024 · 1 repository · arXiv:2406.03953
-
Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data 6 Jun 2024 · 2 repositories · arXiv:2406.03736Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Automating Turkish Educational Quiz Generation Using Large Language Models 5 Jun 2024 · 4 repositories · arXiv:2406.03397
-
Exact Conversion of In-Context Learning to Model Weights in Linearized-Attention Transformers 5 Jun 2024 · 0 repositories · arXiv:2406.02847
-
Exploring Multilingual Large Language Models for Enhanced TNM classification of Radiology Report in lung cancer staging 5 Jun 2024 · 0 repositories · arXiv:2406.06591
-
Missci: Reconstructing Fallacies in Misrepresented Science 5 Jun 2024 · 2 repositories · arXiv:2406.03181
-
PoLYTC: a novel BERT-based classifier to detect political leaning of YouTube videos based on their titles 5 Jun 2024 · 1 repository
-
RICo: Reddit ideological communities 5 Jun 2024 · 1 repository
-
StatBot.Swiss: Bilingual Open Data Exploration in Natural Language 5 Jun 2024 · 0 repositories · arXiv:2406.03170
-
The Good, the Bad, and the Hulk-like GPT: Analyzing Emotional Decisions of Large Language Models in Cooperation and Bargaining Games 5 Jun 2024 · 0 repositories · arXiv:2406.03299
-
Too Big to Fail: Larger Language Models are Disproportionately Resilient to Induction of Dementia-Related Linguistic Anomalies 5 Jun 2024 · 1 repository · arXiv:2406.02830
-
Chain of Agents: Large Language Models Collaborating on Long-Context Tasks 4 Jun 2024 · 0 repositories · arXiv:2406.02818
-
CheckEmbed: Effective Verification of LLM Solutions to Open-Ended Tasks 4 Jun 2024 · 1 repository · arXiv:2406.02524
-
Large Language Model-Enabled Multi-Agent Manufacturing Systems 4 Jun 2024 · 0 repositories · arXiv:2406.01893
-
OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step 4 Jun 2024 · 0 repositories · arXiv:2406.06576
-
Probing the Category of Verbal Aspect in Transformer Language Models 4 Jun 2024 · 0 repositories · arXiv:2406.02335
-
Randomized Geometric Algebra Methods for Convex Neural Networks 4 Jun 2024 · 1 repository · arXiv:2406.02806Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SMS Spam Detection and Classification to Combat Abuse in Telephone Networks Using Natural Language Processing 4 Jun 2024 · 0 repositories · arXiv:2406.06578
-
Synergetic Event Understanding: A Collaborative Approach to Cross-Document Event Coreference Resolution with Large Language Models 4 Jun 2024 · 1 repository · arXiv:2406.02148
-
Towards Effective Time-Aware Language Representation: Exploring Enhanced Temporal Understanding in Language Models 4 Jun 2024 · 0 repositories · arXiv:2406.01863
-
Annotation Guidelines-Based Knowledge Augmentation: Towards Enhancing Large Language Models for Educational Text Classification 3 Jun 2024 · 0 repositories · arXiv:2406.00954
-
Ask-EDA: A Design Assistant Empowered by LLM, Hybrid RAG and Abbreviation De-hallucination 3 Jun 2024 · 0 repositories · arXiv:2406.06575
-
BadRAG: Identifying Vulnerabilities in Retrieval Augmented Generation of Large Language Models 3 Jun 2024 · 0 repositories · arXiv:2406.00083
-
FactGenius: Combining Zero-Shot Prompting and Fuzzy Relation Mining to Improve Fact Verification with Knowledge Graphs 3 Jun 2024 · 1 repository · arXiv:2406.01311
-
Focus on the Core: Efficient Attention via Pruned Token Compression for Document Classification 3 Jun 2024 · 0 repositories · arXiv:2406.01283
-
In-Context Learning of Physical Properties: Few-Shot Adaptation to Out-of-Distribution Molecular Graphs 3 Jun 2024 · 0 repositories · arXiv:2406.01808
-
Luna: An Evaluation Foundation Model to Catch Language Model Hallucinations with High Accuracy and Low Cost 3 Jun 2024 · 0 repositories · arXiv:2406.00975
-
Natural Language Interaction with a Household Electricity Knowledge-based Digital Twin 3 Jun 2024 · 0 repositories · arXiv:2406.06566
-
SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning 3 Jun 2024 · 1 repository · arXiv:2406.01006Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples)
-
SoccerRAG: Multimodal Soccer Information Retrieval via Natural Queries 3 Jun 2024 · 1 repository · arXiv:2406.01273
-
SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models 3 Jun 2024 · 1 repository · arXiv:2406.01584
-
Superhuman performance in urology board questions by an explainable large language model enabled for context integration of the European Association of Urology guidelines: the UroBot study 3 Jun 2024 · 0 repositories · arXiv:2406.01428
-
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding 3 Jun 2024 · 0 repositories · arXiv:2406.01306
-
A Theory for Token-Level Harmonization in Retrieval-Augmented Generation 3 Jun 2024 · 0 repositories · arXiv:2406.00944
-
Value Improved Actor Critic Algorithms 3 Jun 2024 · 0 repositories · arXiv:2406.01423
-
Applying Fine-Tuned LLMs for Reducing Data Needs in Load Profile Analysis 2 Jun 2024 · 0 repositories · arXiv:2406.02479
-
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction 2 Jun 2024 · 1 repository · arXiv:2406.00755Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
FOCUS: Forging Originality through Contrastive Use in Self-Plagiarism for Language Models 2 Jun 2024 · 0 repositories · arXiv:2406.00839
-
Formality Style Transfer in Persian 2 Jun 2024 · 0 repositories · arXiv:2406.00867
-
Pretrained Hybrids with MAD Skills 2 Jun 2024 · 0 repositories · arXiv:2406.00894
-
Domain-specific ReAct for physics-integrated iterative modeling: A case study of LLM agents for gas path analysis of gas turbines 1 Jun 2024 · 0 repositories · arXiv:2406.07572
-
An Evaluation Benchmark for Autoformalization in Lean4 1 Jun 2024 · 0 repositories · arXiv:2406.06555
-
Beyond Metrics: Evaluating LLMs' Effectiveness in Culturally Nuanced, Low-Resource Real-World Scenarios 1 Jun 2024 · 0 repositories · arXiv:2406.00343
-
CASE: Efficient Curricular Data Pre-training for Building Assistive Psychology Expert Models 1 Jun 2024 · 1 repository · arXiv:2406.00314
-
Mix-of-Granularity: Optimize the Chunking Granularity for Retrieval-Augmented Generation 1 Jun 2024 · 1 repository · arXiv:2406.00456
-
Pseudo-label Based Domain Adaptation for Zero-Shot Text Steganalysis 1 Jun 2024 · 0 repositories · arXiv:2406.18565
-
RoBERTa-BiLSTM: A Context-Aware Hybrid Model for Sentiment Analysis 1 Jun 2024 · 1 repository · arXiv:2406.00367
-
A comparison of correspondence analysis with PMI-based word embedding methods 31 May 2024 · 1 repository · arXiv:2405.20895
-
Bi-Directional Transformers vs. word2vec: Discovering Vulnerabilities in Lifted Compiled Code 31 May 2024 · 0 repositories · arXiv:2405.20611
-
Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial Training 31 May 2024 · 1 repository · arXiv:2405.20978Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Generative AI Voting: Fair Collective Choice is Resilient to LLM Biases and Inconsistencies 31 May 2024 · 1 repository · arXiv:2406.11871
-
Hard Cases Detection in Motion Prediction by Vision-Language Foundation Models 31 May 2024 · 1 repository · arXiv:2405.20991
-
Large Language Models are Zero-Shot Next Location Predictors 31 May 2024 · 1 repository · arXiv:2405.20962Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
LOLAMEME: Logic, Language, Memory, Mechanistic Framework 31 May 2024 · 0 repositories · arXiv:2406.02592
-
Multilingual Text Style Transfer: Datasets & Models for Indian Languages 31 May 2024 · 2 repositories · arXiv:2405.20805
-
RAG Does Not Work for Enterprises 31 May 2024 · 0 repositories · arXiv:2406.04369
-
Retrieval Meets Reasoning: Even High-school Textbook Knowledge Benefits Multimodal Reasoning 31 May 2024 · 0 repositories · arXiv:2405.20834
-
The Point of View of a Sentiment: Towards Clinician Bias Detection in Psychiatric Notes 31 May 2024 · 0 repositories · arXiv:2405.20582
-
ANAH: Analytical Annotation of Hallucinations in Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20315Syntology official: no sample here; runs from other or unrecorded repositories · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
AutoBreach: Universal and Adaptive Jailbreaking with Efficient Wordplay-Guided Optimization 30 May 2024 · 0 repositories · arXiv:2405.19668
-
Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation 30 May 2024 · 0 repositories · arXiv:2405.20092
-
Ensemble Model With Bert,Roberta and Xlnet For Molecular property prediction 30 May 2024 · 0 repositories · arXiv:2406.06553
-
GNN-RAG: Graph Neural Retrieval for Large Language Model Reasoning 30 May 2024 · 1 repository · arXiv:2405.20139Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Is My Data in Your Retrieval Database? Membership Inference Attacks Against Retrieval Augmented Generation 30 May 2024 · 0 repositories · arXiv:2405.20446
-
KerasCV and KerasNLP: Vision and Language Power-Ups 30 May 2024 · 0 repositories · arXiv:2405.20247
-
Knowledge Graph Tuning: Real-time Large Language Model Personalization based on Human Feedback 30 May 2024 · 0 repositories · arXiv:2405.19686
-
LLaMEA: A Large Language Model Evolutionary Algorithm for Automatically Generating Metaheuristics 30 May 2024 · 2 repositories · arXiv:2405.20132Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
One Token Can Help! Learning Scalable and Pluggable Virtual Tokens for Retrieval-Augmented Large Language Models 30 May 2024 · 2 repositories · arXiv:2405.19670
-
Phantom: General Trigger Attacks on Retrieval Augmented Language Generation 30 May 2024 · 0 repositories · arXiv:2405.20485
-
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs 30 May 2024 · 0 repositories · arXiv:2405.20179
-
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation 30 May 2024 · 0 repositories · arXiv:2405.19701
-
Student Answer Forecasting: Transformer-Driven Answer Choice Prediction for Language Learning 30 May 2024 · 1 repository · arXiv:2405.20079
-
Towards Ontology-Enhanced Representation Learning for Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20527
-
A Multi-Source Retrieval Question Answering Framework Based on RAG 29 May 2024 · 0 repositories · arXiv:2405.19207
-
Beyond Agreement: Diagnosing the Rationale Alignment of Automated Essay Scoring Methods based on Linguistically-informed Counterfactuals 29 May 2024 · 1 repository · arXiv:2405.19433
-
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension 29 May 2024 · 0 repositories · arXiv:2405.18682
-
CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control 29 May 2024 · 1 repository · arXiv:2405.18727Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
Efficient Model-agnostic Alignment via Bayesian Persuasion 29 May 2024 · 0 repositories · arXiv:2405.18718
-
LMO-DP: Optimizing the Randomization Mechanism for Differentially Private Fine-Tuning (Large) Language Models 29 May 2024 · 0 repositories · arXiv:2405.18776
-
MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series 29 May 2024 · 1 repository · arXiv:2405.19327Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
STAT: Shrinking Transformers After Training 29 May 2024 · 0 repositories · arXiv:2406.00061
-
Toward Conversational Agents with Context and Time Sensitive Long-term Memory 29 May 2024 · 1 repository · arXiv:2406.00057
-
Two-Layer Retrieval-Augmented Generation Framework for Low-Resource Medical Question Answering Using Reddit Data: Proof-of-Concept Study 29 May 2024 · 0 repositories · arXiv:2405.19519
-
Aligning to Thousands of Preferences via System Message Generalization 28 May 2024 · 1 repository · arXiv:2405.17977Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
An Empirical Analysis on Large Language Models in Debate Evaluation 28 May 2024 · 1 repository · arXiv:2406.00050Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)