Methods › General › Regularization › Weight Decay › Papers, page 31
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 31 of 108: papers 3,001 to 3,100 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Reducing hallucination in structured outputs via Retrieval-Augmented Generation 12 Apr 2024 · 0 repositories · arXiv:2404.08189
-
Revisiting Code Similarity Evaluation with Abstract Syntax Tree Edit Distance 12 Apr 2024 · 0 repositories · arXiv:2404.08817
-
Small Models Are (Still) Effective Cross-Domain Argument Extractors 12 Apr 2024 · 1 repository · arXiv:2404.08579
-
AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs 11 Apr 2024 · 1 repository · arXiv:2404.07921
-
Generative Information Retrieval Evaluation 11 Apr 2024 · 0 repositories · arXiv:2404.08137
-
On Training Data Influence of GPT Models 11 Apr 2024 · 2 repositories · arXiv:2404.07840Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Rumour Evaluation with Very Large Language Models 11 Apr 2024 · 1 repository · arXiv:2404.16859
-
Emotion-cause pair extraction method based on multi-granularity information and multi-module interaction 10 Apr 2024 · 0 repositories · arXiv:2404.06812
-
Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness 10 Apr 2024 · 1 repository · arXiv:2404.06714Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation 10 Apr 2024 · 1 repository · arXiv:2404.06809
-
Simpler becomes Harder: Do LLMs Exhibit a Coherent Behavior on Simplified Corpora? 10 Apr 2024 · 1 repository · arXiv:2404.06838
-
Superposition Prompting: Improving and Accelerating Retrieval-Augmented Generation 10 Apr 2024 · 1 repository · arXiv:2404.06910
-
Heuristic-enhanced Candidates Selection strategy for GPTs tackle Few-Shot Aspect-Based Sentiment Analysis 9 Apr 2024 · 1 repository · arXiv:2404.06063
-
Characterizing Multimodal Long-form Summarization: A Case Study on Financial Reports 9 Apr 2024 · 0 repositories · arXiv:2404.06162
-
Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection 9 Apr 2024 · 1 repository · arXiv:2404.06194Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Generative Pre-Trained Transformer for Symbolic Regression Base In-Context Reinforcement Learning 9 Apr 2024 · 0 repositories · arXiv:2404.06330
-
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs 9 Apr 2024 · 0 repositories · arXiv:2404.07242
-
Comprehensive Study on German Language Models for Clinical and Biomedical Text Understanding 8 Apr 2024 · 0 repositories · arXiv:2404.05694
-
Guiding Large Language Models to Generate Computer-Parsable Content 8 Apr 2024 · 0 repositories · arXiv:2404.05499
-
Evaluating Interventional Reasoning Capabilities of Large Language Models 8 Apr 2024 · 0 repositories · arXiv:2404.05545
-
LTNER: Large Language Model Tagging for Named Entity Recognition with Contextualized Entity Marking 8 Apr 2024 · 0 repositories · arXiv:2404.05624
-
MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering 8 Apr 2024 · 0 repositories · arXiv:2404.05590
-
PetKaz at SemEval-2024 Task 3: Advancing Emotion Classification with an LLM for Emotion-Cause Pair Extraction in Conversations 8 Apr 2024 · 1 repository · arXiv:2404.05502
-
Physics of Language Models: Part 3.3, Knowledge Capacity Scaling Laws 8 Apr 2024 · 0 repositories · arXiv:2404.05405
-
Relation Extraction Using Large Language Models: A Case Study on Acupuncture Point Locations 8 Apr 2024 · 0 repositories · arXiv:2404.05415
-
Semantic Stealth: Adversarial Text Attacks on NLP Using Several Methods 8 Apr 2024 · 0 repositories · arXiv:2404.05159
-
A Multi-Level Framework for Accelerating Training Transformer Models 7 Apr 2024 · 1 repository · arXiv:2404.07999Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Morphology-Based Investigation of Positional Encodings 6 Apr 2024 · 0 repositories · arXiv:2404.04530
-
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials 6 Apr 2024 · 1 repository · arXiv:2404.04510
-
RecGPT: Generative Personalized Prompts for Sequential Recommendation via ChatGPT Training Paradigm 6 Apr 2024 · 0 repositories · arXiv:2404.08675
-
Deciphering Political Entity Sentiment in News with Large Language Models: Zero-Shot and Few-Shot Strategies 5 Apr 2024 · 1 repository · arXiv:2404.04361
-
Implicit Bias of AdamW: ℓ_∞ Norm Constrained Optimization 5 Apr 2024 · 0 repositories · arXiv:2404.04454
-
Investigating the Robustness of Modelling Decisions for Few-Shot Cross-Topic Stance Detection: A Preregistered Study 5 Apr 2024 · 1 repository · arXiv:2404.03987
-
Scope Ambiguities in Large Language Models 5 Apr 2024 · 1 repository · arXiv:2404.04332
-
BanglaAutoKG: Automatic Bangla Knowledge Graph Construction with Semantic Neural Graph Filtering 4 Apr 2024 · 1 repository · arXiv:2404.03528Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
CBR-RAG: Case-Based Reasoning for Retrieval Augmented Generation in LLMs for Legal Question Answering 4 Apr 2024 · 1 repository · arXiv:2404.04302
-
CONFLARE: CONFormal LArge language model REtrieval 4 Apr 2024 · 1 repository · arXiv:2404.04287
-
Do Large Language Models Rank Fairly? An Empirical Study on the Fairness of LLMs as Rankers 4 Apr 2024 · 0 repositories · arXiv:2404.03192
-
NLP at UC Santa Cruz at SemEval-2024 Task 5: Legal Answer Validation using Few-Shot Multi-Choice QA 4 Apr 2024 · 1 repository · arXiv:2404.03150
-
Outlier-Efficient Hopfield Layers for Large Transformer-Based Models 4 Apr 2024 · 1 repository · arXiv:2404.03828Syntology official (archive's flag): 10 ran · 10 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
The Death of Feature Engineering? BERT with Linguistic Features on SQuAD 2.0 4 Apr 2024 · 0 repositories · arXiv:2404.03184
-
AI-Tutoring in Software Engineering Education 3 Apr 2024 · 0 repositories · arXiv:2404.02548
-
An Incomplete Loop: Deductive, Inductive, and Abductive Learning in Large Language Models 3 Apr 2024 · 0 repositories · arXiv:2404.03028
-
BCAmirs at SemEval-2024 Task 4: Beyond Words: A Multimodal and Multilingual Exploration of Persuasion in Memes 3 Apr 2024 · 1 repository · arXiv:2404.03022
-
Benchmarking Large Language Models for Persian: A Preliminary Study Focusing on ChatGPT 3 Apr 2024 · 1 repository · arXiv:2404.02403
-
GPT-DETOX: An In-Context Learning-Based Paraphraser for Text Detoxification 3 Apr 2024 · 0 repositories · arXiv:2404.03052
-
uTeBC-NLP at SemEval-2024 Task 9: Can LLMs be Lateral Thinkers? 3 Apr 2024 · 1 repository · arXiv:2404.02474Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Advancing LLM Reasoning Generalists with Preference Trees 2 Apr 2024 · 1 repository · arXiv:2404.02078Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 1 violated, 12 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 20 harvested samples) · 2 pointer-only (licence)
-
Stereotype Detection in LLMs: A Multiclass, Explainable, and Benchmark-Driven Approach 2 Apr 2024 · 0 repositories · arXiv:2404.01768
-
CLAPNQ: Cohesive Long-form Answers from Passages in Natural Questions for RAG systems 2 Apr 2024 · 1 repository · arXiv:2404.02103
-
CMAT: A Multi-Agent Collaboration Tuning Framework for Enhancing Small Language Models 2 Apr 2024 · 1 repository · arXiv:2404.01663
-
Collapse of Self-trained Language Models 2 Apr 2024 · 1 repository · arXiv:2404.02305Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Comparative Study of Domain Driven Terms Extraction Using Large Language Models 2 Apr 2024 · 0 repositories · arXiv:2404.02330
-
Deconstructing In-Context Learning: Understanding Prompts via Corruption 2 Apr 2024 · 1 repository · arXiv:2404.02054
-
GINopic: Topic Modeling with Graph Isomorphism Network 2 Apr 2024 · 1 repository · arXiv:2404.02115Syntology official: no sample here; runs from other or unrecorded repositories · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks 2 Apr 2024 · 1 repository · arXiv:2404.02151Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
METAL: Towards Multilingual Meta-Evaluation 2 Apr 2024 · 0 repositories · arXiv:2404.01667
-
Symbolic Prompt Program Search: A Structure-Aware Approach to Efficient Compile-Time Prompt Optimization 2 Apr 2024 · 1 repository · arXiv:2404.02319
-
Scene Adaptive Sparse Transformer for Event-based Object Detection 2 Apr 2024 · 1 repository · arXiv:2404.01882Syntology official (archive's flag): 19 ran · 19 ran (of which 0 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 23 harvested samples)
-
SGSH: Stimulate Large Language Models with Skeleton Heuristics for Knowledge Base Question Generation 2 Apr 2024 · 1 repository · arXiv:2404.01923
-
Toward Informal Language Processing: Knowledge of Slang in Large Language Models 2 Apr 2024 · 1 repository · arXiv:2404.02323
-
Advancing AI with Integrity: Ethical Challenges and Solutions in Neural Machine Translation 1 Apr 2024 · 0 repositories · arXiv:2404.01070
-
ARAGOG: Advanced RAG Output Grading 1 Apr 2024 · 1 repository · arXiv:2404.01037
-
Artificial Intelligence and the Spatial Documentation of Languages 1 Apr 2024 · 0 repositories · arXiv:2404.01263
-
BERT-Enhanced Retrieval Tool for Homework Plagiarism Detection System 1 Apr 2024 · 0 repositories · arXiv:2404.01582
-
FABLES: Evaluating faithfulness and content selection in book-length summarization 1 Apr 2024 · 3 repositories · arXiv:2404.01261Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Securing Social Spaces: Harnessing Deep Learning to Eradicate Cyberbullying 1 Apr 2024 · 0 repositories · arXiv:2404.03686
-
Unveiling Divergent Inductive Biases of LLMs on Temporal Data 1 Apr 2024 · 1 repository · arXiv:2404.01453
-
CHOPS: CHat with custOmer Profile Systems for Customer Service with LLMs 31 Mar 2024 · 1 repository · arXiv:2404.01343
-
CoUDA: Coherence Evaluation via Unified Data Augmentation 31 Mar 2024 · 1 repository · arXiv:2404.00681
-
Observations on Building RAG Systems for Technical Documents 31 Mar 2024 · 0 repositories · arXiv:2404.00657
-
RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation 31 Mar 2024 · 1 repository · arXiv:2404.00610
-
Training-Free Semantic Segmentation via LLM-Supervision 31 Mar 2024 · 0 repositories · arXiv:2404.00701
-
A Comprehensive Study on NLP Data Augmentation for Hate Speech Detection: Legacy Methods, BERT, and LLMs 30 Mar 2024 · 0 repositories · arXiv:2404.00303
-
Leveraging Pre-trained and Transformer-derived Embeddings from EHRs to Characterize Heterogeneity Across Alzheimer's Disease and Related Dementias 30 Mar 2024 · 0 repositories · arXiv:2404.00464
-
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks 30 Mar 2024 · 0 repositories · arXiv:2404.00376
-
ChatGPT v.s. Media Bias: A Comparative Study of GPT-3.5 and Fine-tuned Language Models 29 Mar 2024 · 0 repositories · arXiv:2403.20158
-
DataAgent: Evaluating Large Language Models' Ability to Answer Zero-Shot, Natural Language Queries 29 Mar 2024 · 0 repositories · arXiv:2404.00188
-
Explainable Deep Learning: A Visual Analytics Approach with Transition Matrices 29 Mar 2024 · 1 repository
-
LayerNorm: A key component in parameter-efficient fine-tuning 29 Mar 2024 · 0 repositories · arXiv:2403.20284
-
ReALM: Reference Resolution As Language Modeling 29 Mar 2024 · 0 repositories · arXiv:2403.20329
-
Shallow Cross-Encoders for Low-Latency Retrieval 29 Mar 2024 · 1 repository · arXiv:2403.20222
-
A Review of Multi-Modal Large Language and Vision Models 28 Mar 2024 · 0 repositories · arXiv:2404.01322
-
AlloyBERT: Alloy Property Prediction with Large Language Models 28 Mar 2024 · 0 repositories · arXiv:2403.19783
-
Are Large Language Models Good at Utility Judgments? 28 Mar 2024 · 1 repository · arXiv:2403.19216Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
FACTOID: FACtual enTailment fOr hallucInation Detection 28 Mar 2024 · 0 repositories · arXiv:2403.19113
-
Generating Multi-Aspect Queries for Conversational Search 28 Mar 2024 · 0 repositories · arXiv:2403.19302
-
Intelligent Classification and Personalized Recommendation of E-commerce Products Based on Machine Learning 28 Mar 2024 · 0 repositories · arXiv:2403.19345
-
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models 28 Mar 2024 · 1 repository · arXiv:2403.19521Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Just-DNA-Seq, open-source personal genomics platform: longevity science for everyone 28 Mar 2024 · 0 repositories · arXiv:2403.19087
-
Risk prediction of pathological gambling on social media 28 Mar 2024 · 0 repositories · arXiv:2403.19358
-
A Novel Corpus of Annotated Medical Imaging Reports and Information Extraction Results Using BERT-based Language Models 27 Mar 2024 · 1 repository · arXiv:2403.18975
-
A Survey on Large Language Models from Concept to Implementation 27 Mar 2024 · 0 repositories · arXiv:2403.18969
-
AcTED: Automatic Acquisition of Typical Event Duration for Semi-supervised Temporal Commonsense QA 27 Mar 2024 · 0 repositories · arXiv:2403.18504
-
Boosting Conversational Question Answering with Fine-Grained Retrieval-Augmentation and Self-Check 27 Mar 2024 · 0 repositories · arXiv:2403.18243
-
CPR: Retrieval Augmented Generation for Copyright Protection 27 Mar 2024 · 0 repositories · arXiv:2403.18920
-
Evaluating Large Language Models for Health-Related Text Classification Tasks with Public Social Media Data 27 Mar 2024 · 0 repositories · arXiv:2403.19031
-
Fusion approaches for emotion recognition from speech using acoustic and text-based features 27 Mar 2024 · 0 repositories · arXiv:2403.18635
-
LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research Practices 27 Mar 2024 · 0 repositories · arXiv:2403.18173
-
Long-form factuality in large language models 27 Mar 2024 · 3 repositories · arXiv:2403.18802Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)