Methods › General › Regularization › Weight Decay › Papers, page 7
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 7 of 108: papers 601 to 700 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Talking to GDELT Through Knowledge Graphs 10 Mar 2025 · 0 repositories · arXiv:2503.07584
-
Effectiveness of Zero-shot-CoT in Japanese Prompts 9 Mar 2025 · 0 repositories · arXiv:2503.06765
-
Fully-Decentralized MADDPG with Networked Agents 9 Mar 2025 · 0 repositories · arXiv:2503.06747
-
Human Cognition Inspired RAG with Knowledge Graph for Complex Problem Solving 9 Mar 2025 · 0 repositories · arXiv:2503.06567
-
Multimodal Emotion Recognition and Sentiment Analysis in Multi-Party Conversation Contexts 9 Mar 2025 · 0 repositories · arXiv:2503.06805
-
Seeing Delta Parameters as JPEG Images: Data-Free Delta Compression with Discrete Cosine Transform 9 Mar 2025 · 0 repositories · arXiv:2503.06676
-
UniGenX: Unified Generation of Sequence and Structure with Autoregressive Diffusion 9 Mar 2025 · 0 repositories · arXiv:2503.06687
-
Breaking Free from MMI: A New Frontier in Rationalization by Probing Input Utilization 8 Mar 2025 · 1 repository · arXiv:2503.06202Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Constructions are Revealed in Word Distributions 8 Mar 2025 · 1 repository · arXiv:2503.06048
-
LimTopic: LLM-based Topic Modeling and Text Summarization for Analyzing Scientific Articles limitations 8 Mar 2025 · 1 repository · arXiv:2503.10658
-
Poisoned-MRAG: Knowledge Poisoning Attacks to Multimodal Retrieval Augmented Generation 8 Mar 2025 · 0 repositories · arXiv:2503.06254
-
Evaluating Large Language Models in Code Generation: INFINITE Methodology for Defining the Inference Index 7 Mar 2025 · 0 repositories · arXiv:2503.05852
-
Simulating and Analysing Human Survey Responses with Large Language Models: A Case Study in Energy Stated Preference 7 Mar 2025 · 0 repositories · arXiv:2503.10652
-
Evaluating open-source Large Language Models for automated fact-checking 7 Mar 2025 · 0 repositories · arXiv:2503.05565
-
FMT:A Multimodal Pneumonia Detection Model Based on Stacking MOE Framework 7 Mar 2025 · 0 repositories · arXiv:2503.05626
-
Leveraging Approximate Caching for Faster Retrieval-Augmented Generation 7 Mar 2025 · 0 repositories · arXiv:2503.05530
-
Leveraging Semantic Type Dependencies for Clinical Named Entity Recognition 7 Mar 2025 · 0 repositories · arXiv:2503.05373
-
Quantifying the Robustness of Retrieval-Augmented Language Models Against Spurious Features in Grounding Data 7 Mar 2025 · 0 repositories · arXiv:2503.05587
-
R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning 7 Mar 2025 · 5 repositories · arXiv:2503.05592Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Zero-shot Medical Event Prediction Using a Generative Pre-trained Transformer on Electronic Health Records 7 Mar 2025 · 0 repositories · arXiv:2503.05893
-
Beyond RAG: Task-Aware KV Cache Compression for Comprehensive Knowledge Reasoning 6 Mar 2025 · 0 repositories · arXiv:2503.04973
-
Collapse of Dense Retrievers: Short, Early, and Literal Biases Outranking Factual Evidence 6 Mar 2025 · 0 repositories · arXiv:2503.05037
-
HILGEN: Hierarchically-Informed Data Generation for Biomedical NER Using Knowledgebases and Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04930
-
In-depth Analysis of Graph-based RAG in a Unified Framework 6 Mar 2025 · 0 repositories · arXiv:2503.04338
-
Incentivizing Multi-Tenant Split Federated Learning for Foundation Models at the Network Edge 6 Mar 2025 · 0 repositories · arXiv:2503.04971
-
Can Frontier LLMs Replace Annotators in Biomedical Text Mining? Analyzing Challenges and Exploring Solutions 5 Mar 2025 · 1 repository · arXiv:2503.03261
-
Intermediate-Task Transfer Learning: Leveraging Sarcasm Detection for Stance Detection 5 Mar 2025 · 0 repositories · arXiv:2503.03172
-
Large language models in finance : what is financial sentiment? 5 Mar 2025 · 0 repositories · arXiv:2503.03612
-
Personalized Federated Fine-tuning for Heterogeneous Data: An Automatic Rank Learning Approach via Two-Level LoRA 5 Mar 2025 · 0 repositories · arXiv:2503.03920
-
Sarcasm Detection as a Catalyst: Improving Stance Detection with Cross-Target Capabilities 5 Mar 2025 · 0 repositories · arXiv:2503.03787
-
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation 5 Mar 2025 · 1 repository · arXiv:2503.03308
-
Effectively Steer LLM To Follow Preference via Building Confident Directions 4 Mar 2025 · 0 repositories · arXiv:2503.02989
-
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02917
-
LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning 4 Mar 2025 · 0 repositories · arXiv:2503.04812
-
Optimizing open-domain question answering with graph-based retrieval augmented generation 4 Mar 2025 · 0 repositories · arXiv:2503.02922
-
PennyLang: Pioneering LLM-Based Quantum Code Generation with a Novel PennyLane-Centric Dataset 4 Mar 2025 · 0 repositories · arXiv:2503.02497
-
Use Me Wisely: AI-Driven Assessment for LLM Prompting Skills Development 4 Mar 2025 · 0 repositories · arXiv:2503.02532
-
Weak-to-Strong Generalization Even in Random Feature Networks, Provably 4 Mar 2025 · 0 repositories · arXiv:2503.02877
-
Wikipedia in the Era of LLMs: Evolution and Risks 4 Mar 2025 · 1 repository · arXiv:2503.02879
-
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders 4 Mar 2025 · 0 repositories · arXiv:2503.02993
-
Enhancing Social Media Rumor Detection: A Semantic and Graph Neural Network Approach for the 2024 Global Election 3 Mar 2025 · 0 repositories · arXiv:2503.01394
-
Machine Learners Should Acknowledge the Legal Implications of Large Language Models as Personal Data 3 Mar 2025 · 0 repositories · arXiv:2503.01630
-
SAGE: A Framework of Precise Retrieval for RAG 3 Mar 2025 · 0 repositories · arXiv:2503.01713
-
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01814
-
Attention Condensation via Sparsity Induced Regularized Training 3 Mar 2025 · 0 repositories · arXiv:2503.01564
-
Boolean-aware Attention for Dense Retrieval 3 Mar 2025 · 0 repositories · arXiv:2503.01753
-
Cancer Type, Stage and Prognosis Assessment from Pathology Reports using LLMs 3 Mar 2025 · 1 repository · arXiv:2503.01194
-
HoH: A Dynamic Benchmark for Evaluating the Impact of Outdated Information on Retrieval-Augmented Generation 3 Mar 2025 · 0 repositories · arXiv:2503.04800Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Retrieval-Augmented Perception: High-Resolution Image Perception Meets Visual RAG 3 Mar 2025 · 1 repository · arXiv:2503.01222
-
SePer: Measure Retrieval Utility Through The Lens Of Semantic Perplexity Reduction 3 Mar 2025 · 1 repository · arXiv:2503.01478Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
SRAG: Structured Retrieval-Augmented Generation for Multi-Entity Question Answering over Wikipedia Graph 3 Mar 2025 · 0 repositories · arXiv:2503.01346
-
M3HF: Multi-agent Reinforcement Learning from Multi-phase Human Feedback of Mixed Quality 3 Mar 2025 · 0 repositories · arXiv:2503.02077
-
Using (Not so) Large Language Models for Generating Simulation Models in a Formal DSL -- A Study on Reaction Networks 3 Mar 2025 · 0 repositories · arXiv:2503.01675
-
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking 2 Mar 2025 · 1 repository · arXiv:2503.00955
-
ER-RAG: Enhance RAG with ER-Based Unified Modeling of Heterogeneous Data Sources 2 Mar 2025 · 0 repositories · arXiv:2504.06271
-
Optimizing Multi-Hop Document Retrieval Through Intermediate Representations 2 Mar 2025 · 0 repositories · arXiv:2503.04796
-
Pseudo-Knowledge Graph: Meta-Path Guided Retrieval and In-Graph Text for RAG-Equipped LLM 1 Mar 2025 · 0 repositories · arXiv:2503.00309
-
BadJudge: Backdoor Vulnerabilities of LLM-as-a-Judge 1 Mar 2025 · 0 repositories · arXiv:2503.00596
-
Hierarchical Multi-Stage BERT Fusion Framework with Dual Attention for Enhanced Cyberbullying Detection in Social Media 1 Mar 2025 · 0 repositories · arXiv:2503.00342
-
Psychological Counseling Ability of Large Language Models 1 Mar 2025 · 0 repositories · arXiv:2503.07627
-
U-NIAH: Unified RAG and LLM Evaluation for Long Context Needle-In-A-Haystack 1 Mar 2025 · 1 repository · arXiv:2503.00353
-
SafeAuto: Knowledge-Enhanced Safe Autonomous Driving with Multimodal Foundation Models 28 Feb 2025 · 1 repository · arXiv:2503.00211
-
Beyond Words: A Latent Memory Approach to Internal Reasoning in LLMs 28 Feb 2025 · 0 repositories · arXiv:2502.21030
-
CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation 28 Feb 2025 · 1 repository · arXiv:2502.21074Syntology 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Ext2Gen: Alignment through Unified Extraction and Generation for Robust Retrieval-Augmented Generation 28 Feb 2025 · 0 repositories · arXiv:2503.04789
-
LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation 28 Feb 2025 · 1 repository · arXiv:2502.20640
-
NutriGen: Personalized Meal Plan Generator Leveraging Large Language Models to Enhance Dietary and Nutritional Adherence 28 Feb 2025 · 1 repository · arXiv:2502.20601
-
Retrieval Augmented Generation for Topic Modeling in Organizational Research: An Introduction with Empirical Demonstration 28 Feb 2025 · 0 repositories · arXiv:2502.20963
-
RuCCoD: Towards Automated ICD Coding in Russian 28 Feb 2025 · 1 repository · arXiv:2502.21263
-
SuperRAG: Beyond RAG with Layout-Aware Graph Modeling 28 Feb 2025 · 0 repositories · arXiv:2503.04790
-
TeleRAG: Efficient Retrieval-Augmented Generation Inference with Lookahead Retrieval 28 Feb 2025 · 0 repositories · arXiv:2502.20969
-
The RAG Paradox: A Black-Box Attack Exploiting Unintentional Vulnerabilities in Retrieval-Augmented Generation Systems 28 Feb 2025 · 0 repositories · arXiv:2502.20995
-
Advanced Deep Learning Techniques for Analyzing Earnings Call Transcripts: Methodologies and Applications 27 Feb 2025 · 0 repositories · arXiv:2503.01886
-
An exploration of features to improve the generalisability of fake news detection models 27 Feb 2025 · 0 repositories · arXiv:2502.20299
-
Bridging Legal Knowledge and AI: Retrieval-Augmented Generation with Vector Stores, Knowledge Graphs, and Hierarchical Non-negative Matrix Factorization 27 Feb 2025 · 1 repository · arXiv:2502.20364
-
Long-Context Inference with Retrieval-Augmented Speculative Decoding 27 Feb 2025 · 1 repository · arXiv:2502.20330
-
SeisMoLLM: Advancing Seismic Monitoring via Cross-modal Transfer with Pre-trained Large Language Model 27 Feb 2025 · 1 repository · arXiv:2502.19960Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Cognitive networks highlight differences and similarities in the STEM mindsets of human and LLM-simulated trainees, experts and academics 26 Feb 2025 · 0 repositories · arXiv:2502.19529
-
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation 26 Feb 2025 · 0 repositories · arXiv:2502.18726
-
Efficient Federated Search for Retrieval-Augmented Generation 26 Feb 2025 · 0 repositories · arXiv:2502.19280
-
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering 26 Feb 2025 · 0 repositories · arXiv:2502.18993
-
NeoBERT: A Next-Generation BERT 26 Feb 2025 · 1 repository · arXiv:2502.19587
-
The Sharpness Disparity Principle in Transformers for Accelerating Language Model Pre-Training 26 Feb 2025 · 0 repositories · arXiv:2502.19002
-
Assessing Large Language Models in Agentic Multilingual National Bias 25 Feb 2025 · 0 repositories · arXiv:2502.17945
-
Broadening Discovery through Structural Models: Multimodal Combination of Local and Structural Properties for Predicting Chemical Features 25 Feb 2025 · 0 repositories · arXiv:2502.17986
-
Detecting Knowledge Boundary of Vision Large Language Models by Sampling-Based Inference 25 Feb 2025 · 0 repositories · arXiv:2502.18023
-
Enhancing Text Classification with a Novel Multi-Agent Collaboration Framework Leveraging BERT 25 Feb 2025 · 0 repositories · arXiv:2502.18653
-
Faster, Cheaper, Better: Multi-Objective Hyperparameter Optimization for LLM and RAG Systems 25 Feb 2025 · 0 repositories · arXiv:2502.18635
-
Independent Mobility GPT (IDM-GPT): A Self-Supervised Multi-Agent Large Language Model Framework for Customized Traffic Mobility Analysis Using Machine Learning Models 25 Feb 2025 · 0 repositories · arXiv:2502.18652
-
LevelRAG: Enhancing Retrieval-Augmented Generation with Multi-hop Logic Planning over Rewriting Augmented Searchers 25 Feb 2025 · 1 repository · arXiv:2502.18139
-
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks 25 Feb 2025 · 1 repository · arXiv:2502.17832
-
MuCoS: Efficient Drug-Target Prediction through Multi-Context-Aware Sampling 25 Feb 2025 · 0 repositories · arXiv:2502.17784
-
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation 25 Feb 2025 · 0 repositories · arXiv:2502.17839
-
Scaling LLM Pre-training with Vocabulary Curriculum 25 Feb 2025 · 0 repositories · arXiv:2502.17910
-
ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents 25 Feb 2025 · 1 repository · arXiv:2502.18017Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
"Actionable Help" in Crises: A Novel Dataset and Resource-Efficient Models for Identifying Request and Offer Social Media Posts 24 Feb 2025 · 0 repositories · arXiv:2502.16839
-
Adversarial Training for Defense Against Label Poisoning Attacks 24 Feb 2025 · 1 repository · arXiv:2502.17121Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Applying LLMs to Active Learning: Towards Cost-Efficient Cross-Task Text Classification without Manually Labeled Data 24 Feb 2025 · 0 repositories · arXiv:2502.16892
-
Are Large Language Models Good Data Preprocessors? 24 Feb 2025 · 0 repositories · arXiv:2502.16790
-
Benchmarking Retrieval-Augmented Generation in Multi-Modal Contexts 24 Feb 2025 · 2 repositories · arXiv:2502.17297