Methods › General › Regularization › Weight Decay › Papers, page 13
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 13 of 108: papers 1,201 to 1,300 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Generating Knowledge Graphs from Large Language Models: A Comparative Study of GPT-4, LLaMA 2, and BERT 10 Dec 2024 · 0 repositories · arXiv:2412.07412
-
GPT-2 Through the Lens of Vector Symbolic Architectures 10 Dec 2024 · 0 repositories · arXiv:2412.07947
-
IntellectSeeker: A Personalized Literature Management System with the Probabilistic Model and Large Language Model 10 Dec 2024 · 1 repository · arXiv:2412.07213
-
RAG-based Question Answering over Heterogeneous Data and Text 10 Dec 2024 · 0 repositories · arXiv:2412.07420
-
Superficial Consciousness Hypothesis for Autoregressive Transformers 10 Dec 2024 · 1 repository · arXiv:2412.07278
-
Towards Predictive Communication with Brain-Computer Interfaces integrating Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07355
-
BatchTopK Sparse Autoencoders 9 Dec 2024 · 2 repositories · arXiv:2412.06410Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit 9 Dec 2024 · 0 repositories · arXiv:2412.06370
-
LLM as HPC Expert: Extending RAG Architecture for HPC Data 9 Dec 2024 · 0 repositories · arXiv:2501.14733
-
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06249
-
SiReRAG: Indexing Similar and Related Information for Multihop Reasoning 9 Dec 2024 · 0 repositories · arXiv:2412.06206Syntology 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 1 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.17821
-
Unseen Attack Detection in Software-Defined Networking Using a BERT-Based Large Language Model 9 Dec 2024 · 0 repositories · arXiv:2412.06239
-
A Collaborative Multi-Agent Approach to Retrieval-Augmented Generation Across Diverse Data 8 Dec 2024 · 0 repositories · arXiv:2412.05838
-
M³-20M: A Large-Scale Multi-Modal Molecule Dataset for AI-driven Drug Design and Discovery 8 Dec 2024 · 1 repository · arXiv:2412.06847
-
Mixture-of-PageRanks: Replacing Long-Context with Real-Time, Sparse GraphRAG 8 Dec 2024 · 0 repositories · arXiv:2412.06078
-
BERTCaps: BERT Capsule for Persian Multi-Domain Sentiment Analysis 7 Dec 2024 · 0 repositories · arXiv:2412.05591
-
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds 7 Dec 2024 · 1 repository · arXiv:2412.05631Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
KG-Retriever: Efficient Knowledge Indexing for Retrieval-Augmented Large Language Models 7 Dec 2024 · 1 repository · arXiv:2412.05547Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
PrivAgent: Agentic-based Red-teaming for LLM Privacy Leakage 7 Dec 2024 · 1 repository · arXiv:2412.05734
-
Shifting NER into High Gear: The Auto-AdvER Approach 7 Dec 2024 · 0 repositories · arXiv:2412.05655
-
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering 7 Dec 2024 · 0 repositories · arXiv:2412.06832
-
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo 6 Dec 2024 · 0 repositories · arXiv:2412.05223
-
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG 6 Dec 2024 · 0 repositories · arXiv:2412.05447
-
Are Frontier Large Language Models Suitable for Q&A in Science Centres? 6 Dec 2024 · 0 repositories · arXiv:2412.05200
-
Enhancing Cross-Language Code Translation via Task-Specific Embedding Alignment in Retrieval-Augmented Generation 6 Dec 2024 · 0 repositories · arXiv:2412.05159
-
NLP-ADBench: NLP Anomaly Detection Benchmark 6 Dec 2024 · 1 repository · arXiv:2412.04784
-
No Free Lunch From Random Feature Ensembles 6 Dec 2024 · 0 repositories · arXiv:2412.05418
-
Privacy-Preserving Retrieval-Augmented Generation with Differential Privacy 6 Dec 2024 · 0 repositories · arXiv:2412.04697
-
QueEn: A Large Language Model for Quechua-English Translation 6 Dec 2024 · 0 repositories · arXiv:2412.05184
-
Addressing Hallucinations with RAG and NMISS in Italian Healthcare LLM Chatbots 5 Dec 2024 · 0 repositories · arXiv:2412.04235
-
Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding 5 Dec 2024 · 0 repositories · arXiv:2412.03980
-
Exploring AI Text Generation, Retrieval-Augmented Generation, and Detection Technologies: a Comprehensive Overview 5 Dec 2024 · 0 repositories · arXiv:2412.03933
-
HEAL: Hierarchical Embedding Alignment Loss for Improved Retrieval and Representation Learning 5 Dec 2024 · 1 repository · arXiv:2412.04661
-
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments? 5 Dec 2024 · 0 repositories · arXiv:2412.03856
-
SoRA: Singular Value Decomposed Low-Rank Adaptation for Domain Generalizable Representation Learning 5 Dec 2024 · 1 repository · arXiv:2412.04077
-
Uniform Discretized Integrated Gradients: An effective attribution based method for explaining large language models 5 Dec 2024 · 0 repositories · arXiv:2412.03886
-
Advancing Conversational Psychotherapy: Integrating Privacy, Dual-Memory, and Domain Expertise with Large Language Models 4 Dec 2024 · 0 repositories · arXiv:2412.02987
-
Controlling the Mutation in Large Language Models for the Efficient Evolution of Algorithms 4 Dec 2024 · 0 repositories · arXiv:2412.03250
-
FANAL -- Financial Activity News Alerting Language Modeling Framework 4 Dec 2024 · 0 repositories · arXiv:2412.03527
-
Multimodal Sentiment Analysis Based on BERT and ResNet 4 Dec 2024 · 0 repositories · arXiv:2412.03625
-
Achieving Semantic Consistency: Contextualized Word Representations for Political Text Analysis 3 Dec 2024 · 0 repositories · arXiv:2412.04505
-
CAISSON: Concept-Augmented Inference Suite of Self-Organizing Neural Networks 3 Dec 2024 · 0 repositories · arXiv:2412.02835
-
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity 3 Dec 2024 · 0 repositories · arXiv:2412.02252
-
CPTQuant -- A Novel Mixed Precision Post-Training Quantization Techniques for Large Language Models 3 Dec 2024 · 0 repositories · arXiv:2412.03599
-
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators 3 Dec 2024 · 1 repository · arXiv:2412.02467
-
Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model 3 Dec 2024 · 0 repositories · arXiv:2412.02802
-
Gracefully Filtering Backdoor Samples for Generative Large Language Models without Retraining 3 Dec 2024 · 1 repository · arXiv:2412.02454
-
Impact of Data Snooping on Deep Learning Models for Locating Vulnerabilities in Lifted Code 3 Dec 2024 · 0 repositories · arXiv:2412.02048
-
OCR Hinders RAG: Evaluating the Cascading Impact of OCR on Retrieval-Augmented Generation 3 Dec 2024 · 1 repository · arXiv:2412.02592
-
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction 3 Dec 2024 · 0 repositories · arXiv:2412.02698
-
Semantic Tokens in Retrieval Augmented Generation 3 Dec 2024 · 0 repositories · arXiv:2412.02563
-
The Asymptotic Behavior of Attention in Transformers 3 Dec 2024 · 0 repositories · arXiv:2412.02682
-
MBA-RAG: a Bandit Approach for Adaptive Retrieval-Augmented Generation through Question Complexity 2 Dec 2024 · 1 repository · arXiv:2412.01572Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models 2 Dec 2024 · 0 repositories · arXiv:2412.01353
-
The Promise and Peril of Generative AI: Evidence from GPT-4 as Sell-Side Analysts 2 Dec 2024 · 0 repositories · arXiv:2412.01069
-
Tokenizing 3D Molecule Structure with Quantized Spherical Coordinates 2 Dec 2024 · 0 repositories · arXiv:2412.01564
-
A Comprehensive Guide to Explainable AI: From Classical Models to LLMs 1 Dec 2024 · 1 repository · arXiv:2412.00800
-
EventGPT: Event Stream Understanding with Multimodal Large Language Models 1 Dec 2024 · 0 repositories · arXiv:2412.00832
-
Lightweight Contenders: Navigating Semi-Supervised Text Mining through Peer Collaboration and Self Transcendence 1 Dec 2024 · 1 repository · arXiv:2412.00883
-
CDEMapper: Enhancing NIH Common Data Element Normalization using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00491
-
Cognitive Biases in Large Language Models: A Survey and Mitigation Experiments 30 Nov 2024 · 0 repositories · arXiv:2412.00323
-
Does Self-Attention Need Separate Weights in Transformers? 30 Nov 2024 · 0 repositories · arXiv:2412.00359
-
Empowering the Deaf and Hard of Hearing Community: Enhancing Video Captions Using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00342
-
Fairness at Every Intersection: Uncovering and Mitigating Intersectional Biases in Multimodal Clinical Predictions 30 Nov 2024 · 0 repositories · arXiv:2412.00606
-
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters 30 Nov 2024 · 0 repositories · arXiv:2412.00530
-
Advanced System Integration: Analyzing OpenAPI Chunking for Retrieval-Augmented Generation 29 Nov 2024 · 0 repositories · arXiv:2411.19804
-
Generating a Low-code Complete Workflow via Task Decomposition and RAG 29 Nov 2024 · 0 repositories · arXiv:2412.00239
-
Know Your RAG: Dataset Taxonomy and Generation Strategies for Evaluating RAG Systems 29 Nov 2024 · 0 repositories · arXiv:2411.19710
-
Knowledge Management for Automobile Failure Analysis Using Graph RAG 29 Nov 2024 · 0 repositories · arXiv:2411.19539
-
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation 29 Nov 2024 · 0 repositories · arXiv:2411.19528
-
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation 29 Nov 2024 · 0 repositories · arXiv:2411.19921
-
Towards Understanding Retrieval Accuracy and Prompt Quality in RAG Systems 29 Nov 2024 · 0 repositories · arXiv:2411.19463
-
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection 28 Nov 2024 · 0 repositories · arXiv:2411.19220
-
Beautimeter: Harnessing GPT for Assessing Architectural and Urban Beauty based on the 15 Properties of Living Structure 28 Nov 2024 · 0 repositories · arXiv:2411.19094
-
DENIAHL: In-Context Features Influence LLM Needle-In-A-Haystack Abilities 28 Nov 2024 · 1 repository · arXiv:2411.19360
-
Efficient Learning Content Retrieval with Knowledge Injection 28 Nov 2024 · 0 repositories · arXiv:2412.00125
-
Habit Coach: Customising RAG-based chatbots to support behavior change 28 Nov 2024 · 0 repositories · arXiv:2411.19229
-
RevPRAG: Revealing Poisoning Attacks in Retrieval-Augmented Generation through LLM Activation Analysis 28 Nov 2024 · 0 repositories · arXiv:2411.18948
-
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation 28 Nov 2024 · 1 repository · arXiv:2411.19067
-
SmartLLMSentry: A Comprehensive LLM Based Smart Contract Vulnerability Detection Framework 28 Nov 2024 · 0 repositories · arXiv:2411.19234
-
The Impact of Example Selection in Few-Shot Prompting on Automated Essay Scoring Using GPT Models 28 Nov 2024 · 0 repositories · arXiv:2411.18924
-
Automated Literature Review Using NLP Techniques and LLM-Based Retrieval-Augmented Generation 27 Nov 2024 · 0 repositories · arXiv:2411.18583
-
Can bidirectional encoder become the ultimate winner for downstream applications of foundation models? 27 Nov 2024 · 0 repositories · arXiv:2411.18021
-
ChatGPT as speechwriter for the French presidents 27 Nov 2024 · 0 repositories · arXiv:2411.18382
-
DRS: Deep Question Reformulation With Structured Output 27 Nov 2024 · 1 repository · arXiv:2411.17993
-
Evaluating and Improving the Robustness of Security Attack Detectors Generated by LLMs 27 Nov 2024 · 1 repository · arXiv:2411.18216
-
Fine-Tuning Large Language Models for Scientific Text Classification: A Comparative Study 27 Nov 2024 · 0 repositories · arXiv:2412.00098
-
Fine-Tuning Small Embeddings for Elevated Performance 27 Nov 2024 · 0 repositories · arXiv:2411.18099
-
On Importance of Code-Mixed Embeddings for Hate Speech Identification 27 Nov 2024 · 0 repositories · arXiv:2411.18577
-
Streamlining Prediction in Bayesian Deep Learning 27 Nov 2024 · 1 repository · arXiv:2411.18425Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Training and Evaluating Language Models with Template-based Data Generation 27 Nov 2024 · 1 repository · arXiv:2411.18104Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Advancing Content Moderation: Evaluating Large Language Models for Detecting Sensitive Content Across Text, Images, and Videos 26 Nov 2024 · 0 repositories · arXiv:2411.17123
-
BERT or FastText? A Comparative Analysis of Contextual as well as Non-Contextual Embeddings 26 Nov 2024 · 1 repository · arXiv:2411.17661
-
Can artificial intelligence predict clinical trial outcomes? 26 Nov 2024 · 0 repositories · arXiv:2411.17595
-
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning 26 Nov 2024 · 1 repository · arXiv:2411.17426
-
Distributed Sign Momentum with Local Steps for Training Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17866
-
Fairness And Performance In Harmony: Data Debiasing Is All You Need 26 Nov 2024 · 0 repositories · arXiv:2411.17374
-
"Give me the code" -- Log Analysis of First-Year CS Students' Interactions With GPT 26 Nov 2024 · 0 repositories · arXiv:2411.17855