Methods › General › Regularization › Weight Decay › Papers, page 16
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 16 of 108: papers 1,501 to 1,600 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BanditCAT and AutoIRT: Machine Learning Approaches to Computerized Adaptive Testing and Item Calibration 28 Oct 2024 · 0 repositories · arXiv:2410.21033
-
BLAST: Block-Level Adaptive Structured Matrices for Efficient Deep Neural Network Inference 28 Oct 2024 · 1 repository · arXiv:2410.21262
-
Calibrated Decision-Making through LLM-Assisted Retrieval 28 Oct 2024 · 0 repositories · arXiv:2411.08891
-
Causal Interventions on Causal Paths: Mapping GPT-2's Reasoning From Syntax to Semantics 28 Oct 2024 · 0 repositories · arXiv:2410.21353
-
Combining Domain-Specific Models and LLMs for Automated Disease Phenotyping from Survey Data 28 Oct 2024 · 0 repositories · arXiv:2410.20695
-
CRAT: A Multi-Agent Framework for Causality-Enhanced Reflective and Retrieval-Augmented Translation with Large Language Models 28 Oct 2024 · 0 repositories · arXiv:2410.21067
-
Deep Learning for Medical Text Processing: BERT Model Fine-Tuning and Comparative Study 28 Oct 2024 · 0 repositories · arXiv:2410.20792
-
Embedding with Large Language Models for Classification of HIPAA Safeguard Compliance Rules 28 Oct 2024 · 0 repositories · arXiv:2410.20664
-
Geo-FuB: A Method for Constructing an Operator-Function Knowledge Base for Geospatial Code Generation Tasks Using Large Language Models 28 Oct 2024 · 1 repository · arXiv:2410.20975
-
Is GPT-4 Less Politically Biased than GPT-3.5? A Renewed Investigation of ChatGPT's Political Biases 28 Oct 2024 · 0 repositories · arXiv:2410.21008
-
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation 28 Oct 2024 · 1 repository · arXiv:2410.20777
-
LinFormer: A Linear-based Lightweight Transformer Architecture For Time-Aware MIMO Channel Prediction 28 Oct 2024 · 0 repositories · arXiv:2410.21351
-
LLMs are Biased Evaluators But Not Biased for Retrieval Augmented Generation 28 Oct 2024 · 1 repository · arXiv:2410.20833
-
MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression 28 Oct 2024 · 1 repository · arXiv:2410.21548
-
Plan×RAG: Planning-guided Retrieval Augmented Generation 28 Oct 2024 · 0 repositories · arXiv:2410.20753
-
Semantic Search Evaluation 28 Oct 2024 · 0 repositories · arXiv:2410.21549
-
Simple Is Effective: The Roles of Graphs and Large Language Models in Knowledge-Graph-Based Retrieval-Augmented Generation 28 Oct 2024 · 1 repository · arXiv:2410.20724Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
Stealthy Jailbreak Attacks on Large Language Models via Benign Data Mirroring 28 Oct 2024 · 0 repositories · arXiv:2410.21083
-
uOttawa at LegalLens-2024: Transformer-based Classification Experiments 28 Oct 2024 · 1 repository · arXiv:2410.21139
-
Deep Learning Based Dense Retrieval: A Comparative Study 27 Oct 2024 · 0 repositories · arXiv:2410.20315
-
LLM Robustness Against Misinformation in Biomedical Question Answering 27 Oct 2024 · 1 repository · arXiv:2410.21330
-
R^3AG: First Workshop on Refined and Reliable Retrieval Augmented Generation 27 Oct 2024 · 0 repositories · arXiv:2410.20598
-
Sequential Large Language Model-Based Hyper-parameter Optimization 27 Oct 2024 · 1 repository · arXiv:2410.20302
-
Mask-based Membership Inference Attacks for Retrieval-Augmented Generation 26 Oct 2024 · 0 repositories · arXiv:2410.20142
-
Think Carefully and Check Again! Meta-Generation Unlocking LLMs for Low-Resource Cross-Lingual Summarization 26 Oct 2024 · 0 repositories · arXiv:2410.20021
-
A Tutorial on Teaching Data Analytics with Generative AI 25 Oct 2024 · 0 repositories · arXiv:2411.07244
-
ChunkRAG: Novel LLM-Chunk Filtering Method for RAG Systems 25 Oct 2024 · 0 repositories · arXiv:2410.19572
-
FISHNET: Financial Intelligence from Sub-querying, Harmonizing, Neural-Conditioning, Expert Swarms, and Task Planning 25 Oct 2024 · 0 repositories · arXiv:2410.19727
-
GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing 25 Oct 2024 · 1 repository · arXiv:2410.19552Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Integrating Large Language Models with Internet of Things Applications 25 Oct 2024 · 0 repositories · arXiv:2410.19223
-
Bielik 7B v0.1: A Polish Language Model -- Development, Insights, and Evaluation 24 Oct 2024 · 0 repositories · arXiv:2410.18565
-
Difficult for Whom? A Study of Japanese Lexical Complexity 24 Oct 2024 · 1 repository · arXiv:2410.18567
-
Iterative Self-Tuning LLMs for Enhanced Jailbreaking Capabilities 24 Oct 2024 · 1 repository · arXiv:2410.18469
-
Little Giants: Synthesizing High-Quality Embedding Data at Scale 24 Oct 2024 · 1 repository · arXiv:2410.18634Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples)
-
PDL: A Declarative Prompt Programming Language 24 Oct 2024 · 1 repository · arXiv:2410.19135
-
Understanding Ranking LLMs: A Mechanistic Analysis for Information Retrieval 24 Oct 2024 · 0 repositories · arXiv:2410.18527
-
Scaling up Masked Diffusion Models on Text 24 Oct 2024 · 1 repository · arXiv:2410.18514Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Understanding Players as if They Are Talking to the Game in a Customized Language: A Pilot Study 24 Oct 2024 · 0 repositories · arXiv:2410.18605
-
An Adaptive Framework for Generating Systematic Explanatory Answer in Online Q&A Platforms 23 Oct 2024 · 1 repository · arXiv:2410.17694
-
Differentially Private Learning Needs Better Model Initialization and Self-Distillation 23 Oct 2024 · 1 repository · arXiv:2410.17566
-
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction 23 Oct 2024 · 0 repositories · arXiv:2410.18160
-
Leveraging the Domain Adaptation of Retrieval Augmented Generation Models for Question Answering and Reducing Hallucination 23 Oct 2024 · 0 repositories · arXiv:2410.17783
-
Small Singular Values Matter: A Random Matrix Analysis of Transformer Models 23 Oct 2024 · 0 repositories · arXiv:2410.17770
-
LongRAG: A Dual-Perspective Retrieval-Augmented Generation Paradigm for Long-Context Question Answering 23 Oct 2024 · 1 repository · arXiv:2410.18050Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
MCUBERT: Memory-Efficient BERT Inference on Commodity Microcontrollers 23 Oct 2024 · 0 repositories · arXiv:2410.17957
-
OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation 23 Oct 2024 · 1 repository · arXiv:2410.17799
-
SimRAG: Self-Improving Retrieval-Augmented Generation for Adapting Large Language Models to Specialized Domains 23 Oct 2024 · 0 repositories · arXiv:2410.17952
-
Dhoroni: Exploring Bengali Climate Change and Environmental Views with a Multi-Perspective News Dataset and Natural Language Processing 22 Oct 2024 · 1 repository · arXiv:2410.17225
-
Distill-SynthKG: Distilling Knowledge Graph Synthesis Workflow for Improved Coverage and Efficiency 22 Oct 2024 · 0 repositories · arXiv:2410.16597
-
DNAHLM -- DNA sequence and Human Language mixed large language Model 22 Oct 2024 · 1 repository · arXiv:2410.16917
-
Exploring Possibilities of AI-Powered Legal Assistance in Bangladesh through Large Language Modeling 22 Oct 2024 · 1 repository · arXiv:2410.17210
-
Scattered Forest Search: Smarter Code Space Exploration with LLMs 22 Oct 2024 · 0 repositories · arXiv:2411.05010
-
SmartRAG: Jointly Learn RAG-Related Tasks From the Environment Feedback 22 Oct 2024 · 0 repositories · arXiv:2410.18141
-
Tracing the Development of the Virtual Particle Concept Using Semantic Change Detection 22 Oct 2024 · 1 repository · arXiv:2410.16855
-
An Efficient System for Automatic Map Storytelling -- A Case Study on Historical Maps 21 Oct 2024 · 1 repository · arXiv:2410.15780
-
Deep Learning and Data Augmentation for Detecting Self-Admitted Technical Debt 21 Oct 2024 · 1 repository · arXiv:2410.15804
-
Developing Retrieval Augmented Generation (RAG) based LLM Systems from PDFs: An Experience Report 21 Oct 2024 · 1 repository · arXiv:2410.15944
-
Exploring Pretraining via Active Forgetting for Improving Cross Lingual Transfer for Decoder Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16168
-
Guardians of Discourse: Evaluating LLMs on Multilingual Offensive Language Detection 21 Oct 2024 · 0 repositories · arXiv:2410.15623
-
Improving Neuron-level Interpretability with White-box Language Models 21 Oct 2024 · 0 repositories · arXiv:2410.16443
-
Leveraging Retrieval-Augmented Generation for Culturally Inclusive Hakka Chatbots: Design Insights and User Perceptions 21 Oct 2024 · 0 repositories · arXiv:2410.15572
-
LightFusionRec: Lightweight Transformers-Based Cross-Domain Recommendation Model 21 Oct 2024 · 0 repositories · arXiv:2410.15656
-
Natural GaLore: Accelerating GaLore for memory-efficient LLM Training and Fine-tuning 21 Oct 2024 · 1 repository · arXiv:2410.16029
-
On Creating an English-Thai Code-switched Machine Translation in Medical Domain 21 Oct 2024 · 1 repository · arXiv:2410.16221
-
RAG4ITOps: A Supervised Fine-Tunable and Comprehensive RAG Framework for IT Operations and Maintenance 21 Oct 2024 · 0 repositories · arXiv:2410.15805
-
SeisLM: a Foundation Model for Seismic Waveforms 21 Oct 2024 · 1 repository · arXiv:2410.15765
-
Using GPT Models for Qualitative and Quantitative News Analytics in the 2024 US Presidental Election Process 21 Oct 2024 · 0 repositories · arXiv:2410.15884
-
Who's Who: Large Language Models Meet Knowledge Conflicts in Practice 21 Oct 2024 · 1 repository · arXiv:2410.15737
-
BRIEF: Bridging Retrieval and Inference for Multi-hop Reasoning via Compression 20 Oct 2024 · 1 repository · arXiv:2410.15277
-
Contextual Augmented Multi-Model Programming (CAMP): A Hybrid Local-Cloud Copilot Framework 20 Oct 2024 · 1 repository · arXiv:2410.15285
-
ConTReGen: Context-driven Tree-structured Retrieval for Open-domain Long-form Text Generation 20 Oct 2024 · 0 repositories · arXiv:2410.15511
-
Do RAG Systems Cover What Matters? Evaluating and Optimizing Responses with Sub-Question Coverage 20 Oct 2024 · 1 repository · arXiv:2410.15531
-
Does ChatGPT Have a Poetic Style? 20 Oct 2024 · 1 repository · arXiv:2410.15299
-
Evaluating Consistencies in LLM responses through a Semantic Clustering of Question Answering 20 Oct 2024 · 0 repositories · arXiv:2410.15440
-
MMDS: A Multimodal Medical Diagnosis System Integrating Image Analysis and Knowledge-based Departmental Consultation 20 Oct 2024 · 0 repositories · arXiv:2410.15403
-
SDP4Bit: Toward 4-bit Communication Quantization in Sharded Data Parallelism for LLM Training 20 Oct 2024 · 0 repositories · arXiv:2410.15526
-
Unveiling and Consulting Core Experts in Retrieval-Augmented MoE-based LLMs 20 Oct 2024 · 0 repositories · arXiv:2410.15438
-
When Machine Unlearning Meets Retrieval-Augmented Generation (RAG): Keep Secret or Forget Knowledge? 20 Oct 2024 · 0 repositories · arXiv:2410.15267
-
Bias Amplification: Language Models as Increasingly Biased Media 19 Oct 2024 · 0 repositories · arXiv:2410.15234
-
Evaluation Of P300 Speller Performance Using Large Language Models Along With Cross-Subject Training 19 Oct 2024 · 1 repository · arXiv:2410.15161
-
MCCoder: Streamlining Motion Control with LLM-Assisted Code Generation and Rigorous Verification 19 Oct 2024 · 1 repository · arXiv:2410.15154
-
Medical-GAT: Cancer Document Classification Leveraging Graph-Based Residual Network for Scenarios with Limited Data 19 Oct 2024 · 0 repositories · arXiv:2410.15198
-
Automated Genre-Aware Article Scoring and Feedback Using Large Language Models 18 Oct 2024 · 0 repositories · arXiv:2410.14165
-
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models 18 Oct 2024 · 0 repositories · arXiv:2410.14479
-
Effects of Soft-Domain Transfer and Named Entity Information on Deception Detection 18 Oct 2024 · 0 repositories · arXiv:2410.14814
-
Graph Contrastive Learning via Cluster-refined Negative Sampling for Semi-supervised Text Classification 18 Oct 2024 · 0 repositories · arXiv:2410.18130
-
How Does Data Diversity Shape the Weight Landscape of Neural Networks? 18 Oct 2024 · 0 repositories · arXiv:2410.14602
-
Implicit Regularization of Sharpness-Aware Minimization for Scale-Invariant Problems 18 Oct 2024 · 0 repositories · arXiv:2410.14802
-
Optimizing Retrieval-Augmented Generation with Elasticsearch for Enhanced Question-Answering Systems 18 Oct 2024 · 0 repositories · arXiv:2410.14167
-
Paths-over-Graph: Knowledge Graph Empowered Large Language Model Reasoning 18 Oct 2024 · 1 repository · arXiv:2410.14211Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
ELOQ: Resources for Enhancing LLM Detection of Out-of-Scope Questions 18 Oct 2024 · 1 repository · arXiv:2410.14567
-
Real-time Fake News from Adversarial Feedback 18 Oct 2024 · 1 repository · arXiv:2410.14651
-
Sentiment Analysis Based on RoBERTa for Amazon Review: An Empirical Study on Decision Making 18 Oct 2024 · 0 repositories · arXiv:2411.00796
-
ST-MoE-BERT: A Spatial-Temporal Mixture-of-Experts Framework for Long-Term Cross-City Mobility Prediction 18 Oct 2024 · 1 repository · arXiv:2410.14099Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples)
-
How Does Knowledge Selection Help Retrieval Augmented Generation? 17 Oct 2024 · 0 repositories · arXiv:2410.13258
-
Detecting AI-Generated Texts in Cross-Domains 17 Oct 2024 · 1 repository · arXiv:2410.13966
-
Enhancing Text Generation in Joint NLG/NLU Learning Through Curriculum Learning, Semi-Supervised Training, and Advanced Optimization Techniques 17 Oct 2024 · 0 repositories · arXiv:2410.13498
-
Evaluating Self-Generated Documents for Enhancing Retrieval-Augmented Generation with Large Language Models 17 Oct 2024 · 0 repositories · arXiv:2410.13192
-
FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs 17 Oct 2024 · 2 repositories · arXiv:2410.13210
-
Help Me Identify: Is an LLM+VQA System All We Need to Identify Visual Concepts? 17 Oct 2024 · 1 repository · arXiv:2410.13651