Methods › General › Regularization › Weight Decay › Papers, page 14
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 14 of 108: papers 1,301 to 1,400 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
On Limitations of LLM as Annotator for Low Resource Languages 26 Nov 2024 · 0 repositories · arXiv:2411.17637
-
Pretrained LLM Adapted with LoRA as a Decision Transformer for Offline RL in Quantitative Trading 26 Nov 2024 · 1 repository · arXiv:2411.17900
-
What Differentiates Educational Literature? A Multimodal Fusion Approach of Transformers and Computational Linguistics 26 Nov 2024 · 0 repositories · arXiv:2411.17593
-
Adaptive Circuit Behavior and Generalization in Mechanistic Interpretability 25 Nov 2024 · 0 repositories · arXiv:2411.16105
-
Are Transformers Truly Foundational for Robotics? 25 Nov 2024 · 0 repositories · arXiv:2411.16917
-
AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning 25 Nov 2024 · 1 repository · arXiv:2411.16495Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16991
-
Fine-Tuning LLMs with Noisy Data for Political Argument Generation and Post Guidance 25 Nov 2024 · 0 repositories · arXiv:2411.16813
-
Human-Calibrated Automated Testing and Validation of Generative Language Models 25 Nov 2024 · 0 repositories · arXiv:2411.16391
-
LaB-RAG: Label Boosted Retrieval Augmented Generation for Radiology Report Generation 25 Nov 2024 · 1 repository · arXiv:2411.16523
-
MarketGPT: Developing a Pre-trained transformer (GPT) for Modeling Financial Time Series 25 Nov 2024 · 1 repository · arXiv:2411.16585Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Predictive Power of LLMs in Financial Markets 25 Nov 2024 · 0 repositories · arXiv:2411.16569
-
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training 25 Nov 2024 · 0 repositories · arXiv:2411.16618
-
Broad Critic Deep Actor Reinforcement Learning for Continuous Control 24 Nov 2024 · 0 repositories · arXiv:2411.15806
-
Development of Pre-Trained Transformer-based Models for the Nepali Language 24 Nov 2024 · 0 repositories · arXiv:2411.15734
-
RAMIE: Retrieval-Augmented Multi-task Information Extraction with Large Language Models on Dietary Supplements 24 Nov 2024 · 0 repositories · arXiv:2411.15700
-
A Comparative Analysis of Transformer and LSTM Models for Detecting Suicidal Ideation on Reddit 23 Nov 2024 · 1 repository · arXiv:2411.15404
-
"All that Glitters": Approaches to Evaluations with Unreliable Model and Human Annotations 23 Nov 2024 · 1 repository · arXiv:2411.15634
-
ChatBCI: A P300 Speller BCI Leveraging Large Language Models for Improved Sentence Composition in Realistic Scenarios 23 Nov 2024 · 0 repositories · arXiv:2411.15395
-
Improving Next Tokens via Second-Last Predictions with Generate and Refine 23 Nov 2024 · 0 repositories · arXiv:2411.15661
-
Inducing Human-like Biases in Moral Reasoning Language Models 23 Nov 2024 · 0 repositories · arXiv:2411.15386
-
Traditional Chinese Medicine Case Analysis System for High-Level Semantic Abstraction: Optimized with Prompt and RAG 23 Nov 2024 · 0 repositories · arXiv:2411.15491
-
Astro-HEP-BERT: A bidirectional language model for studying the meanings of concepts in astrophysics and high energy physics 22 Nov 2024 · 0 repositories · arXiv:2411.14877
-
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective 22 Nov 2024 · 0 repositories · arXiv:2411.14654
-
KBAlign: Efficient Self Adaptation on Specific Knowledge Bases 22 Nov 2024 · 1 repository · arXiv:2411.14790
-
An Experimental Study on Data Augmentation Techniques for Named Entity Recognition on Low-Resource Domains 21 Nov 2024 · 0 repositories · arXiv:2411.14551
-
Assessment of LLM Responses to End-user Security Questions 21 Nov 2024 · 0 repositories · arXiv:2411.14571
-
BERT-Based Approach for Automating Course Articulation Matrix Construction with Explainable AI 21 Nov 2024 · 1 repository · arXiv:2411.14254
-
Evaluating the Robustness of Analogical Reasoning in Large Language Models 21 Nov 2024 · 1 repository · arXiv:2411.14215Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
FastRAG: Retrieval Augmented Generation for Semi-structured Data 21 Nov 2024 · 0 repositories · arXiv:2411.13773
-
G-RAG: Knowledge Expansion in Material Science 21 Nov 2024 · 1 repository · arXiv:2411.14592
-
POS-tagging to highlight the skeletal structure of sentences 21 Nov 2024 · 2 repositories · arXiv:2411.14393
-
Towards Knowledge Checking in Retrieval-augmented Generation: A Representation Perspective 21 Nov 2024 · 0 repositories · arXiv:2411.14572
-
AI-Driven Agents with Prompts Designed for High Agreeableness Increase the Likelihood of Being Mistaken for a Human in the Turing Test 20 Nov 2024 · 0 repositories · arXiv:2411.13749
-
Combining Autoregressive and Autoencoder Language Models for Text Classification 20 Nov 2024 · 1 repository · arXiv:2411.13282
-
DMQR-RAG: Diverse Multi-Query Rewriting for RAG 20 Nov 2024 · 0 repositories · arXiv:2411.13154
-
Exploring Large Language Models for Climate Forecasting 20 Nov 2024 · 0 repositories · arXiv:2411.13724
-
Multimodal large language model for wheat breeding: a new exploration of smart breeding 20 Nov 2024 · 0 repositories · arXiv:2411.15203
-
On the Way to LLM Personalization: Learning to Remember User Conversations 20 Nov 2024 · 0 repositories · arXiv:2411.13405
-
Provably Efficient Action-Manipulation Attack Against Continuous Reinforcement Learning 20 Nov 2024 · 0 repositories · arXiv:2411.13116
-
Retrieval-Augmented Generation for Domain-Specific Question Answering: A Case Study on Pittsburgh and CMU 20 Nov 2024 · 0 repositories · arXiv:2411.13691
-
Unlocking Historical Clinical Trial Data with ALIGN: A Compositional Large Language Model System for Medical Coding 20 Nov 2024 · 0 repositories · arXiv:2411.13163
-
A Combined Encoder and Transformer Approach for Coherent and High-Quality Text Generation 19 Nov 2024 · 0 repositories · arXiv:2411.12157
-
DLBacktrace: A Model Agnostic Explainability for any Deep Learning Models 19 Nov 2024 · 1 repository · arXiv:2411.12643
-
Enhancing Multi-Class Disease Classification: Neoplasms, Cardiovascular, Nervous System, and Digestive Disorders Using Advanced LLMs 19 Nov 2024 · 0 repositories · arXiv:2411.12712
-
Leveraging Virtual Reality and AI Tutoring for Language Learning: A Case Study of a Virtual Campus Environment with OpenAI GPT Integration with Unity 3D 19 Nov 2024 · 0 repositories · arXiv:2411.12619
-
Strengthening Fake News Detection: Leveraging SVM and Sophisticated Text Vectorization Techniques. Defying BERT? 19 Nov 2024 · 0 repositories · arXiv:2411.12703
-
Can Open-source LLMs Enhance Data Synthesis for Toxic Detection?: An Experimental Study 18 Nov 2024 · 0 repositories · arXiv:2411.15175
-
Chapter 7 Review of Data-Driven Generative AI Models for Knowledge Extraction from Scientific Literature in Healthcare 18 Nov 2024 · 0 repositories · arXiv:2411.11635
-
CNMBERT: A Model for Converting Hanyu Pinyin Abbreviations to Chinese Characters 18 Nov 2024 · 1 repository · arXiv:2411.11770
-
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback 18 Nov 2024 · 1 repository · arXiv:2412.03578Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Suicide Risk Assessment on Social Media with Semi-Supervised Learning 18 Nov 2024 · 0 repositories · arXiv:2411.12767
-
Understanding Student Sentiment on Mental Health Support in Colleges Using Large Language Models 18 Nov 2024 · 0 repositories · arXiv:2412.04326
-
VersaTune: An Efficient Data Composition Framework for Training Multi-Capability LLMs 18 Nov 2024 · 1 repository · arXiv:2411.11266
-
A Novel Approach to Eliminating Hallucinations in Large Language Model-Assisted Causal Discovery 16 Nov 2024 · 0 repositories · arXiv:2411.12759
-
Debias-CLR: A Contrastive Learning Based Debiasing Method for Algorithmic Fairness in Healthcare Applications 15 Nov 2024 · 0 repositories · arXiv:2411.10544
-
Does Prompt Formatting Have Any Impact on LLM Performance? 15 Nov 2024 · 0 repositories · arXiv:2411.10541
-
Hysteresis Activation Function for Efficient Inference 15 Nov 2024 · 1 repository · arXiv:2411.10573Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Information Extraction from Clinical Notes: Are We Ready to Switch to Large Language Models? 15 Nov 2024 · 1 repository · arXiv:2411.10020
-
MARS: Unleashing the Power of Variance Reduction for Training Large Models 15 Nov 2024 · 2 repositories · arXiv:2411.10438Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation 15 Nov 2024 · 0 repositories · arXiv:2411.10129
-
SoftLMs: Efficient Adaptive Low-Rank Approximation of Language Models using Soft-Thresholding Mechanism 15 Nov 2024 · 0 repositories · arXiv:2411.10543
-
Take Package as Language: Anomaly Detection Using Transformer 15 Nov 2024 · 0 repositories · arXiv:2412.04473
-
Adopting RAG for LLM-Aided Future Vehicle Design 14 Nov 2024 · 0 repositories · arXiv:2411.09590
-
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency 14 Nov 2024 · 0 repositories · arXiv:2411.09587
-
Beyond Static Tools: Evaluating Large Language Models for Cryptographic Misuse Detection 14 Nov 2024 · 0 repositories · arXiv:2411.09772
-
Comprehensive and Practical Evaluation of Retrieval-Augmented Generation Systems for Medical Question Answering 14 Nov 2024 · 0 repositories · arXiv:2411.09213
-
HateGPT: Unleashing GPT-3.5 Turbo to Combat Hate Speech on X 14 Nov 2024 · 0 repositories · arXiv:2411.09214
-
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework 14 Nov 2024 · 2 repositories · arXiv:2411.09607
-
A Large-Scale Study of Relevance Assessments with Large Language Models: An Initial Look 13 Nov 2024 · 1 repository · arXiv:2411.08275
-
Analyst Reports and Stock Performance: Evidence from the Chinese Market 13 Nov 2024 · 0 repositories · arXiv:2411.08726
-
CamemBERT 2.0: A Smarter French Language Model Aged to Perfection 13 Nov 2024 · 0 repositories · arXiv:2411.08868
-
LLMStinger: Jailbreaking LLMs using RL fine-tuned LLMs 13 Nov 2024 · 0 repositories · arXiv:2411.08862
-
LogLLM: Log-based Anomaly Detection Using Large Language Models 13 Nov 2024 · 1 repository · arXiv:2411.08561
-
Responsible AI in Construction Safety: Systematic Evaluation of Large Language Models and Prompt Engineering 13 Nov 2024 · 0 repositories · arXiv:2411.08320
-
Towards Optimizing a Retrieval Augmented Generation using Large Language Model on Academic Data 13 Nov 2024 · 0 repositories · arXiv:2411.08438
-
VALTEST: Automated Validation of Language Model Generated Test Cases 13 Nov 2024 · 0 repositories · arXiv:2411.08254
-
Controlled Evaluation of Syntactic Knowledge in Multilingual Language Models 12 Nov 2024 · 1 repository · arXiv:2411.07474
-
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis 12 Nov 2024 · 1 repository · arXiv:2411.07529
-
Fair Summarization: Bridging Quality and Diversity in Extractive Summaries 12 Nov 2024 · 1 repository · arXiv:2411.07521Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
LLM App Squatting and Cloning 12 Nov 2024 · 0 repositories · arXiv:2411.07518
-
Query Optimization for Parametric Knowledge Refinement in Retrieval-Augmented Large Language Models 12 Nov 2024 · 0 repositories · arXiv:2411.07820
-
Retrieval Augmented Time Series Forecasting 12 Nov 2024 · 1 repository · arXiv:2411.08249Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders 12 Nov 2024 · 0 repositories · arXiv:2411.07870
-
A Unified Multi-Task Learning Architecture for Hate Detection Leveraging User-Based Information 11 Nov 2024 · 0 repositories · arXiv:2411.06855
-
Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models 11 Nov 2024 · 0 repositories · arXiv:2411.06713
-
AssistRAG: Boosting the Potential of Large Language Models with an Intelligent Information Assistant 11 Nov 2024 · 1 repository · arXiv:2411.06805Syntology official (archive's flag): 10 ran · 10 ran (of which 1 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Autonomous Droplet Microfluidic Design Framework with Large Language Models 11 Nov 2024 · 1 repository · arXiv:2411.06691
-
Cancer-Answer: Empowering Cancer Care with Advanced Large Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.06946
-
Evaluating Large Language Models on Financial Report Summarization: An Empirical Study 11 Nov 2024 · 0 repositories · arXiv:2411.06852
-
Explore the Reasoning Capability of LLMs in the Chess Testbed 11 Nov 2024 · 0 repositories · arXiv:2411.06655
-
Invar-RAG: Invariant LLM-aligned Retrieval for Better Generation 11 Nov 2024 · 0 repositories · arXiv:2411.07021
-
LA4SR: illuminating the dark proteome with generative AI 11 Nov 2024 · 0 repositories · arXiv:2411.06798
-
On Active Privacy Auditing in Supervised Fine-tuning for White-Box Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.07070
-
Quantum Policy Gradient in Reproducing Kernel Hilbert Space 11 Nov 2024 · 0 repositories · arXiv:2411.06650
-
SPARTAN: A Sparse Transformer Learning Local Causation 11 Nov 2024 · 0 repositories · arXiv:2411.06890
-
TempCharBERT: Keystroke Dynamics for Continuous Access Control Based on Pre-trained Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.07224
-
The Backpropagation of the Wave Network 11 Nov 2024 · 0 repositories · arXiv:2411.06989
-
Toward Optimal Search and Retrieval for RAG 11 Nov 2024 · 1 repository · arXiv:2411.07396