Methods › General › Regularization › Weight Decay › Papers, page 40
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 40 of 108: papers 3,901 to 4,000 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Decoding Logic Errors: A Comparative Study on Bug Detection by Students and Large Language Models 27 Nov 2023 · 0 repositories · arXiv:2311.16017
-
Leveraging deep active learning to identify low-resource mobility functioning information in public clinical notes 27 Nov 2023 · 0 repositories · arXiv:2311.15946
-
Real Customization or Just Marketing: Are Customized Versions of Chat GPT Useful? 27 Nov 2023 · 0 repositories · arXiv:2312.03728
-
SSIN: Self-Supervised Learning for Rainfall Spatial Interpolation 27 Nov 2023 · 1 repository · arXiv:2311.15530
-
Leveraging AI-derived Data for Carbon Accounting: Information Extraction from Alternative Sources 26 Nov 2023 · 0 repositories · arXiv:2312.03722
-
Machine-Generated Text Detection using Deep Learning 26 Nov 2023 · 1 repository · arXiv:2311.15425
-
UHGEval: Benchmarking the Hallucination of Chinese Large Language Models via Unconstrained Generation 26 Nov 2023 · 1 repository · arXiv:2311.15296Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Uncertainty-aware Language Modeling for Selective Question Answering 26 Nov 2023 · 0 repositories · arXiv:2311.15451
-
SwiftLearn: A Data-Efficient Training Method of Deep Learning Models using Importance Sampling 25 Nov 2023 · 0 repositories · arXiv:2311.15134
-
CMed-GPT: Prompt Tuning for Entity-Aware Chinese Medical Dialogue Generation 24 Nov 2023 · 0 repositories · arXiv:2311.14539
-
Data-to-Text Bilingual Generation 24 Nov 2023 · 2 repositories · arXiv:2311.14808
-
GPT Struct Me: Probing GPT Models on Narrative Entity Extraction 24 Nov 2023 · 1 repository · arXiv:2311.14583
-
Large Language Models as Automated Aligners for benchmarking Vision-Language Models 24 Nov 2023 · 0 repositories · arXiv:2311.14580
-
Machine Translation for Ge'ez Language 24 Nov 2023 · 0 repositories · arXiv:2311.14530
-
A Cross Attention Approach to Diagnostic Explainability using Clinical Practice Guidelines for Depression 23 Nov 2023 · 1 repository · arXiv:2311.13852
-
A Multi-solution Study on GDPR AI-enabled Completeness Checking of DPAs 23 Nov 2023 · 0 repositories · arXiv:2311.13881
-
Annotation Sensitivity: Training Data Collection Methods Affect Model Performance 23 Nov 2023 · 1 repository · arXiv:2311.14212
-
Hardware Resilience Properties of Text-Guided Image Classifiers 23 Nov 2023 · 1 repository · arXiv:2311.14062Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Minimizing Factual Inconsistency and Hallucination in Large Language Models 23 Nov 2023 · 0 repositories · arXiv:2311.13878
-
Towards Auditing Large Language Models: Improving Text-based Stereotype Detection 23 Nov 2023 · 0 repositories · arXiv:2311.14126
-
Comparison of pipeline, sequence-to-sequence, and GPT models for end-to-end relation extraction: experiments with the rare disease use-case 22 Nov 2023 · 1 repository · arXiv:2311.13729
-
Current Topological and Machine Learning Applications for Bias Detection in Text 22 Nov 2023 · 0 repositories · arXiv:2311.13495
-
Detecting out-of-distribution text using topological features of transformer-based language models 22 Nov 2023 · 1 repository · arXiv:2311.13102
-
Drilling Down into the Discourse Structure with LLMs for Long Document Question Answering 22 Nov 2023 · 0 repositories · arXiv:2311.13565
-
Generation of Explanations for Logic Reasoning 22 Nov 2023 · 0 repositories · arXiv:2311.13455
-
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning 22 Nov 2023 · 0 repositories · arXiv:2311.13721
-
PG-Video-LLaVA: Pixel Grounding Large Video-Language Models 22 Nov 2023 · 1 repository · arXiv:2311.13435Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
AlignedCoT: Prompting Large Language Models via Native-Speaking Demonstrations 22 Nov 2023 · 1 repository · arXiv:2311.13538
-
@ve: A Chatbot for Latin 22 Nov 2023 · 0 repositories · arXiv:2311.14741
-
White-Box Transformers via Sparse Rate Reduction: Compression Is All There Is? 22 Nov 2023 · 1 repository · arXiv:2311.13110Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Survey on Large Language Models for Personalized and Explainable Recommendations 21 Nov 2023 · 0 repositories · arXiv:2311.12338
-
AcademicGPT: Empowering Academic Research 21 Nov 2023 · 0 repositories · arXiv:2311.12315
-
ALPHA: AnomaLous Physiological Health Assessment Using Large Language Models 21 Nov 2023 · 1 repository · arXiv:2311.12524
-
Descriptor and Word Soups: Overcoming the Parameter Efficiency Accuracy Tradeoff for Out-of-Distribution Few-shot Learning 21 Nov 2023 · 1 repository · arXiv:2311.13612
-
Extracting Definienda in Mathematical Scholarly Articles with Transformers 21 Nov 2023 · 2 repositories · arXiv:2311.12448
-
GPT4Motion: Scripting Physical Motions in Text-to-Video Generation via Blender-Oriented GPT Planning 21 Nov 2023 · 0 repositories · arXiv:2311.12631
-
InterPrompt: Interpretable Prompting for Interrelated Interpersonal Risk Factors in Reddit Posts 21 Nov 2023 · 0 repositories · arXiv:2311.12404
-
LowResource at BLP-2023 Task 2: Leveraging BanglaBert for Low Resource Sentiment Analysis of Bangla Language 21 Nov 2023 · 1 repository · arXiv:2311.12735
-
Utilizing Language Models for Tour Itinerary Recommendation 21 Nov 2023 · 0 repositories · arXiv:2311.12355
-
Assessing Prompt Injection Risks in 200+ Custom GPTs 20 Nov 2023 · 1 repository · arXiv:2311.11538
-
Evil Geniuses: Delving into the Safety of LLM-based Agents 20 Nov 2023 · 1 repository · arXiv:2311.11855Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Towards Human-Level Text Coding with LLMs: The Case of Fatherhood Roles in Public Policy Documents 20 Nov 2023 · 1 repository · arXiv:2311.11844
-
LogLead -- Fast and Integrated Log Loader, Enhancer, and Anomaly Detector 20 Nov 2023 · 1 repository · arXiv:2311.11809
-
LQ-LoRA: Low-rank Plus Quantized Matrix Decomposition for Efficient Language Model Finetuning 20 Nov 2023 · 1 repository · arXiv:2311.12023Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
MemoryCompanion: A Smart Healthcare Solution to Empower Efficient Alzheimer's Care Via Unleashing Generative AI 20 Nov 2023 · 0 repositories · arXiv:2311.14730
-
Refactoring Programs Using Large Language Models with Few-Shot Examples 20 Nov 2023 · 0 repositories · arXiv:2311.11690
-
Spot the Bot: Distinguishing Human-Written and Bot-Generated Texts Using Clustering and Information Theory Techniques 19 Nov 2023 · 0 repositories · arXiv:2311.11441
-
Tensor-Aware Energy Accounting 19 Nov 2023 · 1 repository · arXiv:2311.11424
-
Weight Norm Control 19 Nov 2023 · 0 repositories · arXiv:2311.11446
-
Behavior Optimized Image Generation 18 Nov 2023 · 0 repositories · arXiv:2311.10995
-
Compositional Fusion of Signals in Data Embedding 18 Nov 2023 · 0 repositories · arXiv:2311.11085
-
Polynomial-Time Solutions for ReLU Network Training: A Complexity Classification via Max-Cut and Zonotopes 18 Nov 2023 · 0 repositories · arXiv:2311.10972
-
Advancements in Generative AI: A Comprehensive Review of GANs, GPT, Autoencoders, Diffusion Model, and Transformers 17 Nov 2023 · 0 repositories · arXiv:2311.10242
-
Bias A-head? Analyzing Bias in Transformer-Based Language Model Attention Heads 17 Nov 2023 · 0 repositories · arXiv:2311.10395
-
Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2 17 Nov 2023 · 3 repositories · arXiv:2311.10702
-
DynaPipe: Optimizing Multi-task Training through Dynamic Pipelines 17 Nov 2023 · 2 repositories · arXiv:2311.10418
-
Extracting periodontitis diagnosis in clinical notes with RoBERTa and regular expression 17 Nov 2023 · 0 repositories · arXiv:2311.10809
-
Hashing it Out: Predicting Unhealthy Conversations on Twitter 17 Nov 2023 · 1 repository · arXiv:2311.10596
-
Use GPT-J Prompt Generation with RoBERTa for NER Models on Diagnosis Extraction of Periodontal Diagnosis from Electronic Dental Records 17 Nov 2023 · 0 repositories · arXiv:2311.10810
-
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems 16 Nov 2023 · 1 repository · arXiv:2311.09476Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Event Causality Is Key to Computational Story Understanding 16 Nov 2023 · 1 repository · arXiv:2311.09648Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Fumbling in Babel: An Investigation into ChatGPT's Language Identification Ability 16 Nov 2023 · 0 repositories · arXiv:2311.09696
-
Generative AI for Hate Speech Detection: Evaluation and Findings 16 Nov 2023 · 0 repositories · arXiv:2311.09993
-
Human Still Wins over LLM: An Empirical Study of Active Learning on Domain-Specific Annotation Tasks 16 Nov 2023 · 0 repositories · arXiv:2311.09825
-
INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair 16 Nov 2023 · 1 repository · arXiv:2311.09868
-
FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains 16 Nov 2023 · 1 repository · arXiv:2311.09797Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
LifeTox: Unveiling Implicit Toxicity in Life Advice 16 Nov 2023 · 1 repository · arXiv:2311.09585
-
On Retrieval Augmentation and the Limitations of Language Model Training 16 Nov 2023 · 0 repositories · arXiv:2311.09615
-
Predictive Minds: LLMs As Atypical Active Inference Agents 16 Nov 2023 · 0 repositories · arXiv:2311.10215
-
Reducing Privacy Risks in Online Self-Disclosures with Language Models 16 Nov 2023 · 0 repositories · arXiv:2311.09538
-
Sequencing Matters: A Generate-Retrieve-Generate Model for Building Conversational Agents 16 Nov 2023 · 0 repositories · arXiv:2311.09513
-
Chemist-X: Large Language Model-empowered Agent for Reaction Condition Recommendation in Chemical Synthesis 16 Nov 2023 · 0 repositories · arXiv:2311.10776
-
An Eye on Clinical BERT: Investigating Language Model Generalization for Diabetic Eye Disease Phenotyping 15 Nov 2023 · 1 repository · arXiv:2311.08687
-
Can Large Language Models Follow Concept Annotation Guidelines? A Case Study on Scientific and Financial Domains 15 Nov 2023 · 1 repository · arXiv:2311.08704
-
Evaluating Gender Bias in the Translation of Gender-Neutral Languages into English 15 Nov 2023 · 0 repositories · arXiv:2311.08836
-
Exponentially Faster Language Modelling 15 Nov 2023 · 3 repositories · arXiv:2311.10770
-
German FinBERT: A German Pre-trained Language Model 15 Nov 2023 · 0 repositories · arXiv:2311.08793
-
LOKE: Linked Open Knowledge Extraction for Automated Knowledge Graph Construction 15 Nov 2023 · 0 repositories · arXiv:2311.09366
-
Improving Deep Learning Optimization through Constrained Parameter Regularization 15 Nov 2023 · 1 repository · arXiv:2311.09058Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 9 unverified (of 13 harvested samples)
-
ToolTalk: Evaluating Tool-Usage in a Conversational Setting 15 Nov 2023 · 1 repository · arXiv:2311.10775Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
"We Demand Justice!": Towards Social Context Grounding of Political Texts 15 Nov 2023 · 1 repository · arXiv:2311.09106
-
XplainLLM: A Knowledge-Augmented Dataset for Reliable Grounded Explanations in LLMs 15 Nov 2023 · 1 repository · arXiv:2311.08614
-
Unifying the Perspectives of NLP and Software Engineering: A Survey on Language Models for Code 14 Nov 2023 · 1 repository · arXiv:2311.07989
-
AI-generated text boundary detection with RoFT 14 Nov 2023 · 1 repository · arXiv:2311.08349Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
CPopQA: Ranking Cultural Concept Popularity by LLMs 14 Nov 2023 · 0 repositories · arXiv:2311.07897
-
Evaluating LLMs on Document-Based QA: Exact Answer Selection and Numerical Extraction using Cogtale dataset 14 Nov 2023 · 0 repositories · arXiv:2311.07878
-
Exploring Semi-supervised Hierarchical Stacked Encoder for Legal Judgement Prediction 14 Nov 2023 · 1 repository · arXiv:2311.08103
-
Fair Abstractive Summarization of Diverse Perspectives 14 Nov 2023 · 1 repository · arXiv:2311.07884
-
Investigating the Encoding of Words in BERT's Neurons using Feature Textualization 14 Nov 2023 · 0 repositories · arXiv:2311.08240
-
Language Models are Better Bug Detector Through Code-Pair Classification 14 Nov 2023 · 1 repository · arXiv:2311.07957
-
Large Language Model-Driven Classroom Flipping: Empowering Student-Centric Peer Questioning with Flipped Interaction 14 Nov 2023 · 0 repositories · arXiv:2311.14708
-
Memory-efficient Stochastic methods for Memory-based Transformers 14 Nov 2023 · 1 repository · arXiv:2311.08123
-
Do large language models and humans have similar behaviors in causal inference with script knowledge? 13 Nov 2023 · 1 repository · arXiv:2311.07311
-
In-context Learning Generalizes, But Not Always Robustly: The Case of Syntax 13 Nov 2023 · 1 repository · arXiv:2311.07811
-
It's Not Easy Being Wrong: Large Language Models Struggle with Process of Elimination Reasoning 13 Nov 2023 · 1 repository · arXiv:2311.07532
-
Language Model-In-The-Loop: Data Optimal Approach to Learn-To-Recommend Actions in Text Games 13 Nov 2023 · 0 repositories · arXiv:2311.07687
-
MEGAVERSE: Benchmarking Large Language Models Across Languages, Modalities, Models and Tasks 13 Nov 2023 · 0 repositories · arXiv:2311.07463
-
On The Truthfulness of 'Surprisingly Likely' Responses of Large Language Models 13 Nov 2023 · 0 repositories · arXiv:2311.07692
-
Speech-based Slot Filling using Large Language Models 13 Nov 2023 · 0 repositories · arXiv:2311.07418
-
STEER: Unified Style Transfer with Expert Reinforcement 13 Nov 2023 · 1 repository · arXiv:2311.07167