Methods › General › Regularization › Weight Decay › Papers, page 50
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 50 of 108: papers 4,901 to 5,000 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
On "Scientific Debt" in NLP: A Case for More Rigour in Language Model Pre-Training Research 5 Jun 2023 · 0 repositories · arXiv:2306.02870
-
Skill over Scale: The Case for Medium, Domain-Specific Models for SE 5 Jun 2023 · 0 repositories · arXiv:2306.03268
-
Using Sequences of Life-events to Predict Human Lives 5 Jun 2023 · 2 repositories · arXiv:2306.03009
-
Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions 4 Jun 2023 · 1 repository · arXiv:2306.02224Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
SpellMapper: A non-autoregressive neural spellchecker for ASR customization with candidate retrieval based on n-gram mappings 4 Jun 2023 · 1 repository · arXiv:2306.02317
-
Financial sentiment analysis using FinBERT with application in predicting stock movement 3 Jun 2023 · 0 repositories · arXiv:2306.02136
-
MultiLegalPile: A 689GB Multilingual Legal Corpus 3 Jun 2023 · 0 repositories · arXiv:2306.02069
-
Towards Coding Social Science Datasets with Language Models 3 Jun 2023 · 0 repositories · arXiv:2306.02177
-
Can Contextual Biasing Remain Effective with Whisper and GPT-2? 2 Jun 2023 · 1 repository · arXiv:2306.01942
-
Concurrent Classifier Error Detection (CCED) in Large Scale Machine Learning Systems 2 Jun 2023 · 0 repositories · arXiv:2306.01820
-
Establishment of NLP-Based Greenwashing Pattern Detection Service 2 Jun 2023 · 0 repositories
-
Context selectivity with dynamic availability enables lifelong continual learning 2 Jun 2023 · 1 repository · arXiv:2306.01690
-
Word Embeddings for Banking Industry 2 Jun 2023 · 0 repositories · arXiv:2306.01807
-
Adapting Pre-trained Language Models to Vision-Language Tasks via Dynamic Visual Prompting 1 Jun 2023 · 1 repository · arXiv:2306.00409
-
Automatic Glossary of Clinical Terminology: a Large-Scale Dictionary of Biomedical Definitions Generated from Ontological Knowledge 1 Jun 2023 · 0 repositories · arXiv:2306.00665
-
Boosting the Performance of Transformer Architectures for Semantic Textual Similarity 1 Jun 2023 · 0 repositories · arXiv:2306.00708
-
Column Type Annotation using ChatGPT 1 Jun 2023 · 1 repository · arXiv:2306.00745
-
Enhancing Programming eTextbooks with ChatGPT Generated Counterfactual-Thinking-Inspired Questions 1 Jun 2023 · 0 repositories · arXiv:2306.00551
-
Feature Engineering-Based Detection of Buffer Overflow Vulnerability in Source Code Using Neural Networks 1 Jun 2023 · 0 repositories · arXiv:2306.07981
-
Make Pre-trained Model Reversible: From Parameter to Memory Efficient Fine-Tuning 1 Jun 2023 · 1 repository · arXiv:2306.00477Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-Dimensional Evaluation of Text Summarization with In-Context Learning 1 Jun 2023 · 1 repository · arXiv:2306.01200
-
Systematic Evaluation of GPT-3 for Zero-Shot Personality Estimation 1 Jun 2023 · 0 repositories · arXiv:2306.01183
-
TopEx: Topic-based Explanations for Model Comparison 1 Jun 2023 · 0 repositories · arXiv:2306.00976
-
Training-free Neural Architecture Search for RNNs and Transformers 1 Jun 2023 · 1 repository · arXiv:2306.00288Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
UCAS-IIE-NLP at SemEval-2023 Task 12: Enhancing Generalization of Multilingual BERT for Low-resource Sentiment Analysis 1 Jun 2023 · 1 repository · arXiv:2306.01093
-
Supplementary Features of BiLSTM for Enhanced Sequence Labeling 31 May 2023 · 1 repository · arXiv:2305.19928
-
Building Extractive Question Answering System to Support Human-AI Health Coaching Model for Sleep Domain 31 May 2023 · 0 repositories · arXiv:2305.19707
-
Catalysis distillation neural network for the few shot open catalyst challenge 31 May 2023 · 0 repositories · arXiv:2305.19545
-
DeepMerge: Deep-Learning-Based Region-Merging for Image Segmentation 31 May 2023 · 1 repository · arXiv:2305.19787
-
Evaluating GPT's Programming Capability through CodeWars' Katas 31 May 2023 · 0 repositories · arXiv:2306.01784
-
Examining the Emergence of Deductive Reasoning in Generative Language Models 31 May 2023 · 0 repositories · arXiv:2306.01009
-
Harnessing Explanations: LLM-to-LM Interpreter for Enhanced Text-Attributed Graph Representation Learning 31 May 2023 · 3 repositories · arXiv:2305.19523Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Knowledge Base Question Answering for Space Debris Queries 31 May 2023 · 1 repository · arXiv:2305.19734
-
XPhoneBERT: A Pre-trained Multilingual Model for Phoneme Representations for Text-to-Speech 31 May 2023 · 2 repositories · arXiv:2305.19709Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples) · 10 pointer-only (licence)
-
Does Conceptual Representation Require Embodiment? Insights From Large Language Models 30 May 2023 · 0 repositories · arXiv:2305.19103
-
Explaining Hate Speech Classification with Model Agnostic Methods 30 May 2023 · 0 repositories · arXiv:2306.00021
-
Generate then Select: Open-ended Visual Question Answering Guided by World Knowledge 30 May 2023 · 0 repositories · arXiv:2305.18842
-
GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation 30 May 2023 · 0 repositories · arXiv:2305.18997
-
Multitask learning for recognizing stress and depression in social media 30 May 2023 · 0 repositories · arXiv:2305.18907
-
PreQuant: A Task-agnostic Quantization Approach for Pre-trained Language Models 30 May 2023 · 0 repositories · arXiv:2306.00014
-
Research on Multilingual News Clustering Based on Cross-Language Word Embeddings 30 May 2023 · 0 repositories · arXiv:2305.18880
-
ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning 30 May 2023 · 1 repository · arXiv:2305.19426
-
Seeing Seeds Beyond Weeds: Green Teaming Generative AI for Beneficial Uses 30 May 2023 · 0 repositories · arXiv:2306.03097
-
Abstractive Summarization as Augmentation for Document-Level Event Detection 29 May 2023 · 0 repositories · arXiv:2305.18023
-
Check-COVID: Fact-Checking COVID-19 News Claims with Scientific Evidence 29 May 2023 · 1 repository · arXiv:2305.18265
-
Coeditor: Leveraging Contextual Changes for Multi-round Code Auto-editing 29 May 2023 · 0 repositories · arXiv:2305.18584
-
Do Large Language Models Know What They Don't Know? 29 May 2023 · 1 repository · arXiv:2305.18153Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Exploring Effectiveness of GPT-3 in Grammatical Error Correction: A Study on Performance and Controllability in Prompt-Based Methods 29 May 2023 · 0 repositories · arXiv:2305.18156
-
From Adversarial Arms Race to Model-centric Evaluation: Motivating a Unified Automatic Robustness Evaluation Framework 29 May 2023 · 1 repository · arXiv:2305.18503Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LM-CPPF: Paraphrasing-Guided Data Augmentation for Contrastive Prompt-Based Few-Shot Fine-Tuning 29 May 2023 · 1 repository · arXiv:2305.18169
-
Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models 29 May 2023 · 1 repository · arXiv:2305.18189Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ProcessGPT: Transforming Business Process Management with Generative Artificial Intelligence 29 May 2023 · 0 repositories · arXiv:2306.01771
-
SlimFit: Memory-Efficient Fine-Tuning of Transformer-based Models Using Training Dynamics 29 May 2023 · 0 repositories · arXiv:2305.18513
-
Syntax and Semantics Meet in the "Middle": Probing the Syntax-Semantics Interface of LMs Through Agentivity 29 May 2023 · 1 repository · arXiv:2305.18185
-
Test-Time Training on Nearest Neighbors for Large Language Models 29 May 2023 · 1 repository · arXiv:2305.18466Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Transformer Language Models Handle Word Frequency in Prediction Head 29 May 2023 · 0 repositories · arXiv:2305.18294
-
Bridging the Language Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs 28 May 2023 · 0 repositories · arXiv:2305.17740
-
Evaluating GPT-3 Generated Explanations for Hateful Content Moderation 28 May 2023 · 1 repository · arXiv:2305.17680
-
Generating EDU Extracts for Plan-Guided Summary Re-Ranking 28 May 2023 · 1 repository · arXiv:2305.17779Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Knowledge-Augmented Reasoning Distillation for Small Language Models in Knowledge-Intensive Tasks 28 May 2023 · 1 repository · arXiv:2305.18395Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model Application 28 May 2023 · 1 repository · arXiv:2305.17701
-
Mitigating Label Biases for In-context Learning 28 May 2023 · 1 repository · arXiv:2305.19148Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Rethinking Masked Language Modeling for Chinese Spelling Correction 28 May 2023 · 1 repository · arXiv:2305.17721
-
SQuARe: A Large-Scale Dataset of Sensitive Questions and Acceptable Responses Created Through Human-Machine Collaboration 28 May 2023 · 1 repository · arXiv:2305.17696
-
Transfer Learning for Power Outage Detection Task with Limited Training Data 28 May 2023 · 0 repositories · arXiv:2305.17817
-
Diagnosing Transformers: Illuminating Feature Spaces for Clinical Decision-Making 27 May 2023 · 1 repository · arXiv:2305.17588
-
Complementary and Integrative Health Lexicon (CIHLex) and Entity Recognition in the Literature 27 May 2023 · 0 repositories · arXiv:2305.17353
-
DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text 27 May 2023 · 1 repository · arXiv:2305.17359Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
The Curse of Recursion: Training on Generated Data Makes Models Forget 27 May 2023 · 1 repository · arXiv:2305.17493
-
Modeling Adversarial Attack on Pre-trained Language Models as Sequential Decision Making 27 May 2023 · 1 repository · arXiv:2305.17440
-
Reinforcement Learning With Reward Machines in Stochastic Games 27 May 2023 · 0 repositories · arXiv:2305.17372
-
Towards Explainable Conversational Recommender Systems 27 May 2023 · 1 repository · arXiv:2305.18363
-
What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks 27 May 2023 · 1 repository · arXiv:2305.18365
-
Backpack Language Models 26 May 2023 · 1 repository · arXiv:2305.16765Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
Calibration of Transformer-based Models for Identifying Stress and Depression in Social Media 26 May 2023 · 0 repositories · arXiv:2305.16797
-
Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance 26 May 2023 · 1 repository · arXiv:2305.17306Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
ChatGPT: A Study on its Utility for Ubiquitous Software Engineering Tasks 26 May 2023 · 0 repositories · arXiv:2305.16837
-
Counterfactual reasoning: Testing language models' understanding of hypothetical scenarios 26 May 2023 · 1 repository · arXiv:2305.16572
-
Distinguishing Human Generated Text From ChatGPT Generated Text Using Machine Learning 26 May 2023 · 0 repositories · arXiv:2306.01761
-
Do GPTs Produce Less Literal Translations? 26 May 2023 · 1 repository · arXiv:2305.16806
-
Evaluation of Question Generation Needs More References 26 May 2023 · 0 repositories · arXiv:2305.16626
-
Exploring Weight Balancing on Long-Tailed Recognition Problem 26 May 2023 · 1 repository · arXiv:2305.16573
-
GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot Attention for Vision-and-Language Navigation 26 May 2023 · 1 repository · arXiv:2305.17102
-
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing 26 May 2023 · 0 repositories · arXiv:2305.16635
-
Improving accuracy of GPT-3/4 results on biomedical data using a retrieval-augmented language model 26 May 2023 · 0 repositories · arXiv:2305.17116
-
Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization 26 May 2023 · 0 repositories · arXiv:2305.16784
-
KNSE: A Knowledge-aware Natural Language Inference Framework for Dialogue Symptom Status Recognition 26 May 2023 · 0 repositories · arXiv:2305.16833
-
Large Language Models as Tool Makers 26 May 2023 · 1 repository · arXiv:2305.17126
-
Learning and Leveraging Verifiers to Improve Planning Capabilities of Pre-trained Language Models 26 May 2023 · 0 repositories · arXiv:2305.17077
-
LLMs and the Abstraction and Reasoning Corpus: Successes, Failures, and the Importance of Object-based Representations 26 May 2023 · 1 repository · arXiv:2305.18354
-
NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models 26 May 2023 · 2 repositories · arXiv:2305.16986Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Playing repeated games with Large Language Models 26 May 2023 · 0 repositories · arXiv:2305.16867
-
Rotational Equilibrium: How Weight Decay Balances Learning Across Neural Networks 26 May 2023 · 2 repositories · arXiv:2305.17212Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Theoretical and Practical Perspectives on what Influence Functions Do 26 May 2023 · 0 repositories · arXiv:2305.16971
-
Zero is Not Hero Yet: Benchmarking Zero-Shot Performance of LLMs for Financial Tasks 26 May 2023 · 1 repository · arXiv:2305.16633
-
A Survey on ChatGPT: AI-Generated Contents, Challenges, and Solutions 25 May 2023 · 0 repositories · arXiv:2305.18339
-
Comparative Study of Pre-Trained BERT Models for Code-Mixed Hindi-English Data 25 May 2023 · 0 repositories · arXiv:2305.15722
-
Context-aware attention layers coupled with optimal transport domain adaptation and multimodal fusion methods for recognizing dementia from spontaneous speech 25 May 2023 · 0 repositories · arXiv:2305.16406
-
Linguistic Properties of Truthful Response 25 May 2023 · 1 repository · arXiv:2305.15875
-
Not wacky vs. definitely wacky: A study of scalar adverbs in pretrained language models 25 May 2023 · 0 repositories · arXiv:2305.16426