Methods › General › Regularization › Weight Decay › Papers, page 58
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 58 of 108: papers 5,701 to 5,800 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
PromptCap: Prompt-Guided Image Captioning for VQA with GPT-3 1 Jan 2023 · 0 repositories
-
Rethinking with Retrieval: Faithful Large Language Model Inference 31 Dec 2022 · 1 repository · arXiv:2301.00303
-
Sentiment Analysis of COVID-19 Public Activity Restriction (PPKM) Impact using BERT Method 31 Dec 2022 · 0 repositories · arXiv:2301.00096
-
An Analysis of Attention via the Lens of Exchangeability and Latent Variable Models 30 Dec 2022 · 0 repositories · arXiv:2212.14852
-
Distant Reading of the German Coalition Deal: Recognizing Policy Positions with BERT-based Text Classification 30 Dec 2022 · 0 repositories · arXiv:2212.14648
-
Targeted Phishing Campaigns using Large Scale Language Models 30 Dec 2022 · 0 repositories · arXiv:2301.00665
-
GPT Takes the Bar Exam 29 Dec 2022 · 5 repositories · arXiv:2212.14402
-
Maximizing Use-Case Specificity through Precision Model Tuning 29 Dec 2022 · 0 repositories · arXiv:2212.14206
-
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces 29 Dec 2022 · 0 repositories · arXiv:2301.00009
-
Cramming: Training a Language Model on a Single GPU in One Day 28 Dec 2022 · 1 repository · arXiv:2212.14034Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Survey on Knowledge-Enhanced Pre-trained Language Models 27 Dec 2022 · 0 repositories · arXiv:2212.13428
-
DeepCuts: Single-Shot Interpretability based Pruning for BERT 27 Dec 2022 · 1 repository · arXiv:2212.13392
-
TegFormer: Topic-to-Essay Generation with Good Topic Coverage and High Text Coherence 27 Dec 2022 · 0 repositories · arXiv:2212.13456
-
Using Large Language Models to Generate Engaging Captions for Data Visualizations 27 Dec 2022 · 0 repositories · arXiv:2212.14047
-
Biologically Inspired Design Concept Generation Using Generative Pre-Trained Transformers 26 Dec 2022 · 0 repositories · arXiv:2212.13196
-
Benchmark for Uncertainty & Robustness in Self-Supervised Learning 23 Dec 2022 · 1 repository · arXiv:2212.12411
-
Finetuning for Sarcasm Detection with a Pruned Dataset 23 Dec 2022 · 1 repository · arXiv:2212.12213
-
Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times? 23 Dec 2022 · 0 repositories · arXiv:2212.12131
-
CAMeMBERT: Cascading Assistant-Mediated Multilingual BERT 22 Dec 2022 · 0 repositories · arXiv:2212.11456
-
Enhancing the prediction of disease outcomes using electronic health records and pretrained deep learning models 22 Dec 2022 · 0 repositories · arXiv:2212.12067
-
Analyzing Semantic Faithfulness of Language Models via Input Intervention on Question Answering 21 Dec 2022 · 1 repository · arXiv:2212.10696
-
Automatic Emotion Modelling in Written Stories 21 Dec 2022 · 1 repository · arXiv:2212.11382
-
Cross-Linguistic Syntactic Difference in Multilingual BERT: How Good is It and How Does It Affect Transfer? 21 Dec 2022 · 1 repository · arXiv:2212.10879
-
Entropy- and Distance-Based Predictors From GPT-2 Attention Patterns Predict Reading Times Over and Above GPT-2 Surprisal 21 Dec 2022 · 1 repository · arXiv:2212.11185Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 8 pointer-only (licence)
-
JASMINE: Arabic GPT Models for Few-Shot Learning 21 Dec 2022 · 0 repositories · arXiv:2212.10755
-
KL Regularized Normalization Framework for Low Resource Tasks 21 Dec 2022 · 0 repositories · arXiv:2212.11275
-
Towards Efficient Visual Simplification of Computational Graphs in Deep Neural Networks 21 Dec 2022 · 0 repositories · arXiv:2212.10774
-
A Twitter BERT Approach for Offensive Language Detection in Marathi 20 Dec 2022 · 0 repositories · arXiv:2212.10039
-
ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language Models 20 Dec 2022 · 1 repository · arXiv:2212.10474Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Controllable Text Generation with Language Constraints 20 Dec 2022 · 0 repositories · arXiv:2212.10466
-
Do language models have coherent mental models of everyday things? 20 Dec 2022 · 1 repository · arXiv:2212.10029Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
DocAsRef: An Empirical Study on Repurposing Reference-Based Summary Quality Metrics Reference-Freely 20 Dec 2022 · 1 repository · arXiv:2212.10013
-
Generic Temporal Reasoning with Differential Analysis and Explanation 20 Dec 2022 · 0 repositories · arXiv:2212.10467
-
Go-tuning: Improving Zero-shot Learning Abilities of Smaller Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10461
-
Hybrid Rule-Neural Coreference Resolution System based on Actor-Critic Learning 20 Dec 2022 · 0 repositories · arXiv:2212.10087
-
Identifying and Manipulating the Personality Traits of Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10276
-
In and Out-of-Domain Text Adversarial Robustness via Label Smoothing 20 Dec 2022 · 0 repositories · arXiv:2212.10258
-
Is GPT-3 a Good Data Annotator? 20 Dec 2022 · 1 repository · arXiv:2212.10450
-
Evaluating Psychological Safety of Large Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10529
-
Large Language Models Are Reasoning Teachers 20 Dec 2022 · 1 repository · arXiv:2212.10071Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
PairReranker: Pairwise Reranking for Natural Language Generation 20 Dec 2022 · 0 repositories · arXiv:2212.10555
-
Parameter-efficient Zero-shot Transfer for Cross-Language Dense Retrieval with Adapters 20 Dec 2022 · 0 repositories · arXiv:2212.10448
-
Pay Attention to Your Tone: Introducing a New Dataset for Polite Language Rewrite 20 Dec 2022 · 1 repository · arXiv:2212.10190
-
Pretraining Without Attention 20 Dec 2022 · 1 repository · arXiv:2212.10544Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
True Detective: A Deep Abductive Reasoning Benchmark Undoable for GPT-3 and Challenging for GPT-4 20 Dec 2022 · 0 repositories · arXiv:2212.10114
-
Why Can GPT Learn In-Context? Language Models Implicitly Perform Gradient Descent as Meta-Optimizers 20 Dec 2022 · 1 repository · arXiv:2212.10559
-
Do CoNLL-2003 Named Entity Taggers Still Work Well in 2023? 19 Dec 2022 · 1 repository · arXiv:2212.09747
-
Emergent Analogical Reasoning in Large Language Models 19 Dec 2022 · 2 repositories · arXiv:2212.09196Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Enriching Relation Extraction with OpenIE 19 Dec 2022 · 0 repositories · arXiv:2212.09376
-
Evaluating Human-Language Model Interaction 19 Dec 2022 · 1 repository · arXiv:2212.09746
-
Large Language Models are Better Reasoners with Self-Verification 19 Dec 2022 · 1 repository · arXiv:2212.09561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
LENS: A Learnable Evaluation Metric for Text Simplification 19 Dec 2022 · 1 repository · arXiv:2212.09739
-
Less is More: Parameter-Free Text Classification with Gzip 19 Dec 2022 · 0 repositories · arXiv:2212.09410
-
MANTIS at TSAR-2022 Shared Task: Improved Unsupervised Lexical Simplification with Pretrained Encoders 19 Dec 2022 · 0 repositories · arXiv:2212.09855
-
Reasoning with Language Model Prompting: A Survey 19 Dec 2022 · 2 repositories · arXiv:2212.09597
-
The case for 4-bit precision: k-bit Inference Scaling Laws 19 Dec 2022 · 1 repository · arXiv:2212.09720
-
Bort: Towards Explainable Neural Networks with Bounded Orthogonal Constraint 18 Dec 2022 · 1 repository · arXiv:2212.09062Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can Retriever-Augmented Language Models Reason? The Blame Game Between the Retriever and the Language Model 18 Dec 2022 · 1 repository · arXiv:2212.09146
-
Neural Coreference Resolution based on Reinforcement Learning 18 Dec 2022 · 0 repositories · arXiv:2212.09028
-
Neural Rankers for Effective Screening Prioritisation in Medical Systematic Review Literature Search 18 Dec 2022 · 0 repositories · arXiv:2212.09017
-
Exploiting Rich Textual User-Product Context for Improving Sentiment Analysis 17 Dec 2022 · 0 repositories · arXiv:2212.08888
-
LegalRelectra: Mixed-domain Language Modeling for Long-range Legal Text Comprehension 16 Dec 2022 · 0 repositories · arXiv:2212.08204
-
MURMUR: Modular Multi-Step Reasoning for Semi-Structured Data-to-Text Generation 16 Dec 2022 · 0 repositories · arXiv:2212.08607
-
Plansformer: Generating Symbolic Plans using Transformers 16 Dec 2022 · 0 repositories · arXiv:2212.08681
-
POIBERT: A Transformer-based Model for the Tour Recommendation Problem 16 Dec 2022 · 0 repositories · arXiv:2212.13900
-
Preventing RNN from Using Sequence Length as a Feature 16 Dec 2022 · 0 repositories · arXiv:2212.08276
-
ReCo: Reliable Causal Chain Reasoning via Structural Causal Recurrent Neural Networks 16 Dec 2022 · 1 repository · arXiv:2212.08322Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA 16 Dec 2022 · 1 repository · arXiv:2212.08635
-
Utilizing distilBert transformer model for sentiment classification of COVID-19's Persian open-text responses 16 Dec 2022 · 0 repositories · arXiv:2212.08407
-
Efficient Pre-training of Masked Language Model via Concept-based Curriculum Masking 15 Dec 2022 · 1 repository · arXiv:2212.07617Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Revisiting the Gold Standard: Grounding Summarization Evaluation with Robust Human Evaluation 15 Dec 2022 · 2 repositories · arXiv:2212.07981Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Visually-augmented pretrained language models for NLP tasks without images 15 Dec 2022 · 1 repository · arXiv:2212.07937
-
Efficient Self-supervised Learning with Contextualized Target Representations for Vision, Speech and Language 14 Dec 2022 · 5 repositories · arXiv:2212.07525Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Explainability of Text Processing and Retrieval Methods: A Critical Survey 14 Dec 2022 · 0 repositories · arXiv:2212.07126
-
PAC-MAN: Multi-Relation Network in Social Community for Personalized Hashtag Recommendation 14 Dec 2022 · 1 repository
-
CREPE: Can Vision-Language Foundation Models Reason Compositionally? 13 Dec 2022 · 1 repository · arXiv:2212.07796Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Paraphrase Identification with Deep Learning: A Review of Datasets and Methods 13 Dec 2022 · 0 repositories · arXiv:2212.06933
-
DexBERT: Effective, Task-Agnostic and Fine-grained Representation Learning of Android Bytecode 12 Dec 2022 · 1 repository · arXiv:2212.05976
-
Classifying the Ideological Orientation of User-Submitted Texts in Social Media 12 Dec 2022 · 1 repository
-
Off-Policy Deep Reinforcement Learning Algorithms for Handling Various Robotic Manipulator Tasks 11 Dec 2022 · 0 repositories · arXiv:2212.05572
-
Elixir: Train a Large Language Model on a Small GPU Cluster 10 Dec 2022 · 2 repositories · arXiv:2212.05339
-
Thinking Fast and Slow in Large Language Models 10 Dec 2022 · 0 repositories · arXiv:2212.05206
-
Punctuation Restoration for Singaporean Spoken Languages: English, Malay, and Mandarin 10 Dec 2022 · 1 repository · arXiv:2212.05356
-
Structured information extraction from complex scientific text with fine-tuned large language models 10 Dec 2022 · 0 repositories · arXiv:2212.05238
-
Incorporating Emotions into Health Mention Classification Task on Social Media 9 Dec 2022 · 1 repository · arXiv:2212.05039
-
The Turing Deception 9 Dec 2022 · 0 repositories · arXiv:2212.06721
-
TRBLLmaker -- Transformer Reads Between Lyrics Lines maker 9 Dec 2022 · 0 repositories · arXiv:2212.04917
-
Explain to me like I am five -- Sentence Simplification Using Transformers 8 Dec 2022 · 1 repository · arXiv:2212.04595
-
LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models 8 Dec 2022 · 1 repository · arXiv:2212.04088
-
NP4G : Network Programming for Generalization 8 Dec 2022 · 1 repository · arXiv:2212.11118
-
The Role of AI in Drug Discovery: Challenges, Opportunities, and Strategies 8 Dec 2022 · 0 repositories · arXiv:2212.08104
-
Memorization of Named Entities in Fine-tuned BERT Models 7 Dec 2022 · 1 repository · arXiv:2212.03749
-
DeepSpeed Data Efficiency: Improving Deep Learning Model Quality and Training Efficiency via Efficient Data Sampling and Routing 7 Dec 2022 · 1 repository · arXiv:2212.03597
-
Learning-To-Embed: Adopting Transformer based models for E-commerce Products Representation Learning 7 Dec 2022 · 0 repositories · arXiv:2212.03725
-
SimVTP: Simple Video Text Pre-training with Masked Autoencoders 7 Dec 2022 · 0 repositories · arXiv:2212.03490
-
TweetDrought: A Deep-Learning Drought Impacts Recognizer based on Twitter Data 7 Dec 2022 · 0 repositories · arXiv:2212.04001
-
Adaptive Testing of Computer Vision Models 6 Dec 2022 · 1 repository · arXiv:2212.02774Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Counterfactual reasoning: Do language models need world knowledge for causal understanding? 6 Dec 2022 · 1 repository · arXiv:2212.03278
-
CySecBERT: A Domain-Adapted Language Model for the Cybersecurity Domain 6 Dec 2022 · 0 repositories · arXiv:2212.02974
-
Vision Transformer Computation and Resilience for Dynamic Inference 6 Dec 2022 · 0 repositories · arXiv:2212.02687