Methods › General › Regularization › Weight Decay › Papers, page 75
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 75 of 108: papers 7,401 to 7,500 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Model Bias in NLP -- Application to Hate Speech Classification using transfer learning techniques 20 Sep 2021 · 0 repositories · arXiv:2109.09725
-
MirrorWiC: On Eliciting Word-in-Context Representations from Pretrained Language Models 19 Sep 2021 · 1 repository · arXiv:2109.09237
-
Towards Zero-Label Language Learning 19 Sep 2021 · 0 repositories · arXiv:2109.09193
-
Wav-BERT: Cooperative Acoustic and Linguistic Representation Learning for Low-Resource Speech Recognition 19 Sep 2021 · 0 repositories · arXiv:2109.09161
-
What BERT Based Language Models Learn in Spoken Transcripts: An Empirical Study 19 Sep 2021 · 0 repositories · arXiv:2109.09105
-
Complex Temporal Question Answering on Knowledge Graphs 18 Sep 2021 · 1 repository · arXiv:2109.08935
-
DyLex: Incorporating Dynamic Lexicons into BERT for Sequence Labeling 18 Sep 2021 · 1 repository · arXiv:2109.08818
-
Text Detoxification using Large Pre-trained Neural Models 18 Sep 2021 · 1 repository · arXiv:2109.08914Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Commonsense Knowledge-Augmented Pretrained Language Models for Causal Reasoning Classification 17 Sep 2021 · 0 repositories
-
Context vs Target Word: Quantifying Biases When Applying Models to Lexical Semantic Datasets 17 Sep 2021 · 0 repositories
-
Deep Reinforcement Learning Based Multidimensional Resource Management for Energy Harvesting Cognitive NOMA Communications 17 Sep 2021 · 0 repositories · arXiv:2109.09503
-
Defending Textual Neural Networks against Black-Box Adversarial Attacks with Stochastic Multi-Expert Patcher 17 Sep 2021 · 0 repositories
-
Does BERT really agree ? Fine-grained Analysis of Lexical Dependence on a Syntactic Task 17 Sep 2021 · 0 repositories
-
Fine-Tuned Transformers Show Clusters of Similar Representations Across Layers 17 Sep 2021 · 0 repositories · arXiv:2109.08406
-
Grounding Natural Language Instructions: Can Large Language Models Capture Spatial Information? 17 Sep 2021 · 1 repository · arXiv:2109.08634
-
Knowledge Neurons in Pretrained Transformers 17 Sep 2021 · 0 repositories
-
Learning Low-frequency Patterns with A Pre-trained Document-Grounded Conversation Model 17 Sep 2021 · 0 repositories
-
General Cross-Architecture Distillation of Pretrained Language Models into Matrix Embeddings 17 Sep 2021 · 1 repository · arXiv:2109.08449
-
Primer: Searching for Efficient Transformers for Language Modeling 17 Sep 2021 · 4 repositories · arXiv:2109.08668Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Relating Neural Text Degeneration to Exposure Bias 17 Sep 2021 · 0 repositories · arXiv:2109.08705
-
The futility of STILTs for the classification of lexical borrowings in Spanish 17 Sep 2021 · 0 repositories · arXiv:2109.08607
-
Language Models are Few-shot Multilingual Learners 16 Sep 2021 · 1 repository · arXiv:2109.07684Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Let the CAT out of the bag: Contrastive Attributed explanations for Text 16 Sep 2021 · 0 repositories · arXiv:2109.07983
-
RetrievalSum: A Retrieval Enhanced Framework for Abstractive Summarization 16 Sep 2021 · 0 repositories · arXiv:2109.07943
-
Revisiting Tri-training of Dependency Parsers 16 Sep 2021 · 2 repositories · arXiv:2109.08122
-
BERT is Robust! A Case Against Synonym-Based Adversarial Examples in Text Classification 15 Sep 2021 · 0 repositories · arXiv:2109.07403
-
EfficientBERT: Progressively Searching Multilayer Perceptron via Warm-up Knowledge Distillation 15 Sep 2021 · 1 repository · arXiv:2109.07222Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Efficient Domain Adaptation of Language Models via Adaptive Tokenization 15 Sep 2021 · 0 repositories · arXiv:2109.07460
-
Enhancing Clinical Information Extraction with Transferred Contextual Embeddings 15 Sep 2021 · 0 repositories · arXiv:2109.07243
-
Improving Text Auto-Completion with Next Phrase Prediction 15 Sep 2021 · 0 repositories · arXiv:2109.07067
-
Learning to Match Job Candidates Using Multilingual Bi-Encoder BERT 15 Sep 2021 · 0 repositories · arXiv:2109.07157
-
On the Universality of Deep Contextual Language Models 15 Sep 2021 · 0 repositories · arXiv:2109.07140
-
The Unreasonable Effectiveness of the Baseline: Discussing SVMs in Legal Text Classification 15 Sep 2021 · 0 repositories · arXiv:2109.07234
-
Topic Transferable Table Question Answering 15 Sep 2021 · 1 repository · arXiv:2109.07377
-
Transformer-based Language Models for Factoid Question Answering at BioASQ9b 15 Sep 2021 · 1 repository · arXiv:2109.07185
-
A Temporal Variational Model for Story Generation 14 Sep 2021 · 3 repositories · arXiv:2109.06807
-
conSultantBERT: Fine-tuned Siamese Sentence-BERT for Matching Jobs and Job Seekers 14 Sep 2021 · 0 repositories · arXiv:2109.06501
-
Deep learning-based NLP Data Pipeline for EHR Scanned Document Information Extraction 14 Sep 2021 · 0 repositories · arXiv:2110.11864
-
Evaluating Biomedical BERT Models for Vocabulary Alignment at Scale in the UMLS Metathesaurus 14 Sep 2021 · 0 repositories · arXiv:2109.13348
-
Explainable Identification of Dementia from Transcripts using Transformer Networks 14 Sep 2021 · 0 repositories · arXiv:2109.06980
-
Exploring Personality and Online Social Engagement: An Investigation of MBTI Users on Twitter 14 Sep 2021 · 0 repositories · arXiv:2109.06402
-
Frequency Effects on Syntactic Rule Learning in Transformers 14 Sep 2021 · 1 repository · arXiv:2109.07020
-
Learning Bill Similarity with Annotated and Augmented Corpora of Bills 14 Sep 2021 · 1 repository · arXiv:2109.06527
-
Legal Transformer Models May Not Always Help 14 Sep 2021 · 0 repositories · arXiv:2109.06862
-
On the Language-specificity of Multilingual BERT and the Impact of Fine-tuning 14 Sep 2021 · 1 repository · arXiv:2109.06935
-
Semantic Answer Type Prediction using BERT: IAI at the ISWC SMART Task 2020 14 Sep 2021 · 0 repositories · arXiv:2109.06714
-
Tribrid: Stance Classification with Neural Inconsistency Detection 14 Sep 2021 · 1 repository · arXiv:2109.06508
-
YES SIR!Optimizing Semantic Space of Negatives with Self-Involvement Ranker 14 Sep 2021 · 0 repositories · arXiv:2109.06436
-
Effectiveness of Pre-training for Few-shot Intent Classification 13 Sep 2021 · 0 repositories · arXiv:2109.05782
-
Evaluating Transferability of BERT Models on Uralic Languages 13 Sep 2021 · 1 repository · arXiv:2109.06327
-
Exploring a Unified Sequence-To-Sequence Transformer for Medical Product Safety Monitoring in Social Media 13 Sep 2021 · 1 repository · arXiv:2109.05815
-
Keyword Extraction for Improved Document Retrieval in Conversational Search 13 Sep 2021 · 0 repositories · arXiv:2109.05979
-
Mitigating Language-Dependent Ethnic Bias in BERT 13 Sep 2021 · 1 repository · arXiv:2109.05704Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Not All Models Localize Linguistic Knowledge in the Same Place: A Layer-wise Probing on BERToids' Representations 13 Sep 2021 · 0 repositories · arXiv:2109.05958
-
Connecting degree and polarity: An artificial language learning study 13 Sep 2021 · 1 repository · arXiv:2109.06333
-
Phrase-BERT: Improved Phrase Embeddings from BERT with an Application to Corpus Exploration 13 Sep 2021 · 2 repositories · arXiv:2109.06304Syntology official (archive's flag): 2 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 1 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Question Answering over Electronic Devices: A New Benchmark Dataset and a Multi-Task Learning based QA Framework 13 Sep 2021 · 1 repository · arXiv:2109.05897Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
FLiText: A Faster and Lighter Semi-Supervised Text Classification with Convolution Networks 12 Sep 2021 · 1 repository · arXiv:2110.11869
-
TEASEL: A Transformer-Based Speech-Prefixed Language Model 12 Sep 2021 · 1 repository · arXiv:2109.05522Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Clinical Trial Information Extraction with BERT 11 Sep 2021 · 0 repositories · arXiv:2110.10027
-
Multilingual Translation via Grafting Pre-trained Language Models 11 Sep 2021 · 1 repository · arXiv:2109.05256
-
TopicRefine: Joint Topic Prediction and Dialogue Response Generation for Multi-turn End-to-End Dialogue System 11 Sep 2021 · 0 repositories · arXiv:2109.05187
-
An Empirical Study of GPT-3 for Few-Shot Knowledge-Based VQA 10 Sep 2021 · 1 repository · arXiv:2109.05014Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
An Exploratory Study on Long Dialogue Summarization: What Works and What's Next 10 Sep 2021 · 1 repository · arXiv:2109.04609
-
Artificial Text Detection via Examining the Topology of Attention Maps 10 Sep 2021 · 2 repositories · arXiv:2109.04825Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 8 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Block Pruning For Faster Transformers 10 Sep 2021 · 1 repository · arXiv:2109.04838
-
D-REX: Dialogue Relation Extraction with Explanations 10 Sep 2021 · 1 repository · arXiv:2109.05126
-
Enhancing Self-Disclosure In Neural Dialog Models By Candidate Re-ranking 10 Sep 2021 · 0 repositories · arXiv:2109.05090
-
FBERT: A Neural Transformer for Identifying Offensive Content 10 Sep 2021 · 0 repositories · arXiv:2109.05074
-
How May I Help You? Using Neural Text Simplification to Improve Downstream NLP Tasks 10 Sep 2021 · 1 repository · arXiv:2109.04604
-
IndoBERTweet: A Pretrained Language Model for Indonesian Twitter with Effective Domain-Specific Vocabulary Initialization 10 Sep 2021 · 1 repository · arXiv:2109.04607
-
Mixture-of-Partitions: Infusing Large Biomedical Knowledge Graphs into BERT 10 Sep 2021 · 1 repository · arXiv:2109.04810Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
On the validity of pre-trained transformers for natural language processing in the software engineering domain 10 Sep 2021 · 0 repositories · arXiv:2109.04738
-
RoR: Read-over-Read for Long Document Machine Reading Comprehension 10 Sep 2021 · 1 repository · arXiv:2109.04780
-
What Changes Can Large-scale Language Models Bring? Intensive Study on HyperCLOVA: Billions-scale Korean Generative Pretrained Transformers 10 Sep 2021 · 2 repositories · arXiv:2109.04650Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational Quality 9 Sep 2021 · 1 repository · arXiv:2109.04404
-
BERT, mBERT, or BiBERT? A Study on Contextualized Embeddings for Neural Machine Translation 9 Sep 2021 · 2 repositories · arXiv:2109.04588Syntology official (archive's flag): 1 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Generalised Unsupervised Domain Adaptation of Neural Machine Translation with Cross-Lingual Data Selection 9 Sep 2021 · 1 repository · arXiv:2109.04292
-
Improving Video-Text Retrieval by Multi-Stream Corpus Alignment and Dual Softmax Loss 9 Sep 2021 · 2 repositories · arXiv:2109.04290
-
KELM: Knowledge Enhanced Pre-Trained Language Representations with Message Passing on Hierarchical Relational Graphs 9 Sep 2021 · 1 repository · arXiv:2109.04223Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Medically Aware GPT-3 as a Data Generator for Medical Dialogue Summarization 9 Sep 2021 · 0 repositories · arXiv:2110.07356
-
Mining Points of Interest via Address Embeddings: An Unsupervised Approach 9 Sep 2021 · 0 repositories · arXiv:2109.04467
-
Multi-granularity Textual Adversarial Attack with Behavior Cloning 9 Sep 2021 · 1 repository · arXiv:2109.04367
-
Variational Latent-State GPT for Semi-Supervised Task-Oriented Dialog Systems 9 Sep 2021 · 2 repositories · arXiv:2109.04314
-
Word-Level Coreference Resolution 9 Sep 2021 · 1 repository · arXiv:2109.04127
-
Ensemble Fine-tuned mBERT for Translation Quality Estimation 8 Sep 2021 · 0 repositories · arXiv:2109.03914
-
Bag-of-Words vs. Graph vs. Sequence in Text Classification: Questioning the Necessity of Text-Graphs and the Surprising Strength of a Wide MLP 8 Sep 2021 · 2 repositories · arXiv:2109.03777
-
NSP-BERT: A Prompt-based Few-Shot Learner Through an Original Pre-training Task--Next Sentence Prediction 8 Sep 2021 · 1 repository · arXiv:2109.03564
-
Sustainable Modular Debiasing of Language Models 8 Sep 2021 · 0 repositories · arXiv:2109.03646
-
TruthfulQA: Measuring How Models Mimic Human Falsehoods 8 Sep 2021 · 3 repositories · arXiv:2109.07958Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
BERT based classification system for detecting rumours on Twitter 7 Sep 2021 · 0 repositories · arXiv:2109.02975
-
Empathetic Dialogue Generation with Pre-trained RoBERTa-GPT2 and External Knowledge 7 Sep 2021 · 0 repositories · arXiv:2109.03004
-
FHAC at GermEval 2021: Identifying German toxic, engaging, and fact-claiming comments with ensemble learning 7 Sep 2021 · 1 repository · arXiv:2109.03094
-
How much pretraining data do language models need to learn syntax? 7 Sep 2021 · 0 repositories · arXiv:2109.03160
-
Naturalness Evaluation of Natural Language Generation in Task-oriented Dialogues using BERT 7 Sep 2021 · 0 repositories · arXiv:2109.02938
-
NumGPT: Improving Numeracy Ability of Generative Pre-trained Models 7 Sep 2021 · 0 repositories · arXiv:2109.03137
-
PAUSE: Positive and Annealed Unlabeled Sentence Embedding 7 Sep 2021 · 1 repository · arXiv:2109.03155Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Refining BERT Embeddings for Document Hashing via Mutual Information Maximization 7 Sep 2021 · 1 repository · arXiv:2109.02867
-
Text-Free Prosody-Aware Generative Spoken Language Modeling 7 Sep 2021 · 1 repository · arXiv:2109.03264
-
Does BERT Learn as Humans Perceive? Understanding Linguistic Styles through Lexica 6 Sep 2021 · 1 repository · arXiv:2109.02738Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)