Methods › General › Regularization › Weight Decay › Papers, page 92
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 92 of 108: papers 9,101 to 9,200 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Techniques to Improve Q&A Accuracy with Transformer-based models on Large Complex Documents 26 Sep 2020 · 0 repositories · arXiv:2009.12695
-
A little goes a long way: Improving toxic language classification despite data scarcity 25 Sep 2020 · 1 repository · arXiv:2009.12344
-
An Unsupervised Sentence Embedding Method by Mutual Information Maximization 25 Sep 2020 · 1 repository · arXiv:2009.12061
-
BET: A Backtranslation Approach for Easy Data Augmentation in Transformer-based Paraphrase Identification Context 25 Sep 2020 · 1 repository · arXiv:2009.12452
-
HetSeq: Distributed GPU Training on Heterogeneous Infrastructure 25 Sep 2020 · 1 repository · arXiv:2009.14783
-
A Comparative Study of Feature Types for Age-Based Text Classification 24 Sep 2020 · 1 repository · arXiv:2009.11898
-
Adapting BERT for Word Sense Disambiguation with Gloss Selection Objective and Example Sentences 24 Sep 2020 · 1 repository · arXiv:2009.11795Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AnchiBERT: A Pre-Trained Model for Ancient ChineseLanguage Understanding and Generation 24 Sep 2020 · 1 repository · arXiv:2009.11473
-
Toward a Thermodynamics of Meaning 24 Sep 2020 · 1 repository · arXiv:2009.11963
-
A Token-wise CNN-based Method for Sentence Compression 23 Sep 2020 · 0 repositories · arXiv:2009.11260
-
Pruning Convolutional Filters using Batch Bridgeout 23 Sep 2020 · 0 repositories · arXiv:2009.10893
-
AutoRC: Improving BERT Based Relation Classification Models via Architecture Search 22 Sep 2020 · 0 repositories · arXiv:2009.10680
-
Constructing interval variables via faceted Rasch measurement and multitask deep learning: a hate speech application 22 Sep 2020 · 2 repositories · arXiv:2009.10277Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
GRACE: Gradient Harmonized and Cascaded Labeling for Aspect-based Sentiment Analysis 22 Sep 2020 · 1 repository · arXiv:2009.10557
-
On Data Augmentation for Extreme Multi-label Classification 22 Sep 2020 · 0 repositories · arXiv:2009.10778
-
Latin BERT: A Contextual Language Model for Classical Philology 21 Sep 2020 · 1 repository · arXiv:2009.10053
-
Open-set Short Utterance Forensic Speaker Verification using Teacher-Student Network with Explicit Inductive Bias 21 Sep 2020 · 0 repositories · arXiv:2009.09556
-
Profile Consistency Identification for Open-domain Dialogue Agents 21 Sep 2020 · 1 repository · arXiv:2009.09680Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples)
-
"Listen, Understand and Translate": Triple Supervision Decouples End-to-end Speech-to-text Translation 21 Sep 2020 · 1 repository · arXiv:2009.09704
-
"When they say weed causes depression, but it's your fav antidepressant": Knowledge-aware Attention Framework for Relationship Extraction 21 Sep 2020 · 0 repositories · arXiv:2009.10155
-
Dual-path CNN with Max Gated block for Text-Based Person Re-identification 20 Sep 2020 · 1 repository · arXiv:2009.09343
-
Longformer for MS MARCO Document Re-ranking Task 20 Sep 2020 · 1 repository · arXiv:2009.09392
-
Persian Ezafe Recognition Using Transformers and Its Role in Part-Of-Speech Tagging 20 Sep 2020 · 1 repository · arXiv:2009.09474
-
Vicomtech at eHealth-KD Challenge 2020: Deep End-to-End Model for Entity and Relation Extraction in Medical Text 20 Sep 2020 · 0 repositories
-
VirtualFlow: Decoupling Deep Learning Models from the Underlying Hardware 20 Sep 2020 · 0 repositories · arXiv:2009.09523
-
Conditionally Adaptive Multi-Task Learning: Improving Transfer Learning in NLP Using Fewer Parameters & Less Data 19 Sep 2020 · 1 repository · arXiv:2009.09139
-
Nominal Compound Chain Extraction: A New Task for Semantic-enriched Lexical Chain 19 Sep 2020 · 0 repositories · arXiv:2009.09173
-
Prior Art Search and Reranking for Generated Patent Text 19 Sep 2020 · 0 repositories · arXiv:2009.09132
-
fastHan: A BERT-based Multi-Task Toolkit for Chinese NLP 18 Sep 2020 · 1 repository · arXiv:2009.08633
-
Hierarchical GPT with Congruent Transformers for Multi-Sentence Language Models 18 Sep 2020 · 0 repositories · arXiv:2009.08636
-
NEU at WNUT-2020 Task 2: Data Augmentation To Tell BERT That Death Is Not Necessarily Informative 18 Sep 2020 · 0 repositories · arXiv:2009.08590
-
The birth of Romanian BERT 18 Sep 2020 · 1 repository · arXiv:2009.08712
-
Will it Unblend? 18 Sep 2020 · 1 repository · arXiv:2009.09123
-
A Multimodal Memes Classification: A Survey and Open Research Issues 17 Sep 2020 · 0 repositories · arXiv:2009.08395
-
Compositional and Lexical Semantics in RoBERTa, BERT and DistilBERT: A Case Study on CoQA 17 Sep 2020 · 0 repositories · arXiv:2009.08257
-
Cross-Modal Alignment with Mixture Experts Neural Network for Intral-City Retail Recommendation 17 Sep 2020 · 0 repositories · arXiv:2009.09926
-
DSC IIT-ISM at SemEval-2020 Task 6: Boosting BERT with Dependencies for Definition Extraction 17 Sep 2020 · 1 repository · arXiv:2009.08180
-
Efficient Transformer-based Large Scale Language Representations using Hardware-friendly Block Structured Pruning 17 Sep 2020 · 0 repositories · arXiv:2009.08065
-
Knowledge-Assisted Deep Reinforcement Learning in 5G Scheduler Design: From Theoretical Framework to Implementation 17 Sep 2020 · 0 repositories · arXiv:2009.08346
-
MEAL V2: Boosting Vanilla ResNet-50 to 80%+ Top-1 Accuracy on ImageNet without Tricks 17 Sep 2020 · 1 repository · arXiv:2009.08453
-
Multi²OIE: Multilingual Open Information Extraction Based on Multi-Head Attention with BERT 17 Sep 2020 · 1 repository · arXiv:2009.08128Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Deep Learning Approaches for Extracting Adverse Events and Indications of Dietary Supplements from Clinical Text 16 Sep 2020 · 0 repositories · arXiv:2009.07780
-
Simplified TinyBERT: Knowledge Distillation for Document Retrieval 16 Sep 2020 · 4 repositories · arXiv:2009.07531
-
Solomon at SemEval-2020 Task 11: Ensemble Architecture for Fine-Tuned Propaganda Detection in News Articles 16 Sep 2020 · 0 repositories · arXiv:2009.07473
-
UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation 16 Sep 2020 · 1 repository · arXiv:2009.07602Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Real-Time Execution of Large-scale Language Models on Mobile 15 Sep 2020 · 0 repositories · arXiv:2009.06823
-
Augmented Natural Language for Generative Sequence Labeling 15 Sep 2020 · 0 repositories · arXiv:2009.13272
-
BERT-QE: Contextualized Query Expansion for Document Re-ranking 15 Sep 2020 · 1 repository · arXiv:2009.07258Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
Critical Thinking for Language Models 15 Sep 2020 · 1 repository · arXiv:2009.07185Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
DeNERT-KG: Named Entity and Relation Extraction Model Using DQN, Knowledge Graph, and BERT 15 Sep 2020 · 0 repositories
-
Dialogue Response Ranking Training with Large-Scale Human Feedback Data 15 Sep 2020 · 2 repositories · arXiv:2009.06978Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Event Presence Prediction Helps Trigger Detection Across Languages 15 Sep 2020 · 0 repositories · arXiv:2009.07188
-
It's Not Just Size That Matters: Small Language Models Are Also Few-Shot Learners 15 Sep 2020 · 5 repositories · arXiv:2009.07118
-
Lessons Learned from Applying off-the-shelf BERT: There is no Silver Bullet 15 Sep 2020 · 0 repositories · arXiv:2009.07238
-
MLMLM: Link Prediction with Mean Likelihood Masked Language Model 15 Sep 2020 · 0 repositories · arXiv:2009.07058
-
The Radicalization Risks of GPT-3 and Advanced Neural Language Models 15 Sep 2020 · 0 repositories · arXiv:2009.06807
-
Beyond Accuracy: ROI-driven Data Analytics of Empirical Data 14 Sep 2020 · 0 repositories · arXiv:2009.06492
-
On Robustness and Bias Analysis of BERT-based Relation Extraction 14 Sep 2020 · 1 repository · arXiv:2009.06206
-
Efficient Transformers: A Survey 14 Sep 2020 · 0 repositories · arXiv:2009.06732
-
Filling the Gap of Utterance-aware and Speaker-aware Representation for Multi-turn Dialogue 14 Sep 2020 · 1 repository · arXiv:2009.06504Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
GeDi: Generative Discriminator Guided Sequence Generation 14 Sep 2020 · 3 repositories · arXiv:2009.06367Syntology official (archive's flag): 4 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
BoostingBERT:Integrating Multi-Class Boosting into BERT for NLP Tasks 13 Sep 2020 · 0 repositories · arXiv:2009.05959
-
Cluster-Former: Clustering-based Sparse Transformer for Long-Range Dependency Encoding 13 Sep 2020 · 0 repositories · arXiv:2009.06097
-
CIA_NITT at WNUT-2020 Task 2: Classification of COVID-19 Tweets Using Pre-trained Language Models 12 Sep 2020 · 0 repositories · arXiv:2009.05782
-
Country Image in COVID-19 Pandemic: A Case Study of China 12 Sep 2020 · 1 repository · arXiv:2009.05817
-
Fine-tuning Pre-trained Contextual Embeddings for Citation Content Analysis in Scholarly Publication 12 Sep 2020 · 0 repositories · arXiv:2009.05836
-
A Comparison of LSTM and BERT for Small Corpus 11 Sep 2020 · 0 repositories · arXiv:2009.05451
-
Compressed Deep Networks: Goodbye SVD, Hello Robust Low-Rank Approximation 11 Sep 2020 · 1 repository · arXiv:2009.05647
-
Unit Test Case Generation with Transformers and Focal Context 11 Sep 2020 · 1 repository · arXiv:2009.05617
-
UPB at SemEval-2020 Task 11: Propaganda Detection with Domain-Specific Trained BERT 11 Sep 2020 · 0 repositories · arXiv:2009.05289
-
UPB at SemEval-2020 Task 6: Pretrained Language Models for Definition Extraction 11 Sep 2020 · 3 repositories · arXiv:2009.05603
-
Brain2Word: Decoding Brain Activity for Language Generation 10 Sep 2020 · 1 repository · arXiv:2009.04765
-
Do Response Selection Models Really Know What's Next? Utterance Manipulation Strategies for Multi-turn Response Selection 10 Sep 2020 · 1 repository · arXiv:2009.04703
-
Investigating Gender Bias in BERT 10 Sep 2020 · 0 repositories · arXiv:2009.05021
-
Modern Methods for Text Generation 10 Sep 2020 · 2 repositories · arXiv:2009.04968
-
Sparsifying Transformer Models with Trainable Representation Pooling 10 Sep 2020 · 1 repository · arXiv:2009.05169
-
Comparative Study of Language Models on Cross-Domain Data with Model Agnostic Explainability 9 Sep 2020 · 0 repositories · arXiv:2009.04095
-
Pay Attention when Required 9 Sep 2020 · 2 repositories · arXiv:2009.04534
-
ERNIE at SemEval-2020 Task 10: Learning Word Emphasis Selection by Pre-trained Language Model 8 Sep 2020 · 0 repositories · arXiv:2009.03706
-
Black Box to White Box: Discover Model Characteristics Based on Strategic Probing 7 Sep 2020 · 0 repositories · arXiv:2009.03136
-
E-BERT: A Phrase and Product Knowledge Enhanced Language Model for E-commerce 7 Sep 2020 · 0 repositories · arXiv:2009.02835
-
Improving Language Generation with Sentence Coherence Objective 7 Sep 2020 · 1 repository · arXiv:2009.06358
-
Measuring Massive Multitask Language Understanding 7 Sep 2020 · 18 repositories · arXiv:2009.03300Syntology community repositories only · 19 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 26 harvested samples) · 1 pointer-only (licence)
-
EdinburghNLP at WNUT-2020 Task 2: Leveraging Transformers with Generalized Augmentation for Identifying Informativeness in COVID-19 Tweets 6 Sep 2020 · 0 repositories · arXiv:2009.06375
-
QiaoNing at SemEval-2020 Task 4: Commonsense Validation and Explanation system based on ensemble of language model 6 Sep 2020 · 0 repositories · arXiv:2009.02645
-
Accenture at CheckThat! 2020: If you say so: Post-hoc fact-checking of claims using transformer-based models 5 Sep 2020 · 0 repositories · arXiv:2009.02431
-
Comparative Evaluation of Pretrained Transfer Learning Models on Automatic Short Answer Grading 2 Sep 2020 · 1 repository · arXiv:2009.01303
-
Automatic Assignment of Radiology Examination Protocols Using Pre-trained Language Models with Knowledge Distillation 1 Sep 2020 · 1 repository · arXiv:2009.00694
-
Sentimental LIAR: Extended Corpus and Deep Learning Models for Fake Claim Classification 1 Sep 2020 · 2 repositories · arXiv:2009.01047
-
A Bidirectional Tree Tagging Scheme for Joint Medical Relation Extraction 31 Aug 2020 · 0 repositories · arXiv:2008.13339
-
SocCogCom at SemEval-2020 Task 11: Characterizing and Detecting Propaganda using Sentence-Level Emotional Salience Features 29 Aug 2020 · 1 repository · arXiv:2008.13012
-
Knowledge Efficient Deep Learning for Natural Language Processing 28 Aug 2020 · 0 repositories · arXiv:2008.12878
-
Rethinking the Objectives of Extractive Question Answering 28 Aug 2020 · 1 repository · arXiv:2008.12804Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Fast and Robust BERT-based Dialogue State Tracker for Schema-Guided Dialogue Dataset 27 Aug 2020 · 1 repository · arXiv:2008.12335
-
AMBERT: A Pre-trained Language Model with Multi-Grained Tokenization 27 Aug 2020 · 0 repositories · arXiv:2008.11869
-
DAVE: Deriving Automatically Verilog from English 27 Aug 2020 · 0 repositories · arXiv:2009.01026
-
Entity and Evidence Guided Relation Extraction for DocRED 27 Aug 2020 · 0 repositories · arXiv:2008.12283
-
GREEK-BERT: The Greeks visiting Sesame Street 27 Aug 2020 · 1 repository · arXiv:2008.12014Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
MultiGBS: A multi-layer graph approach to biomedical summarization 27 Aug 2020 · 0 repositories · arXiv:2008.11908
-
Query Focused Multi-document Summarisation of Biomedical Texts 27 Aug 2020 · 1 repository · arXiv:2008.11986