Methods › General › Regularization › Weight Decay › Papers, page 83
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 83 of 108: papers 8,201 to 8,300 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Detecting Scenes in Fiction: A new Segmentation Task 1 Apr 2021 · 0 repositories
-
ENPAR:Enhancing Entity and Entity Pair Representations for Joint Entity Relation Extraction 1 Apr 2021 · 1 repository
-
Evaluating language models for the retrieval and categorization of lexical collocations 1 Apr 2021 · 1 repository
-
Evaluating Neural Model Robustness for Machine Comprehension 1 Apr 2021 · 0 repositories
-
HLE-UPC at SemEval-2021 Task 5: Multi-Depth DistilBERT for Toxic Spans Detection 1 Apr 2021 · 1 repository · arXiv:2104.00639
-
How Fast can BERT Learn Simple Natural Language Inference? 1 Apr 2021 · 0 repositories
-
Keep Learning: Self-supervised Meta-learning for Learning from Inference 1 Apr 2021 · 0 repositories
-
Maximal Multiverse Learning for Promoting Cross-Task Generalization of Fine-Tuned Language Models 1 Apr 2021 · 0 repositories
-
Multilingual Entity and Relation Extraction Dataset and Model 1 Apr 2021 · 1 repository
-
Neural-Driven Search-Based Paraphrase Generation 1 Apr 2021 · 0 repositories
-
NLQuAD: A Non-Factoid Long Question Answering Data Set 1 Apr 2021 · 1 repository
-
On the (In)Effectiveness of Images for Text Classification 1 Apr 2021 · 0 repositories
-
Probing for idiomaticity in vector space models 1 Apr 2021 · 1 repository
-
Retrieval, Re-ranking and Multi-task Learning for Knowledge-Base Question Answering 1 Apr 2021 · 0 repositories
-
Russian Paraphrasers: Paraphrase with Transformers 1 Apr 2021 · 2 repositories
-
Spectral decoupling allows training transferable neural networks in medical imaging 31 Mar 2021 · 0 repositories · arXiv:2103.17171
-
An In-depth Analysis of Passage-Level Label Transfer for Contextual Document Ranking 30 Mar 2021 · 1 repository · arXiv:2103.16669
-
Automatic Graph Partitioning for Very Large-scale Deep Learning 30 Mar 2021 · 0 repositories · arXiv:2103.16063
-
Grounding Dialogue Systems via Knowledge Graph Aware Decoding with Pre-trained Transformers 30 Mar 2021 · 1 repository · arXiv:2103.16289
-
Kaleido-BERT: Vision-Language Pre-training on Fashion Domain 30 Mar 2021 · 1 repository · arXiv:2103.16110Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Multi-Scale Vision Longformer: A New Vision Transformer for High-Resolution Image Encoding 29 Mar 2021 · 3 repositories · arXiv:2103.15358Syntology official (archive's flag): 3 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Contextual Text Embeddings for Twi 29 Mar 2021 · 0 repositories · arXiv:2103.15963
-
FixNorm: Dissecting Weight Decay for Training Deep Neural Networks 29 Mar 2021 · 0 repositories · arXiv:2103.15345
-
FocusedDropout for Convolutional Neural Network 29 Mar 2021 · 0 repositories · arXiv:2103.15425
-
Retraining DistilBERT for a Voice Shopping Assistant by Using Universal Dependencies 29 Mar 2021 · 0 repositories · arXiv:2103.15737
-
Whitening Sentence Representations for Better Semantics and Faster Retrieval 29 Mar 2021 · 3 repositories · arXiv:2103.15316Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
PnG BERT: Augmented BERT on Phonemes and Graphemes for Neural TTS 28 Mar 2021 · 0 repositories · arXiv:2103.15060
-
Machine Learning Meets Natural Language Processing -- The story so far 27 Mar 2021 · 0 repositories · arXiv:2104.10213
-
Self-adaptive Torque Vectoring Controller Using Reinforcement Learning 27 Mar 2021 · 1 repository · arXiv:2103.14892
-
Unsupervised Self-Training for Sentiment Analysis of Code-Switched Data 27 Mar 2021 · 0 repositories · arXiv:2103.14797
-
A Practical Survey on Faster and Lighter Transformers 26 Mar 2021 · 0 repositories · arXiv:2103.14636
-
BERT4SO: Neural Sentence Ordering by Fine-tuning BERT 25 Mar 2021 · 0 repositories · arXiv:2103.13584
-
Bertinho: Galician BERT Representations 25 Mar 2021 · 0 repositories · arXiv:2103.13799
-
K-XLNet: A General Method for Combining Explicit Knowledge with Language Model Pretraining 25 Mar 2021 · 0 repositories · arXiv:2104.10649
-
Predicting Directionality in Causal Relations in Text 25 Mar 2021 · 2 repositories · arXiv:2103.13606
-
Visual Grounding Strategies for Text-Only Natural Language Processing 25 Mar 2021 · 0 repositories · arXiv:2103.13942
-
Czert -- Czech BERT-like Model for Language Representation 24 Mar 2021 · 1 repository · arXiv:2103.13031
-
Thinking Aloud: Dynamic Context Generation Improves Zero-Shot Reasoning Performance of GPT-2 24 Mar 2021 · 0 repositories · arXiv:2103.13033
-
Are Neural Language Models Good Plagiarists? A Benchmark for Neural Paraphrase Detection 23 Mar 2021 · 0 repositories · arXiv:2103.12450
-
Detecting Hate Speech with GPT-3 23 Mar 2021 · 2 repositories · arXiv:2103.12407
-
Repairing Pronouns in Translation with BERT-Based Post-Editing 23 Mar 2021 · 0 repositories · arXiv:2103.12838
-
The NLP Cookbook: Modern Recipes for Transformer based Deep Learning Architectures 23 Mar 2021 · 0 repositories · arXiv:2104.10640
-
TMR: Evaluating NER Recall on Tough Mentions 23 Mar 2021 · 0 repositories · arXiv:2103.12312
-
Variable Name Recovery in Decompiled Binary Code using Constrained Masked Language Modeling 23 Mar 2021 · 0 repositories · arXiv:2103.12801
-
BERT: A Review of Applications in Natural Language Processing and Understanding 22 Mar 2021 · 0 repositories · arXiv:2103.11943
-
Bridging the gap between supervised classification and unsupervised topic modelling for social-media assisted crisis management 22 Mar 2021 · 0 repositories · arXiv:2103.11835
-
PatentSBERTa: A Deep NLP based Hybrid Model for Patent Distance and Classification using Augmented SBERT 22 Mar 2021 · 2 repositories · arXiv:2103.11933
-
Identifying Machine-Paraphrased Plagiarism 22 Mar 2021 · 2 repositories · arXiv:2103.11909
-
Open Domain Question Answering over Tables via Dense Retrieval 22 Mar 2021 · 1 repository · arXiv:2103.12011
-
NameRec*: Highly Accurate and Fine-grained Person Name Recognition 21 Mar 2021 · 0 repositories · arXiv:2103.11360
-
ROSITA: Refined BERT cOmpreSsion with InTegrAted techniques 21 Mar 2021 · 1 repository · arXiv:2103.11367Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Cost-effective Deployment of BERT Models in Serverless Environment 19 Mar 2021 · 0 repositories · arXiv:2103.10673
-
Let Your Heart Speak in its Mother Tongue: Multilingual Captioning of Cardiac Signals 19 Mar 2021 · 1 repository · arXiv:2103.11011
-
MuRIL: Multilingual Representations for Indian Languages 19 Mar 2021 · 1 repository · arXiv:2103.10730
-
Play the Shannon Game With Language Models: A Human-Free Approach to Summary Evaluation 19 Mar 2021 · 0 repositories · arXiv:2103.10918
-
GLM: General Language Model Pretraining with Autoregressive Blank Infilling 18 Mar 2021 · 8 repositories · arXiv:2103.10360Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Contextual Biasing of Language Models for Speech Recognition in Goal-Oriented Conversational Agents 18 Mar 2021 · 0 repositories · arXiv:2103.10325
-
GPT Understands, Too 18 Mar 2021 · 10 repositories · arXiv:2103.10385Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Model Extraction and Adversarial Transferability, Your BERT is Vulnerable! 18 Mar 2021 · 1 repository · arXiv:2103.10013
-
Code Word Detection in Fraud Investigations using a Deep-Learning Approach 17 Mar 2021 · 0 repositories · arXiv:2103.09606
-
On the Role of Images for Analyzing Claims in Social Media 17 Mar 2021 · 1 repository · arXiv:2103.09602
-
UniParma at SemEval-2021 Task 5: Toxic Spans Detection Using CharacterBERT and Bag-of-Words Model 17 Mar 2021 · 1 repository · arXiv:2103.09645
-
KGSynNet: A Novel Entity Synonyms Discovery Framework with Knowledge Graph 16 Mar 2021 · 0 repositories · arXiv:2103.08893
-
Robustly Optimized and Distilled Training for Natural Language Understanding 16 Mar 2021 · 0 repositories · arXiv:2103.08809
-
Simulation Studies on Deep Reinforcement Learning for Building Control with Human Interaction 14 Mar 2021 · 0 repositories · arXiv:2103.07919
-
Revisiting ResNets: Improved Training and Scaling Strategies 13 Mar 2021 · 3 repositories · arXiv:2103.07579
-
Text Mining of Stocktwits Data for Predicting Stock Prices 13 Mar 2021 · 0 repositories · arXiv:2103.16388
-
Comparing the Performance of NLP Toolkits and Evaluation measures in Legal Tech 12 Mar 2021 · 0 repositories · arXiv:2103.11792
-
Explaining and Improving BERT Performance on Lexical Semantic Change Detection 12 Mar 2021 · 0 repositories · arXiv:2103.07259
-
Is BERT a Cross-Disciplinary Knowledge Learner? A Surprising Finding of Pre-trained Models' Transferability 12 Mar 2021 · 0 repositories · arXiv:2103.07162
-
Composite Re-Ranking for Efficient Document Search with BERT 11 Mar 2021 · 0 repositories · arXiv:2103.06499
-
Evaluation of Morphological Embeddings for the Russian Language 11 Mar 2021 · 0 repositories · arXiv:2103.06628
-
FairFil: Contrastive Neural Debiasing Method for Pretrained Text Encoders 11 Mar 2021 · 0 repositories · arXiv:2103.06413
-
Improving Bi-encoder Document Ranking Models with Two Rankers and Multi-teacher Distillation 11 Mar 2021 · 1 repository · arXiv:2103.06523
-
LightMBERT: A Simple Yet Effective Method for Multilingual BERT Distillation 11 Mar 2021 · 0 repositories · arXiv:2103.06418
-
Self-supervised Text-to-SQL Learning with Header Alignment Training 11 Mar 2021 · 0 repositories · arXiv:2103.06402
-
Towards Multi-Sense Cross-Lingual Alignment of Contextual Embeddings 11 Mar 2021 · 1 repository · arXiv:2103.06459
-
CEQE: Contextualized Embeddings for Query Expansion 9 Mar 2021 · 0 repositories · arXiv:2103.05256
-
Large Pre-trained Language Models Contain Human-like Biases of What is Right and Wrong to Do 8 Mar 2021 · 1 repository · arXiv:2103.11790Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
SKILLBERT: “SKILLING” THE BERT TO CLASSIFY SKILLS! 8 Mar 2021 · 0 repositories
-
Syntax-BERT: Improving Pre-trained Transformers with Syntax Trees 7 Mar 2021 · 1 repository · arXiv:2103.04350
-
Fine-tuning Pretrained Multilingual BERT Model for Indonesian Aspect-based Sentiment Analysis 5 Mar 2021 · 0 repositories · arXiv:2103.03732
-
MalBERT: Using Transformers for Cybersecurity and Malicious Software Detection 5 Mar 2021 · 0 repositories · arXiv:2103.03806
-
Non-invasive Self-attention for Side Information Fusion in Sequential Recommendation 5 Mar 2021 · 0 repositories · arXiv:2103.03578
-
Hardware Acceleration of Fully Quantized BERT for Efficient Natural Language Processing 4 Mar 2021 · 0 repositories · arXiv:2103.02800
-
Few-shot Learning for Slot Tagging with Attentive Relational Network 3 Mar 2021 · 0 repositories · arXiv:2103.02333
-
Natural Language Understanding for Argumentative Dialogue Systems in the Opinion Building Domain 3 Mar 2021 · 0 repositories · arXiv:2103.02691
-
Data-driven control of room temperature and bidirectional EV charging using deep reinforcement learning: simulations and experiments 2 Mar 2021 · 0 repositories · arXiv:2103.01886
-
Disentangling Syntax and Semantics in the Brain with Deep Networks 2 Mar 2021 · 0 repositories · arXiv:2103.01620
-
Hate Towards the Political Opponent: A Twitter Corpus Study of the 2020 US Elections on the Basis of Offensive Speech and Stance Detection 2 Mar 2021 · 0 repositories · arXiv:2103.01664
-
BERT-based knowledge extraction method of unstructured domain text 1 Mar 2021 · 0 repositories · arXiv:2103.00728
-
BERT based patent novelty search by training claims to their own description 1 Mar 2021 · 0 repositories · arXiv:2103.01126
-
Combat COVID-19 Infodemic Using Explainable Natural Language Processing Models 1 Mar 2021 · 0 repositories · arXiv:2103.00747
-
Long Document Summarization in a Low Resource Setting using Pretrained Language Models 1 Mar 2021 · 0 repositories · arXiv:2103.00751
-
NLP-CUET@DravidianLangTech-EACL2021: Offensive Language Detection from Multilingual Code-Mixed Text using Transformers 28 Feb 2021 · 1 repository · arXiv:2103.00455
-
NLP-CUET@LT-EDI-EACL2021: Multilingual Code-Mixed Hope Speech Detection using Cross-lingual Representation Learner 28 Feb 2021 · 1 repository · arXiv:2103.00464
-
COVID-19 Tweets Analysis through Transformer Language Models 27 Feb 2021 · 1 repository · arXiv:2103.00199
-
Transformers with Competitive Ensembles of Independent Mechanisms 27 Feb 2021 · 0 repositories · arXiv:2103.00336
-
Multi-Agent Path Planning based on MPC and DDPG 26 Feb 2021 · 0 repositories · arXiv:2102.13283
-
Multi-task transfer learning for finding actionable information from crisis-related messages on social media 26 Feb 2021 · 0 repositories · arXiv:2102.13395