Methods › General › Regularization › Weight Decay › Papers, page 67
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 67 of 108: papers 6,601 to 6,700 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Transformer Language Models without Positional Encodings Still Learn Positional Information 30 Mar 2022 · 1 repository · arXiv:2203.16634Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
A Fast Post-Training Pruning Framework for Transformers 29 Mar 2022 · 2 repositories · arXiv:2204.09656Syntology official (archive's flag): 1 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Improving Persian Relation Extraction Models by Data Augmentation 29 Mar 2022 · 0 repositories · arXiv:2203.15323
-
LinkBERT: Pretraining Language Models with Document Links 29 Mar 2022 · 1 repository · arXiv:2203.15827Syntology official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 14 harvested samples)
-
mc-BEiT: Multi-choice Discretization for Image BERT Pre-training 29 Mar 2022 · 1 repository · arXiv:2203.15371Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Training Compute-Optimal Large Language Models 29 Mar 2022 · 2 repositories · arXiv:2203.15556Syntology 8 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
ANNA: Enhanced Language Representation for Question Answering 28 Mar 2022 · 0 repositories · arXiv:2203.14507
-
Continuous Metric Learning For Transferable Speech Emotion Recognition and Embedding Across Low-resource Languages 28 Mar 2022 · 0 repositories · arXiv:2203.14867
-
Hierarchical Transformer Model for Scientific Named Entity Recognition 28 Mar 2022 · 1 repository · arXiv:2203.14710
-
UTSA NLP at SemEval-2022 Task 4: An Exploration of Simple Ensembles of Transformers, Convolutional, and Recurrent Neural Networks 28 Mar 2022 · 0 repositories · arXiv:2203.14920
-
Long-Tailed Recognition via Weight Balancing 27 Mar 2022 · 2 repositories · arXiv:2203.14197Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 14 harvested samples) · 13 pointer-only (licence)
-
Pyramid-BERT: Reducing Complexity via Successive Core-set based Token Selection 27 Mar 2022 · 0 repositories · arXiv:2203.14380
-
StruBERT: Structure-aware BERT for Table Search and Matching 27 Mar 2022 · 1 repository · arXiv:2203.14278Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples)
-
Autoregressive Linguistic Steganography Based on BERT and Consistency Coding 26 Mar 2022 · 0 repositories · arXiv:2203.13972
-
Collaborative Intelligent Reflecting Surface Networks with Multi-Agent Reinforcement Learning 26 Mar 2022 · 0 repositories · arXiv:2203.14152
-
L3Cube-MahaHate: A Tweet-based Marathi Hate Speech Detection Dataset and BERT models 25 Mar 2022 · 1 repository · arXiv:2203.13778
-
MKQ-BERT: Quantized BERT with 4-bits Weights and Activations 25 Mar 2022 · 0 repositories · arXiv:2203.13483
-
Predicting Clinical Intent from Free Text Electronic Health Records 25 Mar 2022 · 0 repositories · arXiv:2204.09594
-
Bailando: 3D Dance Generation by Actor-Critic GPT with Choreographic Memory 24 Mar 2022 · 1 repository · arXiv:2203.13055Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
mcBERT: Momentum Contrastive Learning with BERT for Zero-Shot Slot Filling 24 Mar 2022 · 0 repositories · arXiv:2203.12940
-
minicons: Enabling Flexible Behavioral and Representational Analyses of Transformer Language Models 24 Mar 2022 · 1 repository · arXiv:2203.13112
-
Mono vs Multilingual BERT: A Case Study in Hindi and Marathi Named Entity Recognition 24 Mar 2022 · 0 repositories · arXiv:2203.12907
-
Non-Parametric Stochastic Policy Gradient with Strategic Retreat for Non-Stationary Environment 24 Mar 2022 · 0 repositories · arXiv:2203.14905
-
Token Dropping for Efficient BERT Pretraining 24 Mar 2022 · 0 repositories · arXiv:2203.13240
-
Adversarial Training for Improving Model Robustness? Look at Both Prediction and Interpretation 23 Mar 2022 · 1 repository · arXiv:2203.12709Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ERNIE-SPARSE: Learning Hierarchical Efficient Transformer Through Regularized Self-Attention 23 Mar 2022 · 0 repositories · arXiv:2203.12276
-
Input-specific Attention Subnetworks for Adversarial Detection 23 Mar 2022 · 0 repositories · arXiv:2203.12298Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Computational Approach to Understand Mental Health from Reddit: Knowledge-aware Multitask Learning Framework 22 Mar 2022 · 0 repositories · arXiv:2203.11856
-
Are You Misinformed? A Study of Covid-Related Fake News in Bengali on Facebook 22 Mar 2022 · 0 repositories · arXiv:2203.11669
-
BERT-ASC: Auxiliary-Sentence Construction for Implicit Aspect Learning in Sentiment Analysis 22 Mar 2022 · 1 repository · arXiv:2203.11702Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Factual Consistency of Multilingual Pretrained Language Models 22 Mar 2022 · 1 repository · arXiv:2203.11552
-
Self-supervision through Random Segments with Autoregressive Coding (RandSAC) 22 Mar 2022 · 0 repositories · arXiv:2203.12054
-
Transformer based ensemble for emotion detection 22 Mar 2022 · 0 repositories · arXiv:2203.11899
-
Under the Hood of Transformer Networks for Trajectory Forecasting 22 Mar 2022 · 0 repositories · arXiv:2203.11878
-
A Slot Is Not Built in One Utterance: Spoken Language Dialogs with Sub-Slots 21 Mar 2022 · 1 repository · arXiv:2203.10759
-
An Intellectual Property Entity Recognition Method Based on Transformer and Technological Word Information 21 Mar 2022 · 0 repositories · arXiv:2203.10717
-
Compression of Generative Pre-trained Language Models via Quantization 21 Mar 2022 · 0 repositories · arXiv:2203.10705
-
Neural Token Segmentation for High Token-Internal Complexity 21 Mar 2022 · 0 repositories · arXiv:2203.10845
-
Semantic Similarity Computing for Scientific Academic Conferences fused with domain features 21 Mar 2022 · 0 repositories · arXiv:2203.12593
-
Towards Explainable Evaluation Metrics for Natural Language Generation 21 Mar 2022 · 1 repository · arXiv:2203.11131
-
Build a Robust QA System with Transformer-based Mixture of Experts 20 Mar 2022 · 1 repository · arXiv:2204.09598
-
Cluster & Tune: Boost Cold Start Performance in Text Classification 20 Mar 2022 · 1 repository · arXiv:2203.10581Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
g2pW: A Conditional Weighted Softmax BERT for Polyphone Disambiguation in Mandarin 20 Mar 2022 · 1 repository · arXiv:2203.10430
-
How does the pre-training objective affect what large language models learn about linguistic properties? 20 Mar 2022 · 1 repository · arXiv:2203.10415
-
MicroRacer: a didactic environment for Deep Reinforcement Learning 20 Mar 2022 · 1 repository · arXiv:2203.10494
-
Decision-making of Emergent Incident based on P-MADDPG 19 Mar 2022 · 0 repositories · arXiv:2203.12673
-
Dependency-based Mixture Language Models 19 Mar 2022 · 1 repository · arXiv:2203.10256Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Perturbations in the Wild: Leveraging Human-Written Text Perturbations for Realistic Adversarial Attack and Defense 19 Mar 2022 · 1 repository · arXiv:2203.10346Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Are You Robert or RoBERTa? Deceiving Online Authorship Attribution Models Using Neural Text Generators 18 Mar 2022 · 0 repositories · arXiv:2203.09813
-
Entropy-based Attention Regularization Frees Unintended Bias Mitigation from Lists 17 Mar 2022 · 1 repository · arXiv:2203.09192Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Fine- and Coarse-Granularity Hybrid Self-Attention for Efficient BERT 17 Mar 2022 · 1 repository · arXiv:2203.09055
-
Multilingual Detection of Personal Employment Status on Twitter 17 Mar 2022 · 1 repository · arXiv:2203.09178
-
AdapLeR: Speeding up Inference by Adaptive Length Reduction 16 Mar 2022 · 1 repository · arXiv:2203.08991
-
KinyaBERT: a Morphology-aware Kinyarwanda Language Model 16 Mar 2022 · 1 repository · arXiv:2203.08459Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples)
-
Label Semantics for Few Shot Named Entity Recognition 16 Mar 2022 · 1 repository · arXiv:2203.08985
-
Thinking about GPT-3 In-Context Learning for Biomedical IE? Think Again 16 Mar 2022 · 1 repository · arXiv:2203.08410
-
Data Contamination: From Memorization to Exploitation 15 Mar 2022 · 1 repository · arXiv:2203.08242Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs 15 Mar 2022 · 1 repository · arXiv:2203.07828Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Do Language Models Plagiarize? 15 Mar 2022 · 1 repository · arXiv:2203.07618
-
Imputing Out-of-Vocabulary Embeddings with LOVE Makes Language Models Robust with Little Cost 15 Mar 2022 · 1 repository · arXiv:2203.07860Syntology official (archive's flag): 2 ran · 6 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 2 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Learning to Infer Belief Embedded Communication 15 Mar 2022 · 0 repositories · arXiv:2203.07832
-
The Ghost in the Machine has an American accent: value conflict in GPT-3 15 Mar 2022 · 0 repositories · arXiv:2203.07785
-
Can pre-trained Transformers be used in detecting complex sensitive sentences? -- A Monsanto case study 14 Mar 2022 · 0 repositories · arXiv:2203.06793
-
Contrastive Visual Semantic Pretraining Magnifies the Semantics of Natural Language Representations 14 Mar 2022 · 0 repositories · arXiv:2203.07511
-
GrIPS: Gradient-free, Edit-based Instruction Search for Prompting Large Language Models 14 Mar 2022 · 2 repositories · arXiv:2203.07281Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
The Optimal BERT Surgeon: Scalable and Accurate Second-Order Pruning for Large Language Models 14 Mar 2022 · 1 repository · arXiv:2203.07259
-
VAST: The Valence-Assessing Semantics Test for Contextualizing Language Models 14 Mar 2022 · 1 repository · arXiv:2203.07504
-
WCL-BBCD: A Contrastive Learning and Knowledge Graph Approach to Named Entity Recognition 14 Mar 2022 · 0 repositories · arXiv:2203.06925
-
Investigating the Impact of COVID-19 on Education by Social Network Mining 13 Mar 2022 · 0 repositories · arXiv:2203.06584
-
BiBERT: Accurate Fully Binarized BERT 12 Mar 2022 · 1 repository · arXiv:2203.06390
-
ELLE: Efficient Lifelong Pre-training for Emerging Data 12 Mar 2022 · 1 repository · arXiv:2203.06311Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
FiNER: Financial Numeric Entity Recognition for XBRL Tagging 12 Mar 2022 · 2 repositories · arXiv:2203.06482
-
MarkBERT: Marking Word Boundaries Improves Chinese BERT 12 Mar 2022 · 1 repository · arXiv:2203.06378
-
A Sentence is Worth 128 Pseudo Tokens: A Semantic-Aware Contrastive Learning Framework for Sentence Embeddings 11 Mar 2022 · 1 repository · arXiv:2203.05877Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
BERTopic: Neural topic modeling with a class-based TF-IDF procedure 11 Mar 2022 · 3 repositories · arXiv:2203.05794Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Block-Sparse Adversarial Attack to Fool Transformer-Based Text Classifiers 11 Mar 2022 · 1 repository · arXiv:2203.05948
-
Hierarchical BERT for Medical Document Understanding 11 Mar 2022 · 0 repositories · arXiv:2204.09600
-
Using Word Embeddings to Analyze Protests News 11 Mar 2022 · 0 repositories · arXiv:2203.05875
-
verBERT: Automating Brazilian Case Law Document Multi-label Categorization Using BERT 11 Mar 2022 · 1 repository · arXiv:2203.06224
-
When classifying grammatical role, BERT doesn't care about word order... except when it matters 11 Mar 2022 · 1 repository · arXiv:2203.06204
-
A new approach to calculating BERTScore for automatic assessment of translation quality 10 Mar 2022 · 0 repositories · arXiv:2203.05598
-
Contextualized Sensorimotor Norms: multi-dimensional measures of sensorimotor strength for ambiguous English words, in context 10 Mar 2022 · 1 repository · arXiv:2203.05648
-
Semantic Norm Recognition and its application to Portuguese Law 10 Mar 2022 · 0 repositories · arXiv:2203.05425
-
Speciesist Language and Nonhuman Animal Bias in English Masked Language Models 10 Mar 2022 · 1 repository · arXiv:2203.05140
-
Coarse-to-Fine Sparse Transformer for Hyperspectral Image Reconstruction 9 Mar 2022 · 1 repository · arXiv:2203.04845Syntology official (archive's flag): 13 ran · 13 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
NLX-GPT: A Model for Natural Language Explanations in Vision and Vision-Language Tasks 9 Mar 2022 · 1 repository · arXiv:2203.05081
-
Designing Heterogeneous GNNs with Desired Permutation Properties for Wireless Resource Allocation 8 Mar 2022 · 0 repositories · arXiv:2203.03906
-
Towards Generalized Models for Task-oriented Dialogue Modeling on Spoken Conversations 8 Mar 2022 · 0 repositories · arXiv:2203.04045
-
Pre-trained Token-replaced Detection Model as Few-shot Learner 7 Mar 2022 · 1 repository · arXiv:2203.03235
-
Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer 7 Mar 2022 · 7 repositories · arXiv:2203.03466Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Divide and Conquer: Text Semantic Matching with Disentangled Keywords and Intents 6 Mar 2022 · 1 repository · arXiv:2203.02898
-
Graph Neural Network Enhanced Language Models for Efficient Multilingual Text Classification 6 Mar 2022 · 0 repositories · arXiv:2203.02912
-
Leveraging Pre-trained BERT for Audio Captioning 6 Mar 2022 · 0 repositories · arXiv:2203.02838
-
Detecting Offensive Language on Social Networks: An End-to-end Detection Method based on Graph Attention Networks 4 Mar 2022 · 0 repositories · arXiv:2203.02123
-
LiteTransformerSearch: Training-free Neural Architecture Search for Efficient Language Models 4 Mar 2022 · 1 repository · arXiv:2203.02094Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Training language models to follow instructions with human feedback 4 Mar 2022 · 11 repositories · arXiv:2203.02155Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
BoMD: Bag of Multi-label Descriptors for Noisy Chest X-ray Classification 3 Mar 2022 · 2 repositories · arXiv:2203.01937
-
Discontinuous Constituency and BERT: A Case Study of Dutch 2 Mar 2022 · 1 repository · arXiv:2203.01063
-
Parameter-Efficient Mixture-of-Experts Architecture for Pre-trained Language Models 2 Mar 2022 · 2 repositories · arXiv:2203.01104Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
BERT-LID: Leveraging BERT to Improve Spoken Language Identification 1 Mar 2022 · 1 repository · arXiv:2203.00328