Methods › General › Regularization › Weight Decay › Papers, page 69
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 69 of 108: papers 6,801 to 6,900 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Sentiment Analysis: Predicting Yelp Scores 20 Jan 2022 · 0 repositories · arXiv:2201.07999
-
Transfer Learning Approaches for Building Cross-Language Dense Retrieval Models 20 Jan 2022 · 1 repository · arXiv:2201.08471Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Critic Algorithms using Cooperative Networks 19 Jan 2022 · 0 repositories · arXiv:2201.07839
-
Near-Optimal Sparse Allreduce for Distributed Deep Learning 19 Jan 2022 · 1 repository · arXiv:2201.07598
-
TourBERT: A pretrained language model for the tourism industry 19 Jan 2022 · 0 repositories · arXiv:2201.07449
-
CoAuthor: Designing a Human-AI Collaborative Writing Dataset for Exploring Language Model Capabilities 18 Jan 2022 · 1 repository · arXiv:2201.06796
-
Hierarchical Neural Network Approaches for Long Document Classification 18 Jan 2022 · 0 repositories · arXiv:2201.06774
-
BERT vs ALBERT explained 17 Jan 2022 · 0 repositories
-
MuLVE, A Multi-Language Vocabulary Evaluation Data Set 17 Jan 2022 · 0 repositories · arXiv:2201.06286
-
Unintended Bias in Language Model-driven Conversational Recommendation 17 Jan 2022 · 0 repositories · arXiv:2201.06224
-
A Balanced Data Approach for Evaluating Cross-Lingual Transfer: Mapping the Linguistic Blood Bank 16 Jan 2022 · 0 repositories
-
A Multi-Granularity Opinion Summarization Method 16 Jan 2022 · 0 repositories
-
A Study of Pre-trained Language Models for Analogy Generation 16 Jan 2022 · 0 repositories
-
A Study of the Attention Abnormality in Trojaned BERTs 16 Jan 2022 · 0 repositories
-
An Exploitation of Heterogeneous Graph Neural Network for Extractive Long Document Summarization 16 Jan 2022 · 0 repositories
-
Applying SoftTriple Loss for Supervised Language Model Fine Tuning 16 Jan 2022 · 0 repositories
-
Auto-regressive Text Generation with Pre-Trained Language Models: An Empirical Study on Question-type Short Text Generation 16 Jan 2022 · 0 repositories
-
AutoAttention: Automatic Attention Head Selection Through Differentiable Pruning 16 Jan 2022 · 0 repositories
-
Bridge the Gap Between CV and NLP! A Gradient-based Textual Adversarial Attack Framework 16 Jan 2022 · 0 repositories
-
Can BERT Conduct Logical Reasoning? On the Difficulty of Learning to Reason from Data 16 Jan 2022 · 0 repositories
-
Context-Aware Prompt: Customize A Unique Prompt For Each Input 16 Jan 2022 · 0 repositories
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 16 Jan 2022 · 0 repositories
-
Divide and Conquer: Text Semantic Matching with Disentangled Keywords and Intents 16 Jan 2022 · 0 repositories
-
Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs 16 Jan 2022 · 0 repositories
-
Efficient Hierarchical Domain Adaptation for Pretrained Language Models 16 Jan 2022 · 0 repositories
-
EiCi: A New Method of Dynamic Embedding Incorporating Contextual Information in Chinese NER 16 Jan 2022 · 0 repositories
-
Elastic Weight Consolidation for Reduction of Catastrophic Forgetting in GPT-2 16 Jan 2022 · 0 repositories
-
Event Detection via Derangement Reading Comprehension 16 Jan 2022 · 0 repositories
-
Experiments with adversarial attacks on text genres 16 Jan 2022 · 0 repositories
-
Feasibility of BERT Embeddings For Domain-Specific Knowledge Mining 16 Jan 2022 · 0 repositories
-
FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks 16 Jan 2022 · 0 repositories
-
Few-Shot Semantic Parsing with Language Models Trained On Code 16 Jan 2022 · 0 repositories
-
Global Entity Disambiguation with BERT 16 Jan 2022 · 0 repositories
-
Hierarchical Transformers Are More Efficient Language Models 16 Jan 2022 · 0 repositories
-
Identifying the Source of Vulnerability in Fragile Interpretations: A Case Study in Neural Text Classification 16 Jan 2022 · 0 repositories
-
IMPLI: Investigatng NLI Models' Performance on Figurative Language 16 Jan 2022 · 0 repositories
-
Improving Contextual Representation with Gloss Regularized Pre-training 16 Jan 2022 · 0 repositories
-
Investigating and Explaining Feature and Representation Learning in Translationese Classification 16 Jan 2022 · 0 repositories
-
Investigating the saliency of sentiment expressions in aspect-based sentiment analysis 16 Jan 2022 · 0 repositories
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 16 Jan 2022 · 0 repositories
-
Language Models for Code-switch Detection of te reo Māori and English in a Low-resource Setting 16 Jan 2022 · 0 repositories
-
Magic Pyramid: Accelerating Inference with Early Exiting and Token Pruning 16 Jan 2022 · 0 repositories
-
Measuring Word-Context Biases in Lexical Semantic Datasets 16 Jan 2022 · 0 repositories
-
Memory-assisted prompt editing to improve GPT-3 after deployment 16 Jan 2022 · 1 repository · arXiv:2201.06009Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Minimally-Supervised Relation Induction from Pre-trained Language Model 16 Jan 2022 · 0 repositories
-
Modeling Multi-Granularity Hierarchical Features for Relation Extraction 16 Jan 2022 · 0 repositories
-
Multi-Stage Pre-Training for Math-Understanding: μ²(AL)BERT 16 Jan 2022 · 0 repositories
-
PCEE-BERT: Accelerating BERT Inference via Patient and Confident Early Exiting 16 Jan 2022 · 0 repositories
-
Penguins Don’t Fly: Reasoning about Generics through Instantiations and Exceptions 16 Jan 2022 · 0 repositories
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 16 Jan 2022 · 0 repositories
-
Probing The Linguistic Capacity of Pre-Trained Vision-Language Models 16 Jan 2022 · 0 repositories
-
Progressive Class Semantic Matching for Semi-supervised Text Classification 16 Jan 2022 · 0 repositories
-
Provably Confidential Language Modelling 16 Jan 2022 · 0 repositories
-
Re2G: Retrieve, Rerank, Generate 16 Jan 2022 · 0 repositories
-
Reframing Human-AI Collaboration for Generating Free-Text Explanations 16 Jan 2022 · 0 repositories
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 16 Jan 2022 · 0 repositories
-
Revisiting Additive Compositionality: AND, OR, and NOT Operations with Word Embeddings 16 Jan 2022 · 0 repositories
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 16 Jan 2022 · 0 repositories
-
Roof-BERT: Divide Understanding Labour and Join in Work 16 Jan 2022 · 0 repositories
-
SemAttack: Natural Textual Attacks via Different Semantic Spaces 16 Jan 2022 · 0 repositories
-
Seq-GAN-BERT:Sequence Generative Adversarial Learning for Low-resource Name Entity Recognition 16 Jan 2022 · 0 repositories
-
Simple Local Attentions Remain Competitive for Long-Context Tasks 16 Jan 2022 · 0 repositories
-
TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning 16 Jan 2022 · 0 repositories
-
Tapping BERT for Preposition Sense Disambiguation 16 Jan 2022 · 0 repositories
-
TEMPLATE: TempRel Classification Model Trained with Embedded Temporal Relation Knowledge 16 Jan 2022 · 0 repositories
-
That is a good looking car !: Visual Aspect based Sentiment Controlled Personalized Response Generation 16 Jan 2022 · 0 repositories
-
Tree Knowledge Distillation for Compressing Transformer-Based Language Models 16 Jan 2022 · 0 repositories
-
Uncovering Surprising Event Boundaries in Narratives 16 Jan 2022 · 0 repositories
-
Understand before Answer: Improve Temporal Reading Comprehension via Precise Question Understanding 16 Jan 2022 · 0 repositories
-
UnifiedSKG: Unifying and Multi-Tasking Structured Knowledge Grounding with Text-to-Text Language Models 16 Jan 2022 · 1 repository · arXiv:2201.05966Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
VEE-BERT: Accelerating BERT Inference for Named Entity Recognition via Vote Early Exiting 16 Jan 2022 · 0 repositories
-
WANLI: Worker and AI Collaboration for Natural Language Inference Dataset Creation 16 Jan 2022 · 1 repository · arXiv:2201.05955
-
What do tokens know about their characters and how do they know it? 16 Jan 2022 · 1 repository
-
What Role Does BERT Play in the Neural Machine Translation Encoder? 16 Jan 2022 · 0 repositories
-
When a sentence does not introduce a discourse entity, Transformer-based models still often refer to it 16 Jan 2022 · 0 repositories
-
Why Does Surprisal From Smaller GPT-2 Models Provide Better Fit to Human Reading Times? 16 Jan 2022 · 0 repositories
-
Automatic Correction of Syntactic Dependency Annotation Differences 15 Jan 2022 · 0 repositories · arXiv:2201.05891
-
Automatic Lexical Simplification for Turkish 15 Jan 2022 · 0 repositories · arXiv:2201.05878
-
Machine Learning for Food Review and Recommendation 15 Jan 2022 · 0 repositories · arXiv:2201.10978
-
CommonsenseQA 2.0: Exposing the Limits of AI through Gamification 14 Jan 2022 · 0 repositories · arXiv:2201.05320
-
Polarity and Subjectivity Detection with Multitask Learning and BERT Embedding 14 Jan 2022 · 0 repositories · arXiv:2201.05363
-
Assemble Foundation Models for Automatic Code Summarization 13 Jan 2022 · 1 repository · arXiv:2201.05222
-
Knowledge Graph Augmented Network Towards Multiview Representation Learning for Aspect-based Sentiment Analysis 13 Jan 2022 · 1 repository · arXiv:2201.04831
-
Multi-task Pre-training Language Model for Semantic Network Completion 13 Jan 2022 · 1 repository · arXiv:2201.04843
-
Towards Automated Error Analysis: Learning to Characterize Errors 13 Jan 2022 · 0 repositories · arXiv:2201.05017
-
Diagnosing BERT with Retrieval Heuristics 12 Jan 2022 · 1 repository · arXiv:2201.04458
-
Generative Adversarial Network for Text-to-Face Synthesis and Manipulation with Pretrained BERT Model 12 Jan 2022 · 0 repositories
-
PromptBERT: Improving BERT Sentence Embeddings with Prompts 12 Jan 2022 · 1 repository · arXiv:2201.04337Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; the one sample that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
A Feature Extraction based Model for Hate Speech Identification 11 Jan 2022 · 0 repositories · arXiv:2201.04227
-
Explaining Predictive Uncertainty by Looking Back at Model Explanations 11 Jan 2022 · 0 repositories · arXiv:2201.03742
-
Quantifying Robustness to Adversarial Word Substitutions 11 Jan 2022 · 0 repositories · arXiv:2201.03829
-
BERT for Sentiment Analysis: Pre-trained and Fine-Tuned Alternatives 10 Jan 2022 · 2 repositories · arXiv:2201.03382
-
Black-Box Tuning for Language-Model-as-a-Service 10 Jan 2022 · 2 repositories · arXiv:2201.03514Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Handwriting recognition and automatic scoring for descriptive answers in Japanese language tests 10 Jan 2022 · 0 repositories · arXiv:2201.03215
-
SCROLLS: Standardized CompaRison Over Long Language Sequences 10 Jan 2022 · 2 repositories · arXiv:2201.03533Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Latency Adjustable Transformer Encoder for Language Understanding 10 Jan 2022 · 0 repositories · arXiv:2201.03327
-
Imagined versus Remembered Stories: Quantifying Differences in Narrative Flow 7 Jan 2022 · 0 repositories · arXiv:2201.02662
-
Flow-Guided Sparse Transformer for Video Deblurring 6 Jan 2022 · 1 repository · arXiv:2201.01893Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Self-Training Vision Language BERTs with a Unified Conditional Model 6 Jan 2022 · 0 repositories · arXiv:2201.02010
-
Formal Analysis of Art: Proxy Learning of Visual Concepts from Style Through Language Models 5 Jan 2022 · 0 repositories · arXiv:2201.01819