Methods › General › Regularization › Weight Decay › Papers, page 68
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 68 of 108: papers 6,701 to 6,800 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
E-LANG: Energy-Based Joint Inferencing of Super and Swift Language Models 1 Mar 2022 · 0 repositories · arXiv:2203.00748
-
Exploring and Adapting Chinese GPT to Pinyin Input Method 1 Mar 2022 · 1 repository · arXiv:2203.00249
-
"Is Whole Word Masking Always Better for Chinese BERT?": Probing on Chinese Grammatical Error Correction 1 Mar 2022 · 0 repositories · arXiv:2203.00286
-
The impact of lexical and grammatical processing on generating code from natural language 28 Feb 2022 · 2 repositories · arXiv:2202.13972
-
Enhancing Legal Argument Mining with Domain Pre-training and Neural Networks 27 Feb 2022 · 1 repository · arXiv:2202.13457
-
A Systematic Evaluation of Large Language Models of Code 26 Feb 2022 · 3 repositories · arXiv:2202.13169Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Bi-directional Joint Neural Networks for Intent Classification and Slot Filling 26 Feb 2022 · 0 repositories · arXiv:2202.13079
-
Multi-Level Contrastive Learning for Cross-Lingual Alignment 26 Feb 2022 · 0 repositories · arXiv:2202.13083
-
APEACH: Attacking Pejorative Expressions with Analysis on Crowd-Generated Hate Speech Evaluation Datasets 25 Feb 2022 · 1 repository · arXiv:2202.12459
-
Rethinking the Role of Demonstrations: What Makes In-Context Learning Work? 25 Feb 2022 · 2 repositories · arXiv:2202.12837
-
BERTVision -- A Parameter-Efficient Approach for Question Answering 24 Feb 2022 · 1 repository · arXiv:2202.12210Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Finding Inverse Document Frequency Information in BERT 24 Feb 2022 · 0 repositories · arXiv:2202.12191
-
From Natural Language to Simulations: Applying GPT-3 Codex to Automate Simulation Modeling of Logistics Systems 24 Feb 2022 · 1 repository · arXiv:2202.12107
-
Pretraining without Wordpieces: Learning Over a Vocabulary of Millions of Words 24 Feb 2022 · 0 repositories · arXiv:2202.12142
-
Probing BERT's priors with serial reproduction chains 24 Feb 2022 · 1 repository · arXiv:2202.12226Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Sky Computing: Accelerating Geo-distributed Computing in Federated Learning 24 Feb 2022 · 1 repository · arXiv:2202.11836
-
TrimBERT: Tailoring BERT for Trade-offs 24 Feb 2022 · 0 repositories · arXiv:2202.12411
-
Using calibrator to improve robustness in Machine Reading Comprehension 24 Feb 2022 · 0 repositories · arXiv:2202.11865
-
Consistent Dropout for Policy Gradient Reinforcement Learning 23 Feb 2022 · 0 repositories · arXiv:2202.11818
-
Improving CTC-based speech recognition via knowledge transferring from pre-trained language models 22 Feb 2022 · 1 repository · arXiv:2203.03582
-
JAMES: Normalizing Job Titles with Multi-Aspect Graph Embeddings and Reasoning 22 Feb 2022 · 0 repositories · arXiv:2202.10739
-
Items from Psychometric Tests as Training Data for Personality Profiling Models of Twitter Users 21 Feb 2022 · 0 repositories · arXiv:2202.10415
-
Contextual Semantic Embeddings for Ontology Subsumption Prediction 20 Feb 2022 · 2 repositories · arXiv:2202.09791
-
Do Transformers know symbolic rules, and would we know if they did? 19 Feb 2022 · 0 repositories · arXiv:2203.00162
-
Evaluating the Construct Validity of Text Embeddings with Application to Survey Questions 18 Feb 2022 · 1 repository · arXiv:2202.09166
-
SGPT: GPT Sentence Embeddings for Semantic Search 17 Feb 2022 · 1 repository · arXiv:2202.08904Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
When BERT Meets Quantum Temporal Convolution Learning for Text Classification in Heterogeneous Computing 17 Feb 2022 · 0 repositories · arXiv:2203.03550
-
A Survey of Pretraining on Graphs: Taxonomy, Methods, and Applications 16 Feb 2022 · 3 repositories · arXiv:2202.07893
-
No One Left Behind: Inclusive Federated Learning over Heterogeneous Devices 16 Feb 2022 · 0 repositories · arXiv:2202.08036
-
The NLP Task Effectiveness of Long-Range Transformers 16 Feb 2022 · 0 repositories · arXiv:2202.07856
-
BLUE at Memotion 2.0 2022: You have my Image, my Text and my Transformer 15 Feb 2022 · 0 repositories · arXiv:2202.07543
-
Defending against Reconstruction Attacks with Rényi Differential Privacy 15 Feb 2022 · 0 repositories · arXiv:2202.07623
-
One Configuration to Rule Them All? Towards Hyperparameter Transfer in Topic Models using Multi-Objective Bayesian Optimization 15 Feb 2022 · 1 repository · arXiv:2202.07631
-
Tomayto, Tomahto. Beyond Token-level Answer Equivalence for Question Answering Evaluation 15 Feb 2022 · 1 repository · arXiv:2202.07654
-
Toxic Comments Hunter : Score Severity of Toxic Comments 15 Feb 2022 · 0 repositories · arXiv:2203.03548
-
Punctuation restoration in Swedish through fine-tuned KB-BERT 14 Feb 2022 · 0 repositories · arXiv:2202.06769
-
UserBERT: Modeling Long- and Short-Term User Preferences via Self-Supervision 14 Feb 2022 · 0 repositories · arXiv:2202.07605
-
Assessment of contextualised representations in detecting outcome phrases in clinical trials 13 Feb 2022 · 0 repositories · arXiv:2203.03547
-
Automatic Issue Classifier: A Transfer Learning Framework for Classifying Issue Reports 12 Feb 2022 · 1 repository · arXiv:2202.06149
-
Maximizing Communication Efficiency for Large-scale Training via 0/1 Adam 12 Feb 2022 · 1 repository · arXiv:2202.06009
-
White-Box Attacks on Hate-speech BERT Classifiers in German with Explicit and Implicit Character Level Defense 11 Feb 2022 · 1 repository · arXiv:2202.05778
-
A Multi-task Learning Framework for Product Ranking with BERT 10 Feb 2022 · 0 repositories · arXiv:2202.05317
-
AI-based Robust Resource Allocation in End-to-End Network Slicing under Demand and CSI Uncertainties 10 Feb 2022 · 0 repositories · arXiv:2202.05131
-
Exact Solutions of a Deep Linear Network 10 Feb 2022 · 0 repositories · arXiv:2202.04777
-
Slovene SuperGLUE Benchmark: Translation and Evaluation 10 Feb 2022 · 0 repositories · arXiv:2202.04994
-
Can Open Domain Question Answering Systems Answer Visual Knowledge Questions? 9 Feb 2022 · 0 repositories · arXiv:2202.04306
-
Social Media as an Instant Source of Feedback on Water Quality 9 Feb 2022 · 0 repositories · arXiv:2202.04462
-
pNLP-Mixer: an Efficient all-MLP Architecture for Language 9 Feb 2022 · 1 repository · arXiv:2202.04350
-
Do Language Models Learn Position-Role Mappings? 8 Feb 2022 · 0 repositories · arXiv:2202.03611
-
HistBERT: A Pre-trained Language Model for Diachronic Lexical Semantic Analysis 8 Feb 2022 · 1 repository · arXiv:2202.03612
-
Logical Reasoning for Task Oriented Dialogue Systems 8 Feb 2022 · 0 repositories · arXiv:2202.04161
-
Semantic features of object concepts generated with GPT-3 8 Feb 2022 · 1 repository · arXiv:2202.03753
-
What are the best systems? New perspectives on NLP Benchmarking 8 Feb 2022 · 1 repository · arXiv:2202.03799
-
Cedille: A large autoregressive French language model 7 Feb 2022 · 1 repository · arXiv:2202.03371
-
OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework 7 Feb 2022 · 4 repositories · arXiv:2202.03052Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating natural language processing models with generalization metrics that do not need access to any training or testing data 6 Feb 2022 · 1 repository · arXiv:2202.02842
-
Classification on Sentence Embeddings for Legal Assistance 5 Feb 2022 · 0 repositories · arXiv:2202.02639
-
Ethics, Rules of Engagement, and AI: Neural Narrative Mapping Using Large Transformer Language Models 5 Feb 2022 · 0 repositories · arXiv:2202.02647
-
A Benchmark Corpus for the Detection of Automatically Generated Text in Academic Publications 4 Feb 2022 · 1 repository · arXiv:2202.02013
-
StonkBERT: Can Language Models Predict Medium-Run Stock Price Movements? 4 Feb 2022 · 0 repositories · arXiv:2202.02268
-
Temporal Attention for Language Models 4 Feb 2022 · 1 repository · arXiv:2202.02093
-
ASR-Aware End-to-end Neural Diarization 2 Feb 2022 · 0 repositories · arXiv:2202.01286
-
Co-training Improves Prompt-based Learning for Large Language Models 2 Feb 2022 · 1 repository · arXiv:2202.00828Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
GatorTron: A Large Clinical Language Model to Unlock Patient Information from Unstructured Electronic Health Records 2 Feb 2022 · 0 repositories · arXiv:2203.03540
-
L3Cube-MahaCorpus and MahaBERT: Marathi Monolingual Corpus, Marathi BERT Language Models, and Resources 2 Feb 2022 · 1 repository · arXiv:2202.01159
-
RescoreBERT: Discriminative Speech Recognition Rescoring with BERT 2 Feb 2022 · 0 repositories · arXiv:2202.01094
-
Robust Training of Neural Networks Using Scale Invariant Architectures 2 Feb 2022 · 0 repositories · arXiv:2202.00980
-
An Adaptive Deep Clustering Pipeline to Inform Text Labeling at Scale 1 Feb 2022 · 0 repositories · arXiv:2202.01211
-
A Semi-Supervised Deep Clustering Pipeline for Mining Intentions From Texts 1 Feb 2022 · 0 repositories · arXiv:2202.00802
-
Improving BERT-based Query-by-Document Retrieval with Multi-Task Optimization 1 Feb 2022 · 0 repositories · arXiv:2202.00373
-
Transformer-based Models of Text Normalization for Speech Applications 1 Feb 2022 · 0 repositories · arXiv:2202.00153
-
Memory-Efficient Backpropagation through Large Linear Layers 31 Jan 2022 · 2 repositories · arXiv:2201.13195
-
A Frustratingly Simple Approach for End-to-End Image Captioning 30 Jan 2022 · 0 repositories · arXiv:2201.12723
-
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models 28 Jan 2022 · 19 repositories · arXiv:2201.11903Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Electra: Conditional Generative Model based Predicate-Aware Query Approximation 28 Jan 2022 · 0 repositories · arXiv:2201.12420
-
Describing Differences between Text Distributions with Natural Language 28 Jan 2022 · 1 repository · arXiv:2201.12323Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Clinical-Longformer and Clinical-BigBird: Transformers for long clinical sequences 27 Jan 2022 · 1 repository · arXiv:2201.11838
-
Going Extreme: Comparative Analysis of Hate Speech in Parler and Gab 27 Jan 2022 · 1 repository · arXiv:2201.11770
-
Grad2Task: Improved Few-shot Text Classification Using Gradients for Task Representation 27 Jan 2022 · 1 repository · arXiv:2201.11576
-
DiscoScore: Evaluating Text Generation with BERT and Discourse Coherence 26 Jan 2022 · 1 repository · arXiv:2201.11176
-
DNNFuser: Generative Pre-Trained Transformer as a Generalized Mapper for Layer Fusion in DNN Accelerators 26 Jan 2022 · 0 repositories · arXiv:2201.11218
-
FiNCAT: Financial Numeral Claim Analysis Tool 26 Jan 2022 · 1 repository · arXiv:2202.00631
-
Neural Grapheme-to-Phoneme Conversion with Pre-trained Grapheme Models 26 Jan 2022 · 1 repository · arXiv:2201.10716
-
Self-supervised 3D Semantic Representation Learning for Vision-and-Language Navigation 26 Jan 2022 · 0 repositories · arXiv:2201.10788
-
Synchromesh: Reliable code generation from pre-trained language models 26 Jan 2022 · 2 repositories · arXiv:2201.11227Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
BERTHA: Video Captioning Evaluation Via Transfer-Learned Human Assessment 25 Jan 2022 · 1 repository · arXiv:2201.10243
-
Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models 25 Jan 2022 · 0 repositories · arXiv:2201.10103
-
Pre-Trained Language Transformers are Universal Image Classifiers 25 Jan 2022 · 0 repositories · arXiv:2201.10182
-
Whose Language Counts as High Quality? Measuring Language Ideologies in Text Data Selection 25 Jan 2022 · 0 repositories · arXiv:2201.10474
-
Emotion-based Modeling of Mental Disorders on Social Media 24 Jan 2022 · 0 repositories · arXiv:2201.09451
-
Polyphone disambiguation and accent prediction using pre-trained language models in Japanese TTS front-end 24 Jan 2022 · 0 repositories · arXiv:2201.09427
-
Synthetic Books 24 Jan 2022 · 0 repositories · arXiv:2201.09518
-
Unified Multimodal Punctuation Restoration Framework for Mixed-Modality Corpus 24 Jan 2022 · 1 repository · arXiv:2202.00468
-
A Large and Diverse Arabic Corpus for Language Modeling 23 Jan 2022 · 0 repositories · arXiv:2201.09227
-
An Application of Pseudo-Log-Likelihoods to Natural Language Scoring 23 Jan 2022 · 0 repositories · arXiv:2201.09377
-
Black-box Prompt Learning for Pre-trained Language Models 21 Jan 2022 · 1 repository · arXiv:2201.08531Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Dual Contrastive Learning: Text Classification via Label-Aware Data Augmentation 21 Jan 2022 · 2 repositories · arXiv:2201.08702
-
Less is Less: When Are Snippets Insufficient for Human vs Machine Relevance Estimation? 21 Jan 2022 · 0 repositories · arXiv:2201.08721
-
Cheating Automatic Short Answer Grading: On the Adversarial Usage of Adjectives and Adverbs 20 Jan 2022 · 1 repository · arXiv:2201.08318
-
DDPG-Driven Deep-Unfolding with Adaptive Depth for Channel Estimation with Sparse Bayesian Learning 20 Jan 2022 · 0 repositories · arXiv:2201.08477