Methods › General › Regularization › Weight Decay › Papers, page 65
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 65 of 108: papers 6,401 to 6,500 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Multiple Domain Cyberspace Attack and Defense Game Based on Reward Randomization Reinforcement Learning 23 May 2022 · 0 repositories · arXiv:2205.10990
-
On the Paradox of Learning to Reason from Data 23 May 2022 · 1 repository · arXiv:2205.11502Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Outliers Dimensions that Disrupt Transformers Are Driven by Frequency 23 May 2022 · 1 repository · arXiv:2205.11380Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 16 harvested samples)
-
Parameter-Efficient Sparsity for Large Language Models Fine-Tuning 23 May 2022 · 2 repositories · arXiv:2205.11005Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Penguins Don't Fly: Reasoning about Generics through Instantiations and Exceptions 23 May 2022 · 0 repositories · arXiv:2205.11658
-
Prompt Tuning for Discriminative Pre-trained Language Models 23 May 2022 · 1 repository · arXiv:2205.11166Syntology official (archive's flag): 7 ran · 7 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples)
-
RL with KL penalties is better viewed as Bayesian inference 23 May 2022 · 0 repositories · arXiv:2205.11275
-
The Diminishing Returns of Masked Language Models to Science 23 May 2022 · 0 repositories · arXiv:2205.11342
-
Simple Recurrence Improves Masked Language Models 23 May 2022 · 0 repositories · arXiv:2205.11588
-
A Graph Enhanced BERT Model for Event Prediction 22 May 2022 · 0 repositories · arXiv:2205.10822
-
GraphMAE: Self-Supervised Masked Graph Autoencoders 22 May 2022 · 3 repositories · arXiv:2205.10803Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Instruction Induction: From Few Examples to Natural Language Task Descriptions 22 May 2022 · 1 repository · arXiv:2205.10782Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
A Study on Transformer Configuration and Training Objective 21 May 2022 · 0 repositories · arXiv:2205.10505
-
Least-to-Most Prompting Enables Complex Reasoning in Large Language Models 21 May 2022 · 1 repository · arXiv:2205.10625
-
Life after BERT: What do Other Muppets Understand about Language? 21 May 2022 · 1 repository · arXiv:2205.10696
-
Pre-training Data Quality and Quantity for a Low-Resource Language: New Corpus and BERT Models for Maltese 21 May 2022 · 1 repository · arXiv:2205.10517
-
Current Trends and Approaches in Synonyms Extraction: Potential Adaptation to Arabic 20 May 2022 · 0 repositories · arXiv:2205.10412
-
Exploring Extreme Parameter Compression for Pre-trained Language Models 20 May 2022 · 1 repository · arXiv:2205.10036
-
Pre-training Transformer Models with Sentence-Level Objectives for Answer Sentence Selection 20 May 2022 · 0 repositories · arXiv:2205.10455
-
Progressive Class Semantic Matching for Semi-supervised Text Classification 20 May 2022 · 1 repository · arXiv:2205.10189
-
Prototypical Calibration for Few-shot Learning of Language Models 20 May 2022 · 1 repository · arXiv:2205.10183
-
ArabGlossBERT: Fine-Tuning BERT on Context-Gloss Pairs for WSD 19 May 2022 · 0 repositories · arXiv:2205.09685
-
Automated Scoring for Reading Comprehension via In-context BERT Tuning 19 May 2022 · 1 repository · arXiv:2205.09864
-
Overcoming Language Disparity in Online Content Classification with Multimodal Learning 19 May 2022 · 1 repository · arXiv:2205.09744
-
Psychiatric Scale Guided Risky Post Screening for Early Detection of Depression 19 May 2022 · 1 repository · arXiv:2205.09497
-
Towards Understanding Gender-Seniority Compound Bias in Natural Language Generation 19 May 2022 · 1 repository · arXiv:2205.09830
-
Evaluation of Transfer Learning for Polish with a Text-to-Text Model 18 May 2022 · 0 repositories · arXiv:2205.08808
-
Learning Rate Curriculum 18 May 2022 · 1 repository · arXiv:2205.09180
-
Persian Natural Language Inference: A Meta-learning approach 18 May 2022 · 1 repository · arXiv:2205.08755
-
Feature Aggregation in Zero-Shot Cross-Lingual Transfer Using Multilingual BERT 17 May 2022 · 0 repositories · arXiv:2205.08497
-
M6-Rec: Generative Pretrained Language Models are Open-Ended Recommender Systems 17 May 2022 · 0 repositories · arXiv:2205.08084
-
SAMU-XLSR: Semantically-Aligned Multimodal Utterance-level Cross-Lingual Speech Representation 17 May 2022 · 0 repositories · arXiv:2205.08180
-
SEMI-FND: Stacked Ensemble Based Multimodal Inference For Faster Fake News Detection 17 May 2022 · 0 repositories · arXiv:2205.08159
-
ViralBERT: A User Focused BERT-Based Approach to Virality Prediction 17 May 2022 · 1 repository · arXiv:2206.10298
-
Chemical transformer compression for accelerating both training and inference of molecular modeling 16 May 2022 · 1 repository · arXiv:2205.07582
-
Harnessing Multilingual Resources to Question Answering in Arabic 16 May 2022 · 0 repositories · arXiv:2205.08024
-
Heroes, Villains, and Victims, and GPT-3: Automated Extraction of Character Roles Without Training Data 16 May 2022 · 0 repositories · arXiv:2205.07557
-
The AI Teacher Test: Measuring the Pedagogical Ability of Blender and GPT-3 in Educational Dialogues 16 May 2022 · 1 repository · arXiv:2205.07540
-
What GPT Knows About Who is Who 16 May 2022 · 1 repository · arXiv:2205.07407
-
Discovering Latent Concepts Learned in BERT 15 May 2022 · 0 repositories · arXiv:2205.07237
-
Learning Lip-Based Audio-Visual Speaker Embeddings with AV-HuBERT 15 May 2022 · 1 repository · arXiv:2205.07180
-
Fake News Quick Detection on Dynamic Heterogeneous Information Networks 14 May 2022 · 0 repositories · arXiv:2205.07039
-
Naturalistic Causal Probing for Morpho-Syntax 14 May 2022 · 1 repository · arXiv:2205.07043
-
A Study of the Attention Abnormality in Trojaned BERTs 13 May 2022 · 1 repository · arXiv:2205.08305Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Improving Contextual Representation with Gloss Regularized Pre-training 13 May 2022 · 0 repositories · arXiv:2205.06603
-
AppTek's Submission to the IWSLT 2022 Isometric Spoken Language Translation Task 12 May 2022 · 0 repositories · arXiv:2205.05807
-
Is the Computation of Abstract Sameness Relations Human-Like in Neural Language Models? 12 May 2022 · 0 repositories · arXiv:2205.06149
-
NER-MQMRC: Formulating Named Entity Recognition as Multi Question Machine Reading Comprehension 12 May 2022 · 0 repositories · arXiv:2205.05904
-
A time-varying study of Chinese investor sentiment, stock market liquidity and volatility: Based on deep learning BERT model and TVP-VAR model 11 May 2022 · 0 repositories · arXiv:2205.05719
-
Clinical Prompt Learning with Frozen Language Models 11 May 2022 · 1 repository · arXiv:2205.05535
-
Query-Based Keyphrase Extraction from Long Documents 11 May 2022 · 1 repository · arXiv:2205.05391
-
Towards the Generation of Musical Explanations with GPT-3 11 May 2022 · 1 repository · arXiv:2206.08264
-
Deep learning based Chinese text sentiment mining and stock market correlation research 10 May 2022 · 0 repositories · arXiv:2205.04743
-
Hybrid Reinforcement Learning for STAR-RISs: A Coupled Phase-Shift Model Based Beamformer 10 May 2022 · 0 repositories · arXiv:2205.05029
-
Problems with Cosine as a Measure of Embedding Similarity for High Frequency Words 10 May 2022 · 2 repositories · arXiv:2205.05092
-
Ratatouille: A tool for Novel Recipe Generation 10 May 2022 · 0 repositories · arXiv:2206.08267
-
Reducing Activation Recomputation in Large Transformer Models 10 May 2022 · 4 repositories · arXiv:2205.05198Syntology official: harvested, nothing ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
UL2: Unifying Language Learning Paradigms 10 May 2022 · 2 repositories · arXiv:2205.05131Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples)
-
Automated Evaluation for Student Argumentative Writing: A Survey 9 May 2022 · 0 repositories · arXiv:2205.04083
-
Long Document Re-ranking with Modular Re-ranker 9 May 2022 · 1 repository · arXiv:2205.04275
-
Multi-segment preserving sampling for deep manifold sampler 9 May 2022 · 0 repositories · arXiv:2205.04259
-
Research on the correlation between text emotion mining and stock market based on deep learning 9 May 2022 · 0 repositories · arXiv:2205.06675
-
On the Use of BERT for Automated Essay Scoring: Joint Learning of Multi-Scale Essay Representation 8 May 2022 · 1 repository · arXiv:2205.03835
-
AKI-BERT: a Pre-trained Clinical Language Model for Early Prediction of Acute Kidney Injury 7 May 2022 · 1 repository · arXiv:2205.03695
-
EmotionFlow: Capture the Dialogue Level Emotion Transitions 7 May 2022 · 1 repository
-
Improving Downstream Task Performance by Treating Numbers as Entities 7 May 2022 · 0 repositories · arXiv:2205.03559
-
A Data Cartography based MixUp for Pre-trained Language Models 6 May 2022 · 1 repository · arXiv:2205.03403
-
Stock Price Prediction Based on Natural Language Processing 6 May 2022 · 2 repositories
-
The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning 6 May 2022 · 1 repository · arXiv:2205.03401
-
When a sentence does not introduce a discourse entity, Transformer-based models still sometimes refer to it 6 May 2022 · 1 repository · arXiv:2205.03472
-
BORT: Back and Denoising Reconstruction for End-to-End Task-Oriented Dialog 5 May 2022 · 1 repository · arXiv:2205.02471
-
Exploiting Global and Local Hierarchies for Hierarchical Text Classification 5 May 2022 · 1 repository · arXiv:2205.02613
-
Hyperbolic Relevance Matching for Neural Keyphrase Extraction 4 May 2022 · 1 repository · arXiv:2205.02047
-
Provably Confidential Language Modelling 4 May 2022 · 1 repository · arXiv:2205.01863Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Using virtual edges to extract keywords from texts modeled as complex networks 4 May 2022 · 0 repositories · arXiv:2205.02172
-
Contrastive Learning for Prompt-Based Few-Shot Language Learners 3 May 2022 · 1 repository · arXiv:2205.01308Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
Efficient Fine-Tuning of BERT Models on the Edge 3 May 2022 · 0 repositories · arXiv:2205.01541
-
Explain and Conquer: Personalised Text-based Reviews to Achieve Transparency 3 May 2022 · 0 repositories · arXiv:2205.01759
-
Finding patterns in Knowledge Attribution for Transformers 3 May 2022 · 0 repositories · arXiv:2205.01366
-
Mixed-effects transformers for hierarchical adaptation 3 May 2022 · 1 repository · arXiv:2205.01749
-
Predicting Issue Types with seBERT 3 May 2022 · 1 repository · arXiv:2205.01335
-
SemAttack: Natural Textual Attacks via Different Semantic Spaces 3 May 2022 · 1 repository · arXiv:2205.01287Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
BERTops: Studying BERT Representations under a Topological Lens 2 May 2022 · 1 repository · arXiv:2205.00953
-
Entity-aware Transformers for Entity Search 2 May 2022 · 1 repository · arXiv:2205.00820
-
Improving Students' Academic Performance with AI and Semantic Technologies 2 May 2022 · 1 repository · arXiv:2206.03213
-
OPT: Open Pre-trained Transformer Language Models 2 May 2022 · 11 repositories · arXiv:2205.01068Syntology official (archive's flag): 9 ran · 14 ran (of which 2 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified (of 24 harvested samples) · 17 pointer-only (licence)
-
Detecting COVID-19 Conspiracy Theories with Transformers and TF-IDF 1 May 2022 · 0 repositories · arXiv:2205.00377
-
StorSeismic: A new paradigm in deep learning for seismic processing 30 Apr 2022 · 1 repository · arXiv:2205.00222
-
ExaASC: A General Target-Based Stance Detection Corpus in Arabic Language 29 Apr 2022 · 1 repository · arXiv:2204.13979
-
Training Language Models with Language Feedback 29 Apr 2022 · 0 repositories · arXiv:2204.14146
-
QRelScore: Better Evaluating Generated Questions with Deeper Understanding of Context-aware Relevance 29 Apr 2022 · 0 repositories · arXiv:2204.13921
-
Inferring Implicit Relations in Complex Questions with Language Models 28 Apr 2022 · 1 repository · arXiv:2204.13778
-
On the Effect of Pretraining Corpora on In-context Learning by a Large-scale Language Model 28 Apr 2022 · 0 repositories · arXiv:2204.13509
-
RobBERTje: a Distilled Dutch BERT Model 28 Apr 2022 · 0 repositories · arXiv:2204.13511
-
Tailor: A Prompt-Based Approach to Attribute-Based Controlled Text Generation 28 Apr 2022 · 0 repositories · arXiv:2204.13362
-
An End-to-End Dialogue Summarization System for Sales Calls 27 Apr 2022 · 0 repositories · arXiv:2204.12951
-
Better Query Graph Selection for Knowledge Base Question Answering 27 Apr 2022 · 0 repositories · arXiv:2204.12662
-
Modern Baselines for SPARQL Semantic Parsing 27 Apr 2022 · 1 repository · arXiv:2204.12793
-
RigoBERTa: A State-of-the-Art Language Model For Spanish 27 Apr 2022 · 0 repositories · arXiv:2205.10233
-
SkillSpan: Hard and Soft Skill Extraction from English Job Postings 27 Apr 2022 · 1 repository · arXiv:2204.12811