Methods › General › Regularization › Weight Decay › Papers, page 62
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 62 of 108: papers 6,101 to 6,200 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Unsupervised Early Exit in DNNs with Multiple Exits 20 Sep 2022 · 1 repository · arXiv:2209.09480
-
Meta-Adapters: Parameter Efficient Few-shot Fine-tuning through Meta-Learning 19 Sep 2022 · 1 repository
-
Will It Blend? Mixing Training Paradigms & Prompting for Argument Quality Prediction 19 Sep 2022 · 0 repositories · arXiv:2209.08966
-
Detecting Generated Scientific Papers using an Ensemble of Transformer Models 17 Sep 2022 · 1 repository · arXiv:2209.08283
-
CodeQueries: A Dataset of Semantic Queries over Code 17 Sep 2022 · 1 repository · arXiv:2209.08372
-
Changing the Representation: Examining Language Representation for Neural Sign Language Production 16 Sep 2022 · 0 repositories · arXiv:2210.06312
-
Psychologically-informed chain-of-thought prompts for metaphor understanding in large language models 16 Sep 2022 · 1 repository · arXiv:2209.08141Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Text and Patterns: For Effective Chain of Thought, It Takes Two to Tango 16 Sep 2022 · 0 repositories · arXiv:2209.07686
-
Machine Reading, Fast and Slow: When Do Models "Understand" Language? 15 Sep 2022 · 0 repositories · arXiv:2209.07430
-
uChecker: Masked Pretrained Language Models as Unsupervised Chinese Spelling Checkers 15 Sep 2022 · 0 repositories · arXiv:2209.07068
-
Automated Fidelity Assessment for Strategy Training in Inpatient Rehabilitation using Natural Language Processing 14 Sep 2022 · 0 repositories · arXiv:2209.06727
-
BERT-based Ensemble Approaches for Hate Speech Detection 14 Sep 2022 · 0 repositories · arXiv:2209.06505
-
Efficient Quantized Sparse Matrix Operations on Tensor Cores 14 Sep 2022 · 1 repository · arXiv:2209.06979
-
Out of One, Many: Using Language Models to Simulate Human Samples 14 Sep 2022 · 0 repositories · arXiv:2209.06899
-
Pre-training for Information Retrieval: Are Hyperlinks Fully Explored? 14 Sep 2022 · 0 repositories · arXiv:2209.06583
-
CNN-Trans-Enc: A CNN-Enhanced Transformer-Encoder On Top Of Static BERT representations for Document Classification 13 Sep 2022 · 0 repositories · arXiv:2209.06344
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 13 Sep 2022 · 0 repositories · arXiv:2209.05707
-
SkIn: Skimming-Intensive Long-Text Classification Using BERT for Medical Corpus 13 Sep 2022 · 0 repositories · arXiv:2209.05741
-
A new hazard event classification model via deep learning and multifractal 12 Sep 2022 · 0 repositories · arXiv:2209.05263
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 12 Sep 2022 · 0 repositories · arXiv:2209.05286
-
Chain of Explanation: New Prompting Method to Generate Higher Quality Natural Language Explanation for Implicit Hate Speech 11 Sep 2022 · 0 repositories · arXiv:2209.04889
-
Probing for Understanding of English Verb Classes and Alternations in Large Pre-trained Language Models 11 Sep 2022 · 0 repositories · arXiv:2209.04811
-
Yes, DLGM! A novel hierarchical model for hazard classification 10 Sep 2022 · 0 repositories · arXiv:2209.04576
-
EchoCoTr: Estimation of the Left Ventricular Ejection Fraction from Spatiotemporal Echocardiography 9 Sep 2022 · 1 repository · arXiv:2209.04242
-
Trigger Warnings: Bootstrapping a Violence Detector for FanFiction 9 Sep 2022 · 0 repositories · arXiv:2209.04409
-
CLaCLab at SocialDisNER: Using Medical Gazetteers for Named-Entity Recognition of Disease Mentions in Spanish Tweets 8 Sep 2022 · 1 repository · arXiv:2209.03528
-
FADE: Enabling Federated Adversarial Training on Heterogeneous Resource-Constrained Edge Devices 8 Sep 2022 · 0 repositories · arXiv:2209.03839
-
5q032e@SMM4H'22: Transformer-based classification of premise in tweets related to COVID-19 8 Sep 2022 · 0 repositories · arXiv:2209.03851
-
Why So Toxic? Measuring and Triggering Toxic Behavior in Open-Domain Chatbots 7 Sep 2022 · 0 repositories · arXiv:2209.03463
-
Multilingual Bidirectional Unsupervised Translation Through Multilingual Finetuning and Back-Translation 6 Sep 2022 · 1 repository · arXiv:2209.02821
-
A New Approach to Training Multiple Cooperative Agents for Autonomous Driving 5 Sep 2022 · 0 repositories · arXiv:2209.02157
-
ChemBERTa-2: Towards Chemical Foundation Models 5 Sep 2022 · 2 repositories · arXiv:2209.01712
-
Distilling the Knowledge of BERT for CTC-based ASR 5 Sep 2022 · 0 repositories · arXiv:2209.02030
-
Evaluating the Susceptibility of Pre-Trained Language Models via Handcrafted Adversarial Examples 5 Sep 2022 · 0 repositories · arXiv:2209.02128
-
Do Large Language Models know what humans know? 4 Sep 2022 · 1 repository · arXiv:2209.01515
-
Every picture tells a story: Image-grounded controllable stylistic story generation 4 Sep 2022 · 0 repositories · arXiv:2209.01638
-
Generalization in Neural Networks: A Broad Survey 4 Sep 2022 · 0 repositories · arXiv:2209.01610
-
Elaboration-Generating Commonsense Question Answering at Scale 2 Sep 2022 · 1 repository · arXiv:2209.01232
-
FOLIO: Natural Language Reasoning with First-Order Logic 2 Sep 2022 · 1 repository · arXiv:2209.00840
-
GReS: Graphical Cross-domain Recommendation for Supply Chain Platform 2 Sep 2022 · 0 repositories · arXiv:2209.01031
-
Optimal bump functions for shallow ReLU networks: Weight decay, depth separation and the curse of dimensionality 2 Sep 2022 · 0 repositories · arXiv:2209.01173
-
Enhancing Semantic Understanding with Self-supervised Methods for Abstractive Dialogue Summarization 1 Sep 2022 · 0 repositories · arXiv:2209.00278
-
Isotropic Representation Can Improve Dense Retrieval 1 Sep 2022 · 1 repository · arXiv:2209.00218
-
Distilling Multi-Scale Knowledge for Event Temporal Relation Extraction 1 Sep 2022 · 0 repositories · arXiv:2209.00568
-
Negation detection in Dutch clinical texts: an evaluation of rule-based and machine learning methods 1 Sep 2022 · 1 repository · arXiv:2209.00470
-
Few-Shot Learning for Clinical Natural Language Processing Using Siamese Neural Networks 31 Aug 2022 · 0 repositories · arXiv:2208.14923
-
Large-scale Multi-granular Concept Extraction Based on Machine Reading Comprehension 30 Aug 2022 · 1 repository · arXiv:2208.14139
-
No means ‘No’; a non-im-proper modeling approach, with embedded speculative context 30 Aug 2022 · 0 repositories
-
SwiftPruner: Reinforced Evolutionary Pruning for Efficient Ad Relevance 30 Aug 2022 · 0 repositories · arXiv:2209.00625
-
Transformers with Learnable Activation Functions 30 Aug 2022 · 2 repositories · arXiv:2208.14111Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Multi-dimensional Racism Classification during COVID-19: Stigmatization, Offensiveness, Blame, and Exclusion 29 Aug 2022 · 0 repositories · arXiv:2208.13318
-
Goal-Conditioned Q-Learning as Knowledge Distillation 28 Aug 2022 · 1 repository · arXiv:2208.13298
-
Building the Intent Landscape of Real-World Conversational Corpora with Extractive Question-Answering Transformers 26 Aug 2022 · 0 repositories · arXiv:2208.12886
-
Task-specific Pre-training and Prompt Decomposition for Knowledge Graph Population with Language Models 26 Aug 2022 · 1 repository · arXiv:2208.12539
-
On Reality and the Limits of Language Data: Aligning LLMs with Human Norms 25 Aug 2022 · 0 repositories · arXiv:2208.11981
-
Addressing Token Uniformity in Transformers via Singular Value Transformation 24 Aug 2022 · 1 repository · arXiv:2208.11790Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Evaluate Confidence Instead of Perplexity for Zero-shot Commonsense Reasoning 23 Aug 2022 · 0 repositories · arXiv:2208.11007
-
Prompting as Probing: Using Language Models for Knowledge Base Construction 23 Aug 2022 · 1 repository · arXiv:2208.11057Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
A Syntax Aware BERT for Identifying Well-Formed Queries in a Curriculum Framework 21 Aug 2022 · 0 repositories · arXiv:2208.09912
-
CMSBERT-CLR: Context-driven Modality Shifting BERT with Contrastive Learning for linguistic, visual, acoustic Representations 21 Aug 2022 · 0 repositories · arXiv:2209.07424
-
BSpell: A CNN-Blended BERT Based Bangla Spell Checker 20 Aug 2022 · 1 repository · arXiv:2208.09709
-
Combining Compressions for Multiplicative Size Scaling on Natural Language Tasks 20 Aug 2022 · 0 repositories · arXiv:2208.09684
-
Pretrained Language Encoders are Natural Tagging Frameworks for Aspect Sentiment Triplet Extraction 20 Aug 2022 · 0 repositories · arXiv:2208.09617
-
SPOT: Knowledge-Enhanced Language Representations for Information Extraction 20 Aug 2022 · 0 repositories · arXiv:2208.09625
-
Graph-Augmented Cyclic Learning Framework for Similarity Estimation of Medical Clinical Notes 19 Aug 2022 · 0 repositories · arXiv:2208.09437
-
UniCausal: Unified Benchmark and Repository for Causal Text Mining 19 Aug 2022 · 1 repository · arXiv:2208.09163
-
MulZDG: Multilingual Code-Switching Framework for Zero-shot Dialogue Generation 18 Aug 2022 · 1 repository · arXiv:2208.08629
-
Using Large Language Models to Simulate Multiple Humans and Replicate Human Subject Studies 18 Aug 2022 · 2 repositories · arXiv:2208.10264Syntology community repositories only · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
VAuLT: Augmenting the Vision-and-Language Transformer for Sentiment Classification on Social Media 18 Aug 2022 · 1 repository · arXiv:2208.09021Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
EmoMent: An Emotion Annotated Mental Health Corpus from two South Asian Countries 17 Aug 2022 · 0 repositories · arXiv:2208.08486
-
Neural Embeddings for Text 17 Aug 2022 · 1 repository · arXiv:2208.08386
-
Transformer Encoder for Social Science 17 Aug 2022 · 1 repository · arXiv:2208.08005
-
Continuous Active Learning Using Pretrained Transformers 15 Aug 2022 · 0 repositories · arXiv:2208.06955
-
MoCapAct: A Multi-Task Dataset for Simulated Humanoid Control 15 Aug 2022 · 1 repository · arXiv:2208.07363Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Targeted Honeyword Generation with Language Models 15 Aug 2022 · 0 repositories · arXiv:2208.06946
-
Teacher Guided Training: An Efficient Framework for Knowledge Transfer 14 Aug 2022 · 0 repositories · arXiv:2208.06825
-
Text Difficulty Study: Do machines behave the same as humans regarding text difficulty? 14 Aug 2022 · 0 repositories · arXiv:2208.14509
-
Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models 13 Aug 2022 · 9 repositories · arXiv:2208.06677Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Interpreting BERT-based Text Similarity via Activation and Saliency Maps 13 Aug 2022 · 0 repositories · arXiv:2208.06612
-
Is Your Model Sensitive? SPeDaC: A New Benchmark for Detecting and Classifying Sensitive Personal Data 12 Aug 2022 · 0 repositories · arXiv:2208.06216
-
Pre-training Tasks for User Intent Detection and Embedding Retrieval in E-commerce Search 12 Aug 2022 · 1 repository · arXiv:2208.06150
-
A Model of Anaphoric Ambiguities using Sheaf Theoretic Quantum-like Contextuality and BERT 11 Aug 2022 · 0 repositories · arXiv:2208.05720
-
A Twitter-Driven Deep Learning Mechanism for the Determination of Vehicle Hijacking Spots in Cities 11 Aug 2022 · 0 repositories · arXiv:2208.10280
-
Bayesian Soft Actor-Critic: A Directed Acyclic Strategy Graph Based Deep Reinforcement Learning 11 Aug 2022 · 2 repositories · arXiv:2208.06033
-
Searching for chromate replacements using natural language processing and machine learning algorithms 11 Aug 2022 · 0 repositories · arXiv:2208.05672
-
Plug-and-Play Model-Agnostic Counterfactual Policy Synthesis for Deep Reinforcement Learning based Recommendation 10 Aug 2022 · 0 repositories · arXiv:2208.05142
-
A Multimodal Transformer: Fusing Clinical Notes with Structured EHR Data for Interpretable In-Hospital Mortality Prediction 9 Aug 2022 · 0 repositories · arXiv:2208.10240
-
E2EG: End-to-End Node Classification Using Graph Topology and Text-based Node Attributes 9 Aug 2022 · 1 repository · arXiv:2208.04609Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 4 pointer-only (licence)
-
Emotion Detection From Tweets Using a BERT and SVM Ensemble Model 9 Aug 2022 · 1 repository · arXiv:2208.04547
-
Exploring Hate Speech Detection with HateXplain and BERT 9 Aug 2022 · 1 repository · arXiv:2208.04489
-
Debiased Large Language Models Still Associate Muslims with Uniquely Violent Acts 8 Aug 2022 · 0 repositories · arXiv:2208.04417
-
Efficient Fine-Tuning of Compressed Language Models with Learners 3 Aug 2022 · 0 repositories · arXiv:2208.02070
-
A Comparative Study on COVID-19 Fake News Detection Using Different Transformer Based Models 2 Aug 2022 · 0 repositories · arXiv:2208.01355
-
Automatic Classification of Bug Reports Based on Multiple Text Information and Reports' Intention 2 Aug 2022 · 0 repositories · arXiv:2208.01274
-
Debiasing Gender Bias in Information Retrieval Models 2 Aug 2022 · 0 repositories · arXiv:2208.01755
-
giMLPs: Gate with Inhibition Mechanism in MLPs 1 Aug 2022 · 1 repository · arXiv:2208.00929
-
Performance Comparison of Deep RL Algorithms for Energy Systems Optimal Scheduling 1 Aug 2022 · 1 repository · arXiv:2208.00728
-
Interacting with next-phrase suggestions: How suggestion systems aid and influence the cognitive processes of writing 1 Aug 2022 · 0 repositories · arXiv:2208.00636
-
What Can Transformers Learn In-Context? A Case Study of Simple Function Classes 1 Aug 2022 · 2 repositories · arXiv:2208.01066
-
Aggretriever: A Simple Approach to Aggregate Textual Representations for Robust Dense Passage Retrieval 31 Jul 2022 · 1 repository · arXiv:2208.00511