Methods › General › Regularization › Weight Decay › Papers, page 64
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 64 of 108: papers 6,301 to 6,400 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Domain-Adaptive Text Classification with Structured Knowledge from Unlabeled Data 20 Jun 2022 · 1 repository · arXiv:2206.09591Syntology official (archive's flag): 4 ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models 20 Jun 2022 · 2 repositories · arXiv:2206.09557
-
When Does Re-initialization Work? 20 Jun 2022 · 0 repositories · arXiv:2206.10011
-
Two-Hop Age of Information Scheduling for Multi-UAV Assisted Mobile Edge Computing: FRL vs MADDPG 19 Jun 2022 · 0 repositories · arXiv:2206.09488
-
Argumentative Text Generation in Economic Domain 18 Jun 2022 · 1 repository · arXiv:2206.09251
-
Automatic Summarization of Russian Texts: Comparison of Extractive and Abstractive Methods 18 Jun 2022 · 0 repositories · arXiv:2206.09253
-
RuArg-2022: Argument Mining Evaluation 18 Jun 2022 · 0 repositories · arXiv:2206.09249
-
BITS Pilani at HinglishEval: Quality Evaluation for Code-Mixed Hinglish Text Using Transformers 17 Jun 2022 · 0 repositories · arXiv:2206.08680
-
Video Sparse Transformer With Attention-Guided Memory for Video Object Detection 17 Jun 2022 · 1 repository
-
An Open-Domain QA System for e-Governance 16 Jun 2022 · 0 repositories · arXiv:2206.08046
-
Autonomous Platoon Control with Integrated Deep Reinforcement Learning and Dynamic Programming 15 Jun 2022 · 0 repositories · arXiv:2206.07536
-
Detecting Harmful Online Conversational Content towards LGBTQIA+ Individuals 15 Jun 2022 · 1 repository · arXiv:2207.10032
-
Estimating Confidence of Predictions of Individual Classifiers and Their Ensembles for the Genre Classification Task 15 Jun 2022 · 0 repositories · arXiv:2206.07427
-
RDU: A Region-based Approach to Form-style Document Understanding 14 Jun 2022 · 0 repositories · arXiv:2206.06890
-
SBERT studies Meaning Representations: Decomposing Sentence Embeddings into Explainable Semantic Features 14 Jun 2022 · 1 repository · arXiv:2206.07023Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Computation Offloading and Resource Allocation in F-RANs: A Federated Deep Reinforcement Learning Approach 13 Jun 2022 · 0 repositories · arXiv:2206.05881
-
Optimal Clipping and Magnitude-aware Differentiation for Improved Quantization-aware Training 13 Jun 2022 · 1 repository · arXiv:2206.06501Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
Transition-based Abstract Meaning Representation Parsing with Contextual Embeddings 13 Jun 2022 · 0 repositories · arXiv:2206.06229
-
SGD and Weight Decay Secretly Minimize the Rank of Your Neural Network 12 Jun 2022 · 0 repositories · arXiv:2206.05794
-
From Human Days to Machine Seconds: Automatically Answering and Generating Machine Learning Final Exams 11 Jun 2022 · 0 repositories · arXiv:2206.05442
-
Comparative Snippet Generation 11 Jun 2022 · 1 repository · arXiv:2206.05473
-
Putting GPT-3's Creativity to the (Alternative Uses) Test 10 Jun 2022 · 1 repository · arXiv:2206.08932
-
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models 9 Jun 2022 · 6 repositories · arXiv:2206.04615Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Joint Encoder-Decoder Self-Supervised Pre-training for ASR 9 Jun 2022 · 0 repositories · arXiv:2206.04465
-
Traditional and context-specific spam detection in low resource settings 9 Jun 2022 · 1 repository
-
Abstraction not Memory: BERT and the English Article System 8 Jun 2022 · 1 repository · arXiv:2206.04184
-
SciDeBERTa: Learning DeBERTa for Science Technology Documents and Fine-Tuning Information Extraction Tasks 8 Jun 2022 · 1 repository
-
Always Keep your Target in Mind: Studying Semantics and Improving Performance of Neural Lexical Substitution 7 Jun 2022 · 1 repository · arXiv:2206.11815
-
An Empirical Study of IoT Security Aspects at Sentence-Level in Developer Textual Discussions 7 Jun 2022 · 0 repositories · arXiv:2206.03079
-
DynaMaR: Dynamic Prompt with Mask Token Representation 7 Jun 2022 · 0 repositories · arXiv:2206.02982
-
OCHADAI at SemEval-2022 Task 2: Adversarial Training for Multilingual Idiomaticity Detection 7 Jun 2022 · 0 repositories · arXiv:2206.03025
-
A computational psycholinguistic evaluation of the syntactic abilities of Galician BERT models at the interface of dependency resolution and training time 6 Jun 2022 · 1 repository · arXiv:2206.02440
-
Balancing Profit, Risk, and Sustainability for Portfolio Management 6 Jun 2022 · 0 repositories · arXiv:2207.02134
-
Making Large Language Models Better Reasoners with Step-Aware Verifier 6 Jun 2022 · 0 repositories · arXiv:2206.02336
-
Spam Detection Using BERT 6 Jun 2022 · 0 repositories · arXiv:2206.02443
-
What do tokens know about their characters and how do they know it? 6 Jun 2022 · 1 repository · arXiv:2206.02608
-
Sentiment Analysis of Online Travel Reviews Based on Capsule Network and Sentiment Lexicon 5 Jun 2022 · 0 repositories · arXiv:2206.02160
-
Speech Detection Task Against Asian Hate: BERT the Central, While Data-Centric Studies the Crucial 5 Jun 2022 · 0 repositories · arXiv:2206.02114
-
Comparing Performance of Different Linguistically-Backed Word Embeddings for Cyberbullying Detection 4 Jun 2022 · 0 repositories · arXiv:2206.01950
-
Extreme Compression for Pre-trained Transformers Made Simple and Efficient 4 Jun 2022 · 1 repository · arXiv:2206.01859
-
ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale Transformers 4 Jun 2022 · 3 repositories · arXiv:2206.01861Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
Automatic Generation of Programming Exercises and Code Explanations using Large Language Models 3 Jun 2022 · 0 repositories · arXiv:2206.11861
-
Differentially Private Model Compression 3 Jun 2022 · 0 repositories · arXiv:2206.01838
-
Extracting Similar Questions From Naturally-occurring Business Conversations 3 Jun 2022 · 0 repositories · arXiv:2206.01585
-
Joint Energy Dispatch and Unit Commitment in Microgrids Based on Deep Reinforcement Learning 3 Jun 2022 · 0 repositories · arXiv:2206.01663
-
TCE at Qur'an QA 2022: Arabic Language Question Answering Over Holy Qur'an Using a Post-Processed Ensemble of BERT-based Models 3 Jun 2022 · 1 repository · arXiv:2206.01550
-
Decentralized Training of Foundation Models in Heterogeneous Environments 2 Jun 2022 · 1 repository · arXiv:2206.01288Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
MMTM: Multi-Tasking Multi-Decoder Transformer for Math Word Problems 2 Jun 2022 · 0 repositories · arXiv:2206.01268
-
Assessing Group-level Gender Bias in Professional Evaluations: The Case of Medical Student End-of-Shift Feedback 1 Jun 2022 · 0 repositories · arXiv:2206.00234
-
BERT-Sort: A Zero-shot MLM Semantic Encoder on Ordinal Features for AutoML 1 Jun 2022 · 1 repository
-
Order-sensitive Shapley Values for Evaluating Conceptual Soundness of NLP Models 1 Jun 2022 · 0 repositories · arXiv:2206.00192
-
Romantic-Computing 1 Jun 2022 · 0 repositories · arXiv:2206.11864
-
A Unified Framework for Emotion Identification and Generation in Dialogues 31 May 2022 · 0 repositories · arXiv:2205.15513
-
Knowledge Graph - Deep Learning: A Case Study in Question Answering in Aviation Safety Domain 31 May 2022 · 0 repositories · arXiv:2205.15952
-
Automatic Short Math Answer Grading via In-context Meta-learning 30 May 2022 · 1 repository · arXiv:2205.15219
-
Billions of Parameters Are Worth More Than In-domain Training Data: A case study in the Legal Case Entailment Task 30 May 2022 · 1 repository · arXiv:2205.15172
-
Multi-Agent Reinforcement Learning is a Sequence Modeling Problem 30 May 2022 · 1 repository · arXiv:2205.14953
-
Prompting ELECTRA: Few-Shot Learning with Discriminative Pre-Trained Models 30 May 2022 · 1 repository · arXiv:2205.15223
-
Truly Deterministic Policy Optimization 30 May 2022 · 1 repository · arXiv:2205.15379
-
COVID-19 Literature Mining and Retrieval using Text Mining Approaches 29 May 2022 · 0 repositories · arXiv:2205.14781
-
CPED: A Large-Scale Chinese Personalized and Emotional Dialogue Dataset for Conversational AI 29 May 2022 · 1 repository · arXiv:2205.14727
-
Micro-Expression Recognition Based on Attribute Information Embedding and Cross-modal Contrastive Learning 29 May 2022 · 0 repositories · arXiv:2205.14643
-
Urdu News Article Recommendation Model using Natural Language Processing Techniques 29 May 2022 · 0 repositories · arXiv:2206.11862
-
Teaching Models to Express Their Uncertainty in Words 28 May 2022 · 1 repository · arXiv:2205.14334Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness 27 May 2022 · 13 repositories · arXiv:2205.14135Syntology official: no sample here; runs from other or unrecorded repositories · 24 ran (of which 2 constructed an object rather than computing a result; 18 with no instrument failure: 0 honoured, 1 violated, 17 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 30 harvested samples) · 1 pointer-only (licence)
-
Multimodal Masked Autoencoders Learn Transferable Representations 27 May 2022 · 3 repositories · arXiv:2205.14204Syntology official (archive's flag): 3 ran · 12 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples)
-
kNN-Prompt: Nearest Neighbor Zero-Shot Inference 27 May 2022 · 1 repository · arXiv:2205.13792
-
What Dense Graph Do You Need for Self-Attention? 27 May 2022 · 1 repository · arXiv:2205.14014
-
Federated Split BERT for Heterogeneous Text Classification 26 May 2022 · 0 repositories · arXiv:2205.13299
-
Learning Dialogue Representations from Consecutive Utterances 26 May 2022 · 1 repository · arXiv:2205.13568
-
Leveraging Dependency Grammar for Fine-Grained Offensive Language Detection using Graph Convolutional Networks 26 May 2022 · 1 repository · arXiv:2205.13164
-
The Document Vectors Using Cosine Similarity Revisited 26 May 2022 · 1 repository · arXiv:2205.13357
-
Training and Inference on Any-Order Autoregressive Models the Right Way 26 May 2022 · 1 repository · arXiv:2205.13554
-
BiT: Robustly Binarized Multi-distilled Transformer 25 May 2022 · 3 repositories · arXiv:2205.13016Syntology official (archive's flag): 2 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Conditional set generation using Seq2seq models 25 May 2022 · 0 repositories · arXiv:2205.12485
-
Large Language Models are Few-Shot Clinical Information Extractors 25 May 2022 · 0 repositories · arXiv:2205.12689
-
Lifelong Learning Natural Language Processing Approach for Multilingual Data Classification 25 May 2022 · 0 repositories · arXiv:2206.11867
-
NaturalProver: Grounded Mathematical Proof Generation with Language Models 25 May 2022 · 1 repository · arXiv:2205.12910Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ORCA: Interpreting Prompted Language Models via Locating Supporting Data Evidence in the Ocean of Pretraining Data 25 May 2022 · 0 repositories · arXiv:2205.12600
-
RobustLR: Evaluating Robustness to Logical Perturbation in Deductive Reasoning 25 May 2022 · 1 repository · arXiv:2205.12598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Text-to-Face Generation with StyleGAN2 25 May 2022 · 0 repositories · arXiv:2205.12512
-
Do we need Label Regularization to Fine-tune Pre-trained Language Models? 25 May 2022 · 0 repositories · arXiv:2205.12428
-
Train Flat, Then Compress: Sharpness-Aware Minimization Learns More Compressible Models 25 May 2022 · 0 repositories · arXiv:2205.12694
-
Transcormer: Transformer for Sentence Scoring with Sliding Language Modeling 25 May 2022 · 1 repository · arXiv:2205.12986Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
VulBERTa: Simplified Source Code Pre-Training for Vulnerability Detection 25 May 2022 · 1 repository · arXiv:2205.12424
-
FLUTE: Figurative Language Understanding through Textual Explanations 24 May 2022 · 1 repository · arXiv:2205.12404
-
Formulating Few-shot Fine-tuning Towards Language Model Pre-training: A Pilot Study on Named Entity Recognition 24 May 2022 · 1 repository · arXiv:2205.11799
-
Garden-Path Traversal in GPT-2 24 May 2022 · 1 repository · arXiv:2205.12302Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
The Authenticity Gap in Human Evaluation 24 May 2022 · 0 repositories · arXiv:2205.11930
-
K-12BERT: BERT for K-12 education 24 May 2022 · 1 repository · arXiv:2205.12335
-
On the Role of Bidirectionality in Language Model Pre-Training 24 May 2022 · 0 repositories · arXiv:2205.11726
-
Partial-input baselines show that NLI models can ignore context, but they don't 24 May 2022 · 1 repository · arXiv:2205.12181
-
RetroMAE: Pre-Training Retrieval-oriented Language Models Via Masked Auto-Encoder 24 May 2022 · 1 repository · arXiv:2205.12035
-
Sparse Mixers: Combining MoE and Mixing to build a more efficient BERT 24 May 2022 · 1 repository · arXiv:2205.12399
-
Word-order typology in Multilingual BERT: A case study in subordinate-clause detection 24 May 2022 · 0 repositories · arXiv:2205.11987
-
Artificial intelligence for topic modelling in Hindu philosophy: mapping themes between the Upanishads and the Bhagavad Gita 23 May 2022 · 1 repository · arXiv:2205.11020
-
Improving Short Text Classification With Augmented Data Using GPT-3 23 May 2022 · 0 repositories · arXiv:2205.10981
-
KOLD: Korean Offensive Language Dataset 23 May 2022 · 1 repository · arXiv:2205.11315
-
Learning to Ignore Adversarial Attacks 23 May 2022 · 0 repositories · arXiv:2205.11551
-
Looking for a Handsome Carpenter! Debiasing GPT-3 Job Advertisements 23 May 2022 · 1 repository · arXiv:2205.11374