Methods › General › Regularization › Weight Decay › Papers, page 102
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 102 of 108: papers 10,101 to 10,200 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Phenotyping of Clinical Notes with Improved Document Classification Models Using Contextualized Neural Language Models 30 Oct 2019 · 2 repositories · arXiv:1910.13664
-
Time to Take Emoji Seriously: They Vastly Improve Casual Conversational Models 30 Oct 2019 · 0 repositories · arXiv:1910.13793
-
Sentence Embeddings for Russian NLU 29 Oct 2019 · 1 repository · arXiv:1910.13291
-
Inducing brain-relevant bias in natural language processing models 29 Oct 2019 · 1 repository · arXiv:1911.03268Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Learning Rich Image Region Representation for Visual Question Answering 29 Oct 2019 · 0 repositories · arXiv:1910.13077
-
A BERT-Based Transfer Learning Approach for Hate Speech Detection in Online Social Media 28 Oct 2019 · 2 repositories · arXiv:1910.12574Syntology 0 ran · 1 unverified (of 1 harvested sample)
-
A Simple but Effective BERT Model for Dialog State Tracking on Resource-Limited Systems 28 Oct 2019 · 0 repositories · arXiv:1910.12995
-
Sequence-to-sequence Automatic Speech Recognition with Word Embedding Regularization and Fused Decoding 28 Oct 2019 · 1 repository · arXiv:1910.12740
-
What does BERT Learn from Multiple-Choice Reading Comprehension Datasets? 28 Oct 2019 · 0 repositories · arXiv:1910.12391
-
Word-level Textual Adversarial Attacking as Combinatorial Optimization 27 Oct 2019 · 1 repository · arXiv:1910.12196Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Thieves on Sesame Street! Model Extraction of BERT-based APIs 27 Oct 2019 · 1 repository · arXiv:1910.12366
-
DENS: A Dataset for Multi-class Emotion Analysis 25 Oct 2019 · 0 repositories · arXiv:1910.11769
-
HUBERT Untangles BERT to Improve Transfer across NLP Tasks 25 Oct 2019 · 1 repository · arXiv:1910.12647
-
L2RS: A Learning-to-Rescore Mechanism for Automatic Speech Recognition 25 Oct 2019 · 0 repositories · arXiv:1910.11496
-
On the Cross-lingual Transferability of Monolingual Representations 25 Oct 2019 · 7 repositories · arXiv:1910.11856Syntology community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering 25 Oct 2019 · 0 repositories · arXiv:1910.11559
-
An Empirical Study of Efficient ASR Rescoring with Transformers 24 Oct 2019 · 0 repositories · arXiv:1910.11450
-
Combining Acoustics, Content and Interaction Features to Find Hot Spots in Meetings 24 Oct 2019 · 0 repositories · arXiv:1910.10869
-
Emergent Properties of Finetuned Language Representation Models 23 Oct 2019 · 0 repositories · arXiv:1910.10832
-
Hierarchical Transformers for Long Document Classification 23 Oct 2019 · 3 repositories · arXiv:1910.10781
-
Relation Module for Non-answerable Prediction on Question Answering 23 Oct 2019 · 0 repositories · arXiv:1910.10843
-
Speech-XLNet: Unsupervised Acoustic Model Pretraining For Self-Attention Networks 23 Oct 2019 · 0 repositories · arXiv:1910.10387
-
MRQA 2019 Shared Task: Evaluating Generalization in Reading Comprehension 22 Oct 2019 · 1 repository · arXiv:1910.09753Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MonaLog: a Lightweight System for Natural Language Inference Based on Monotonicity 19 Oct 2019 · 1 repository · arXiv:1910.08772
-
A Mutual Information Maximization Perspective of Language Representation Learning 18 Oct 2019 · 0 repositories · arXiv:1910.08350
-
Model Compression with Two-stage Multi-teacher Knowledge Distillation for Web Question Answering System 18 Oct 2019 · 0 repositories · arXiv:1910.08381
-
BIG MOOD: Relating Transformers to Explicit Commonsense Knowledge 17 Oct 2019 · 0 repositories · arXiv:1910.07713
-
Measuring semantic similarity of clinical trial outcomes using deep pre-trained language representations 17 Oct 2019 · 0 repositories
-
Universal Text Representation from BERT: An Empirical Study 17 Oct 2019 · 0 repositories · arXiv:1910.07973
-
An Exponential Learning Rate Schedule for Deep Learning 16 Oct 2019 · 0 repositories · arXiv:1910.07454
-
BERTRAM: Improved Word Embeddings Have Big Impact on Contextualized Model Performance 16 Oct 2019 · 1 repository · arXiv:1910.07181Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 3 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Bridging the Knowledge Gap: Enhancing Question Answering with World and Domain Knowledge 16 Oct 2019 · 0 repositories · arXiv:1910.07429
-
Evolution of transfer learning in natural language processing 16 Oct 2019 · 0 repositories · arXiv:1910.07370
-
Aligning Cross-Lingual Entities with Multi-Aspect Information 15 Oct 2019 · 1 repository · arXiv:1910.06575
-
Answering Complex Open-domain Questions Through Iterative Query Generation 15 Oct 2019 · 1 repository · arXiv:1910.07000Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Structured Pruning of a BERT-based Question Answering Model 14 Oct 2019 · 0 repositories · arXiv:1910.06360
-
Q8BERT: Quantized 8Bit BERT 14 Oct 2019 · 5 repositories · arXiv:1910.06188Syntology official (archive's flag): 5 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
STANCY: Stance Classification Based on Consistency Cues 14 Oct 2019 · 1 repository · arXiv:1910.06048
-
Whatcha lookin' at? DeepLIFTing BERT's Attention in Question Answering 14 Oct 2019 · 1 repository · arXiv:1910.06431
-
Progress Notes Classification and Keyword Extraction using Attention-based Deep Learning Models with BERT 13 Oct 2019 · 0 repositories · arXiv:1910.05786
-
vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations 12 Oct 2019 · 3 repositories · arXiv:1910.05453Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Automatic segmentation of texts into units of meaning for reading assistance 11 Oct 2019 · 0 repositories · arXiv:1910.05014
-
BiPaR: A Bilingual Parallel Dataset for Multilingual and Cross-lingual Reading Comprehension on Novels 11 Oct 2019 · 1 repository · arXiv:1910.05040
-
exBERT: A Visual Analysis Tool to Explore Learned Representations in Transformers Models 11 Oct 2019 · 1 repository · arXiv:1910.05276
-
Cross-lingual Alignment vs Joint Training: A Comparative Study and A Simple Unified Framework 10 Oct 2019 · 2 repositories · arXiv:1910.04708
-
FUSE: Multi-Faceted Set Expansion by Coherent Clustering of Skip-grams 10 Oct 2019 · 3 repositories · arXiv:1910.04345
-
Multi-label Categorization of Accounts of Sexism using a Neural Framework 10 Oct 2019 · 1 repository · arXiv:1910.04602
-
Multilingual Question Answering from Formatted Text applied to Conversational Agents 10 Oct 2019 · 0 repositories · arXiv:1910.04659
-
Structured Pruning of Large Language Models 10 Oct 2019 · 2 repositories · arXiv:1910.04732Syntology community repositories only · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 2 pointer-only (licence)
-
Visual Natural Language Query Auto-Completion for Estimating Instance Probabilities 10 Oct 2019 · 1 repository · arXiv:1910.04887
-
A Closer Look At Feature Space Data Augmentation For Few-Shot Intent Classification 9 Oct 2019 · 0 repositories · arXiv:1910.04176
-
Alternating Recurrent Dialog Model with Large-scale Pre-trained Language Models 9 Oct 2019 · 1 repository · arXiv:1910.03756
-
Ctrl-Z: Recovering from Instability in Reinforcement Learning 9 Oct 2019 · 0 repositories · arXiv:1910.03732
-
Is Multilingual BERT Fluent in Language Generation? 9 Oct 2019 · 1 repository · arXiv:1910.03806
-
ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks 8 Oct 2019 · 13 repositories · arXiv:1910.03151Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Knowledge Distillation from Internal Representations 8 Oct 2019 · 0 repositories · arXiv:1910.03723
-
SesameBERT: Attention for Anywhere 8 Oct 2019 · 0 repositories · arXiv:1910.03176
-
BERT for Evidence Retrieval and Claim Verification 7 Oct 2019 · 2 repositories · arXiv:1910.02655Syntology 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Deformable Kernels: Adapting Effective Receptive Fields for Object Deformation 7 Oct 2019 · 2 repositories · arXiv:1910.02940
-
Parallel Iterative Edit Models for Local Sequence Transduction 7 Oct 2019 · 1 repository · arXiv:1910.02893
-
Named Entity Recognition -- Is there a glass ceiling? 6 Oct 2019 · 1 repository · arXiv:1910.02403
-
Distilling BERT into Simple Neural Networks with Unlabeled Transfer Data 4 Oct 2019 · 0 repositories · arXiv:1910.01769
-
Fine-grained Sentiment Classification using BERT 4 Oct 2019 · 1 repository · arXiv:1910.03474Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
ZeRO: Memory Optimizations Toward Training Trillion Parameter Models 4 Oct 2019 · 10 repositories · arXiv:1910.02054Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Towards Understanding of Medical Randomized Controlled Trials by Conclusion Generation 3 Oct 2019 · 1 repository · arXiv:1910.01462
-
Cracking the Contextual Commonsense Code: Understanding Commonsense Reasoning Aptitude of Deep Contextual Representations 2 Oct 2019 · 0 repositories · arXiv:1910.01157
-
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter 2 Oct 2019 · 37 repositories · arXiv:1910.01108Syntology official (archive's flag): 1 ran · 21 ran (of which 5 constructed an object rather than computing a result; 13 with no instrument failure: 3 honoured, 1 violated, 9 with no contract checked; 8 where Syntology's instrument failed) · 6 unverified (of 27 harvested samples) · 2 pointer-only (licence)
-
Exploiting BERT for End-to-End Aspect-based Sentiment Analysis 2 Oct 2019 · 1 repository · arXiv:1910.00883
-
Linking artificial and human neural representations of language 2 Oct 2019 · 1 repository · arXiv:1910.01244Syntology 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
QuaRL: Quantization for Fast and Environmentally Sustainable Reinforcement Learning 2 Oct 2019 · 1 repository · arXiv:1910.01055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
BERT for Question Generation 1 Oct 2019 · 0 repositories
-
Specializing Word Embeddings (for Parsing) by Information Bottleneck 1 Oct 2019 · 1 repository · arXiv:1910.00163
-
TMLab: Generative Enhanced Model (GEM) for adversarial attacks 1 Oct 2019 · 0 repositories · arXiv:1910.00337
-
VAE-PGN based Abstractive Model in Multi-stage Architecture for Text Summarization 1 Oct 2019 · 0 repositories
-
End-to-End Resume Parsing and Finding Candidates for a Job Description using BERT 30 Sep 2019 · 0 repositories · arXiv:1910.03089
-
RandAugment: Practical automated data augmentation with a reduced search space 30 Sep 2019 · 19 repositories · arXiv:1909.13719Syntology 58 ran (of which 1 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 51 where Syntology's instrument failed) · 7 unverified (of 65 harvested samples) · 17 pointer-only (licence)
-
Fake news detection using Deep Learning 29 Sep 2019 · 2 repositories · arXiv:1910.03496
-
Global Sparse Momentum SGD for Pruning Very Deep Neural Networks 27 Sep 2019 · 4 repositories · arXiv:1909.12778Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
HateMonitors: Language Agnostic Abuse Detection in Social Media 27 Sep 2019 · 1 repository · arXiv:1909.12642
-
On the use of BERT for Neural Machine Translation 27 Sep 2019 · 0 repositories · arXiv:1909.12744
-
Reweighted Proximal Pruning for Large-Scale Language Representation 27 Sep 2019 · 0 repositories · arXiv:1909.12486
-
ALBERT: A Lite BERT for Self-supervised Learning of Language Representations 26 Sep 2019 · 48 repositories · arXiv:1909.11942Syntology official (archive's flag): 6 ran · 81 ran (of which 17 constructed an object rather than computing a result; 59 with no instrument failure: 4 honoured, 0 violated, 55 with no contract checked; 22 where Syntology's instrument failed) · 45 unverified (of 126 harvested samples) · 28 pointer-only (licence)
-
Aspect and Opinion Term Extraction for Hotel Reviews using Transfer Learning and Auxiliary Labels 26 Sep 2019 · 0 repositories · arXiv:1909.11879
-
Biomedical relation extraction with pre-trained language representations and minimal task-specific architecture 26 Sep 2019 · 0 repositories · arXiv:1909.12411
-
Improving Pre-Trained Multilingual Models with Vocabulary Expansion 26 Sep 2019 · 0 repositories · arXiv:1909.12440
-
Extremely Small BERT Models from Mixed-Vocabulary Training 25 Sep 2019 · 0 repositories · arXiv:1909.11687
-
Learning to Detect Opinion Snippet for Aspect-Based Sentiment Analysis 25 Sep 2019 · 0 repositories · arXiv:1909.11297
-
Mixout: Effective Regularization to Finetune Large-scale Pretrained Language Models 25 Sep 2019 · 2 repositories · arXiv:1909.11299
-
Technical report on Conversational Question Answering 24 Sep 2019 · 0 repositories · arXiv:1909.10772
-
Understanding Semantics from Speech Through Pre-training 24 Sep 2019 · 0 repositories · arXiv:1909.10924
-
TinyBERT: Distilling BERT for Natural Language Understanding 23 Sep 2019 · 10 repositories · arXiv:1909.10351Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Does BERT Make Any Sense? Interpretable Word Sense Disambiguation with Contextualized Embeddings 23 Sep 2019 · 1 repository · arXiv:1909.10430
-
Automatic Identification and Normalisation of Physical Measurements in Scientific Literature 23 Sep 2019 · 1 repository
-
Constrained Attractor Selection Using Deep Reinforcement Learning 23 Sep 2019 · 0 repositories · arXiv:1909.10500
-
Portuguese Named Entity Recognition using BERT-CRF 23 Sep 2019 · 1 repository · arXiv:1909.10649Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
BERT Meets Chinese Word Segmentation 20 Sep 2019 · 0 repositories · arXiv:1909.09292
-
Absum: Simple Regularization Method for Reducing Structural Sensitivity of Convolutional Neural Networks 19 Sep 2019 · 0 repositories · arXiv:1909.08830
-
AllenNLP Interpret: A Framework for Explaining Predictions of NLP Models 19 Sep 2019 · 1 repository · arXiv:1909.09251
-
How Additional Knowledge can Improve Natural Language Commonsense Question Answering? 19 Sep 2019 · 0 repositories · arXiv:1909.08855
-
Summary Level Training of Sentence Rewriting for Abstractive Summarization 19 Sep 2019 · 0 repositories · arXiv:1909.08752