Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 65
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 65 of 71: papers 6,401 to 6,500 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
X-Stance: A Multilingual Multi-Target Dataset for Stance Detection 18 Mar 2020 · 1 repository · arXiv:2003.08385
-
Author2Vec: A Framework for Generating User Embedding 17 Mar 2020 · 0 repositories · arXiv:2003.11627
-
Calibration of Pre-trained Transformers 17 Mar 2020 · 1 repository · arXiv:2003.07892
-
PO-EMO: Conceptualization, Annotation, and Modeling of Aesthetic Emotions in German and English Poetry 17 Mar 2020 · 1 repository · arXiv:2003.07723
-
A Survey on Contextual Embeddings 16 Mar 2020 · 0 repositories · arXiv:2003.07278
-
Cost-Sensitive BERT for Generalisable Sentence Classification with Imbalanced Data 16 Mar 2020 · 1 repository · arXiv:2003.11563
-
TRANS-BLSTM: Transformer with Bidirectional LSTM for Language Understanding 16 Mar 2020 · 0 repositories · arXiv:2003.07000
-
Document Ranking with a Pretrained Sequence-to-Sequence Model 14 Mar 2020 · 2 repositories · arXiv:2003.06713
-
Finnish Language Modeling with Deep Transformer Models 14 Mar 2020 · 0 repositories · arXiv:2003.11562
-
Hurtful Words: Quantifying Biases in Clinical Contextual Word Embeddings 11 Mar 2020 · 1 repository · arXiv:2003.11515
-
Investigating Entity Knowledge in BERT with Simple Neural End-To-End Entity Linking 11 Mar 2020 · 1 repository · arXiv:2003.05473Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Keyword-Attentive Deep Semantic Matching 11 Mar 2020 · 1 repository · arXiv:2003.11516
-
Efficient Intent Detection with Dual Sentence Encoders 10 Mar 2020 · 5 repositories · arXiv:2003.04807Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Sensitive Data Detection and Classification in Spanish Clinical Text: Experiments with BERT 6 Mar 2020 · 0 repositories · arXiv:2003.03106
-
Transfer Learning for Information Extraction with Limited Data 6 Mar 2020 · 0 repositories · arXiv:2003.03064
-
BERT as a Teacher: Contextual Embeddings for Sequence-Level Reward 5 Mar 2020 · 0 repositories · arXiv:2003.02738
-
HypoNLI: Exploring the Artificial Patterns of Hypothesis-only Bias in Natural Language Inference 5 Mar 2020 · 0 repositories · arXiv:2003.02756
-
What the [MASK]? Making Sense of Language-Specific BERT Models 5 Mar 2020 · 0 repositories · arXiv:2003.02912
-
A Study on Efficiency, Accuracy and Document Structure for Answer Sentence Selection 4 Mar 2020 · 0 repositories · arXiv:2003.02349
-
Data Augmentation using Pre-trained Transformer Models 4 Mar 2020 · 4 repositories · arXiv:2003.02245Syntology community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
jiant: A Software Toolkit for Research on General-Purpose Text Understanding Models 4 Mar 2020 · 6 repositories · arXiv:2003.02249Syntology official (archive's flag): 6 ran · 6 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Kleister: A novel task for Information Extraction involving Long Documents with Complex Layout 4 Mar 2020 · 0 repositories · arXiv:2003.02356
-
CLUECorpus2020: A Large-scale Chinese Corpus for Pre-training Language Model 3 Mar 2020 · 2 repositories · arXiv:2003.01355
-
Hierarchical Context Enhanced Multi-Domain Dialogue System for Multi-domain Task Completion 3 Mar 2020 · 0 repositories · arXiv:2003.01338
-
AraBERT: Transformer-based Model for Arabic Language Understanding 28 Feb 2020 · 4 repositories · arXiv:2003.00104
-
DC-BERT: Decoupling Question and Document for Efficient Contextual Encoding 28 Feb 2020 · 0 repositories · arXiv:2002.12591
-
TextBrewer: An Open-Source Knowledge Distillation Toolkit for Natural Language Processing 28 Feb 2020 · 1 repository · arXiv:2002.12620
-
A Primer in BERTology: What we know about how BERT works 27 Feb 2020 · 0 repositories · arXiv:2002.12327
-
Adv-BERT: BERT is not robust on misspellings! Generating nature adversarial samples on BERT 27 Feb 2020 · 0 repositories · arXiv:2003.04985
-
Compressing Large-Scale Transformer-Based Models: A Case Study on BERT 27 Feb 2020 · 0 repositories · arXiv:2002.11985
-
MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers 25 Feb 2020 · 1 repository · arXiv:2002.10957
-
What BERT Sees: Cross-Modal Transfer for Visual Question Generation 25 Feb 2020 · 0 repositories · arXiv:2002.10832
-
Exploring BERT Parameter Efficiency on the Stanford Question Answering Dataset v2.0 25 Feb 2020 · 0 repositories · arXiv:2002.10670
-
Improving BERT Fine-Tuning via Self-Ensemble and Self-Distillation 24 Feb 2020 · 1 repository · arXiv:2002.10345Syntology 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Predicting Subjective Features of Questions of QA Websites using BERT 24 Feb 2020 · 5 repositories · arXiv:2002.10107
-
Federated pretraining and fine tuning of BERT using clinical notes from multiple silos 20 Feb 2020 · 0 repositories · arXiv:2002.08562
-
Compressing BERT: Studying the Effects of Weight Pruning on Transfer Learning 19 Feb 2020 · 1 repository · arXiv:2002.08307
-
The Microsoft Toolkit of Multi-Task Deep Neural Networks for Natural Language Understanding 19 Feb 2020 · 3 repositories · arXiv:2002.07972
-
From English To Foreign Languages: Transferring Pre-trained Language Models 18 Feb 2020 · 1 repository · arXiv:2002.07306Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A Financial Service Chatbot based on Deep Bidirectional Transformers 17 Feb 2020 · 0 repositories · arXiv:2003.04987
-
Incorporating BERT into Neural Machine Translation 17 Feb 2020 · 3 repositories · arXiv:2002.06823
-
SBERT-WK: A Sentence Embedding Method by Dissecting BERT-based Word Models 16 Feb 2020 · 3 repositories · arXiv:2002.06652
-
The Utility of General Domain Transfer Learning for Medical Language Tasks 16 Feb 2020 · 0 repositories · arXiv:2002.06670
-
Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping 15 Feb 2020 · 4 repositories · arXiv:2002.06305
-
UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation 15 Feb 2020 · 2 repositories · arXiv:2002.06353Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
FQuAD: French Question Answering Dataset 14 Feb 2020 · 0 repositories · arXiv:2002.06071
-
Stress Test Evaluation of Transformer-based Models in Natural Language Understanding Tasks 14 Feb 2020 · 0 repositories · arXiv:2002.06261
-
Transformer on a Diet 14 Feb 2020 · 1 repository · arXiv:2002.06170
-
TwinBERT: Distilling Knowledge to Twin-Structured BERT Models for Efficient Retrieval 14 Feb 2020 · 2 repositories · arXiv:2002.06275
-
Understanding patient complaint characteristics using contextual clinical BERT embeddings 14 Feb 2020 · 0 repositories · arXiv:2002.05902
-
A Simple Framework for Contrastive Learning of Visual Representations 13 Feb 2020 · 96 repositories · arXiv:2002.05709Syntology official (archive's flag): 2 ran · 115 ran (of which 33 constructed an object rather than computing a result; 90 with no instrument failure: 1 honoured, 1 violated, 88 with no contract checked; 25 where Syntology's instrument failed) · 22 unverified (of 137 harvested samples) · 52 pointer-only (licence)
-
Training Large Neural Networks with Constant Memory using a New Execution Algorithm 13 Feb 2020 · 2 repositories · arXiv:2002.05645Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Learning to Compare for Better Training and Evaluation of Open Domain Natural Language Generation Models 12 Feb 2020 · 0 repositories · arXiv:2002.05058
-
Utilizing BERT Intermediate Layers for Aspect Based Sentiment Analysis and Natural Language Inference 12 Feb 2020 · 1 repository · arXiv:2002.04815
-
Multilingual Alignment of Contextual Word Representations 10 Feb 2020 · 1 repository · arXiv:2002.03518Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Application of Pre-training Models in Named Entity Recognition 9 Feb 2020 · 0 repositories · arXiv:2002.08902
-
Momentum Improves Normalized SGD 9 Feb 2020 · 0 repositories · arXiv:2002.03305
-
BERT-of-Theseus: Compressing BERT by Progressive Module Replacing 7 Feb 2020 · 2 repositories · arXiv:2002.02925Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
K-Adapter: Infusing Knowledge into Pre-Trained Models with Adapters 5 Feb 2020 · 2 repositories · arXiv:2002.01808Syntology 11 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Rapid Adaptation of BERT for Information Extraction on Domain-Specific Business Documents 5 Feb 2020 · 1 repository · arXiv:2002.01861
-
Interpretable & Time-Budget-Constrained Contextualization for Re-Ranking 4 Feb 2020 · 1 repository · arXiv:2002.01854
-
Bertrand-DR: Improving Text-to-SQL using a Discriminative Re-ranker 3 Feb 2020 · 1 repository · arXiv:2002.00557
-
Beat the AI: Investigating Adversarial Human Annotation for Reading Comprehension 2 Feb 2020 · 1 repository · arXiv:2002.00293
-
Fine-Tuning BERT for Schema-Guided Zero-Shot Dialogue State Tracking 1 Feb 2020 · 0 repositories · arXiv:2002.00181
-
Pretrained Transformers for Simple Question Answering over Knowledge Graphs 31 Jan 2020 · 1 repository · arXiv:2001.11985
-
Adversarial Training for Aspect-Based Sentiment Analysis with BERT 30 Jan 2020 · 4 repositories · arXiv:2001.11316
-
On the Importance of Word Order Information in Cross-lingual Sequence Labeling 30 Jan 2020 · 0 repositories · arXiv:2001.11164
-
PEL-BERT: A Joint Model for Protocol Entity Linking 28 Jan 2020 · 0 repositories · arXiv:2002.00744
-
BERT's output layer recognizes all hidden layers? Some Intriguing Phenomena and a simple way to boost BERT 25 Jan 2020 · 0 repositories · arXiv:2001.09309
-
Generation-Distillation for Efficient Natural Language Understanding in Low-Data Settings 25 Jan 2020 · 0 repositories · arXiv:2002.00733
-
PoWER-BERT: Accelerating BERT Inference via Progressive Word-vector Elimination 24 Jan 2020 · 1 repository · arXiv:2001.08950Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Navigation-Based Candidate Expansion and Pretrained Language Models for Citation Recommendation 23 Jan 2020 · 0 repositories · arXiv:2001.08687
-
A multimodal deep learning approach for named entity recognition from social media 19 Jan 2020 · 0 repositories · arXiv:2001.06888
-
Deep Learning for Hindi Text Classification: A Comparison 19 Jan 2020 · 0 repositories · arXiv:2001.10340
-
Capturing Evolution in Word Usage: Just Add More Clusters? 18 Jan 2020 · 0 repositories · arXiv:2001.06629
-
RobBERT: a Dutch RoBERTa-based Language Model 17 Jan 2020 · 1 repository · arXiv:2001.06286Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Schema2QA: High-Quality and Low-Cost Q&A Agents for the Structured Web 16 Jan 2020 · 3 repositories · arXiv:2001.05609
-
FGN: Fusion Glyph Network for Chinese Named Entity Recognition 15 Jan 2020 · 1 repository · arXiv:2001.05272
-
A BERT based Sentiment Analysis and Key Entity Detection Approach for Online Financial Texts 14 Jan 2020 · 0 repositories · arXiv:2001.05326
-
AdaBERT: Task-Adaptive BERT Compression with Differentiable Neural Architecture Search 13 Jan 2020 · 1 repository · arXiv:2001.04246Syntology 0 ran · 3 unverified (of 3 harvested samples)
-
Représentations lexicales pour la détection non supervisée d'événements dans un flux de tweets : étude sur des corpus français et anglais 13 Jan 2020 · 1 repository · arXiv:2001.04139
-
Exploring and Improving Robustness of Multi Task Deep Neural Networks via Domain Agnostic Defenses 11 Jan 2020 · 1 repository · arXiv:2001.05286
-
Resolving the Scope of Speculation and Negation using Transformer-Based Architectures 9 Jan 2020 · 1 repository · arXiv:2001.02885
-
To Transfer or Not to Transfer: Misclassification Attacks Against Transfer Learned Text Classifiers 8 Jan 2020 · 0 repositories · arXiv:2001.02438
-
Improving Entity Linking by Modeling Latent Entity Type Information 6 Jan 2020 · 0 repositories · arXiv:2001.01447
-
Multi-Layer Content Interaction Through Quaternion Product For Visual Question Answering 3 Jan 2020 · 0 repositories · arXiv:2001.05840
-
BERT-AL: BERT for Arbitrarily Long Document Understanding 1 Jan 2020 · 0 repositories
-
Stacked DeBERT: All Attention in Incomplete Data for Text Classification 1 Jan 2020 · 1 repository · arXiv:2001.00137Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
oLMpics -- On what Language Model Pre-training Captures 31 Dec 2019 · 2 repositories · arXiv:1912.13283
-
AutoDiscern: Rating the Quality of Online Health Information with Hierarchical Encoder Attention-based Neural Networks 30 Dec 2019 · 1 repository · arXiv:1912.12999
-
Clinical XLNet: Modeling Sequential Clinical Notes and Predicting Prolonged Mechanical Ventilation 27 Dec 2019 · 3 repositories · arXiv:1912.11975
-
Harnessing Evolution of Multi-Turn Conversations for Effective Answer Retrieval 22 Dec 2019 · 1 repository · arXiv:1912.10554
-
Learning and Evaluating Contextual Embedding of Source Code 21 Dec 2019 · 2 repositories · arXiv:2001.00059
-
Pretrained Encyclopedia: Weakly Supervised Knowledge-Pretrained Language Model 20 Dec 2019 · 0 repositories · arXiv:1912.09637
-
Shareable Representations for Search Query Understanding 20 Dec 2019 · 0 repositories · arXiv:2001.04345
-
BERTje: A Dutch BERT Model 19 Dec 2019 · 2 repositories · arXiv:1912.09582Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
CJRC: A Reliable Human-Annotated Benchmark DataSet for Chinese Judicial Reading Comprehension 19 Dec 2019 · 0 repositories · arXiv:1912.09156
-
Neural Simile Recognition with Cyclic Multitask Learning and Local Attention 19 Dec 2019 · 1 repository · arXiv:1912.09084
-
A Multi-task Learning Model for Chinese-oriented Aspect Polarity Classification and Aspect Term Extraction 17 Dec 2019 · 6 repositories · arXiv:1912.07976
-
Cross-Lingual Ability of Multilingual BERT: An Empirical Study 17 Dec 2019 · 0 repositories · arXiv:1912.07840