Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 57
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 57 of 71: papers 5,601 to 5,700 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Accelerating Training of Transformer-Based Language Models with Progressive Layer Dropping 26 Oct 2020 · 1 repository · arXiv:2010.13369
-
Fine-grained Information Status Classification Using Discourse Context-Aware BERT 26 Oct 2020 · 1 repository · arXiv:2010.14759
-
Semi-Supervised Spoken Language Understanding via Self-Supervised Speech and Language Model Pretraining 26 Oct 2020 · 1 repository · arXiv:2010.13826
-
UPB at SemEval-2020 Task 12: Multilingual Offensive Language Detection on Social Media by Fine-tuning a Variety of BERT-based Models 26 Oct 2020 · 0 repositories · arXiv:2010.13609
-
Commonsense knowledge adversarial dataset that challenges ELECTRA 25 Oct 2020 · 0 repositories · arXiv:2010.13049
-
Contextualized Word Embeddings Encode Aspects of Human-Like Word Sense Knowledge 25 Oct 2020 · 0 repositories · arXiv:2010.13057
-
CRAB: Class Representation Attentive BERT for Hate Speech Identification in Social Media 25 Oct 2020 · 0 repositories · arXiv:2010.13028
-
Two-stage Textual Knowledge Distillation for End-to-End Spoken Language Understanding 25 Oct 2020 · 1 repository · arXiv:2010.13105
-
Char2Subword: Extending the Subword Embedding Space Using Robust Character Compositionality 24 Oct 2020 · 0 repositories · arXiv:2010.12730
-
COUGH: A Challenge Dataset and Models for COVID-19 FAQ Retrieval 24 Oct 2020 · 1 repository · arXiv:2010.12800
-
Jointly Optimizing State Operation Prediction and Value Generation for Dialogue State Tracking 24 Oct 2020 · 2 repositories · arXiv:2010.14061
-
Pre-trained Summarization Distillation 24 Oct 2020 · 1 repository · arXiv:2010.13002
-
BARThez: a Skilled Pretrained French Sequence-to-Sequence Model 23 Oct 2020 · 5 repositories · arXiv:2010.12321
-
Did You Ask a Good Question? A Cross-Domain Question Intention Classification Benchmark for Text-to-SQL 23 Oct 2020 · 1 repository · arXiv:2010.12634
-
ERNIE-Gram: Pre-Training with Explicitly N-Gram Masked Language Modeling for Natural Language Understanding 23 Oct 2020 · 2 repositories · arXiv:2010.12148
-
GiBERT: Introducing Linguistic Knowledge into BERT through a Lightweight Gated Injection Method 23 Oct 2020 · 0 repositories · arXiv:2010.12532
-
HateBERT: Retraining BERT for Abusive Language Detection in English 23 Oct 2020 · 1 repository · arXiv:2010.12472
-
LightSeq: A High Performance Inference Library for Transformers 23 Oct 2020 · 1 repository · arXiv:2010.13887
-
Long Document Ranking with Query-Directed Sparse Transformer 23 Oct 2020 · 1 repository · arXiv:2010.12683
-
On the Transformer Growth for Progressive BERT Training 23 Oct 2020 · 0 repositories · arXiv:2010.12562
-
Posterior Differential Regularization with f-divergence for Improving Model Robustness 23 Oct 2020 · 2 repositories · arXiv:2010.12638
-
Pre-training with Meta Learning for Chinese Word Segmentation 23 Oct 2020 · 0 repositories · arXiv:2010.12272
-
ST-BERT: Cross-modal Language Model Pre-training For End-to-end Spoken Language Understanding 23 Oct 2020 · 0 repositories · arXiv:2010.12283
-
Topic Modeling with Contextualized Word Representation Clusters 23 Oct 2020 · 0 repositories · arXiv:2010.12626
-
Distilling Dense Representations for Ranking using Tightly-Coupled Teachers 22 Oct 2020 · 2 repositories · arXiv:2010.11386
-
Improving BERT Performance for Aspect-Based Sentiment Analysis 22 Oct 2020 · 2 repositories · arXiv:2010.11731Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Exploiting News Article Structure for Automatic Corpus Generation of Entailment Datasets 22 Oct 2020 · 1 repository · arXiv:2010.11574
-
Knowledge Distillation for BERT Unsupervised Domain Adaptation 22 Oct 2020 · 1 repository · arXiv:2010.11478
-
Language Models are Open Knowledge Graphs 22 Oct 2020 · 2 repositories · arXiv:2010.11967Syntology 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Self-Alignment Pretraining for Biomedical Entity Representations 22 Oct 2020 · 1 repository · arXiv:2010.11784
-
Towards Fully Bilingual Deep Language Modeling 22 Oct 2020 · 0 repositories · arXiv:2010.11639
-
UniCase -- Rethinking Casing in Language Models 22 Oct 2020 · 0 repositories · arXiv:2010.11936
-
Detection of COVID-19 informative tweets using RoBERTa 21 Oct 2020 · 0 repositories · arXiv:2010.11238
-
A Simple and Efficient Multi-Task Learning Approach for Conditioned Dialogue Generation 21 Oct 2020 · 1 repository · arXiv:2010.11140
-
German's Next Language Model 21 Oct 2020 · 1 repository · arXiv:2010.10906
-
Latte-Mix: Measuring Sentence Semantic Similarity with Latent Categorical Mixtures 21 Oct 2020 · 0 repositories · arXiv:2010.11351
-
AutoMeTS: The Autocomplete for Medical Text Simplification 20 Oct 2020 · 1 repository · arXiv:2010.10573
-
BERT2DNN: BERT Distillation with Massive Unlabeled Data for Online E-Commerce Search 20 Oct 2020 · 0 repositories · arXiv:2010.10442
-
CharacterBERT: Reconciling ELMo and BERT for Word-Level Open-Vocabulary Representations From Characters 20 Oct 2020 · 2 repositories · arXiv:2010.10392Syntology community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 4 pointer-only (licence)
-
ConjNLI: Natural Language Inference Over Conjunctive Sentences 20 Oct 2020 · 1 repository · arXiv:2010.10418Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
CoRT: Complementary Rankings from Transformers 20 Oct 2020 · 1 repository · arXiv:2010.10252
-
Looking for Clues of Language in Multilingual BERT to Improve Cross-lingual Generalization 20 Oct 2020 · 1 repository · arXiv:2010.10041
-
Optimal Subarchitecture Extraction For BERT 20 Oct 2020 · 3 repositories · arXiv:2010.10499
-
Performance of Transfer Learning Model vs. Traditional Neural Network in Low System Resource Environment 20 Oct 2020 · 0 repositories · arXiv:2011.07962
-
PROP: Pre-training with Representative Words Prediction for Ad-hoc Retrieval 20 Oct 2020 · 1 repository · arXiv:2010.10137
-
Text Classification of Manifestos and COVID-19 Press Briefings using BERT and Convolutional Neural Networks 20 Oct 2020 · 0 repositories · arXiv:2010.10267
-
What makes multilingual BERT multilingual? 20 Oct 2020 · 0 repositories · arXiv:2010.10938
-
BERTnesia: Investigating the capture and forgetting of knowledge in BERT 19 Oct 2020 · 1 repository · arXiv:2010.09313
-
Better Distractions: Transformer-based Distractor Generation and Multiple Choice Question Filtering 19 Oct 2020 · 0 repositories · arXiv:2010.09598
-
Cold-start Active Learning through Self-supervised Language Modeling 19 Oct 2020 · 1 repository · arXiv:2010.09535Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ColloQL: Robust Cross-Domain Text-to-SQL Over Search Queries 19 Oct 2020 · 1 repository · arXiv:2010.09927
-
Cross-Lingual Transfer in Zero-Shot Cross-Language Entity Linking 19 Oct 2020 · 1 repository · arXiv:2010.09828
-
Drug Repurposing for COVID-19 via Knowledge Graph Completion 19 Oct 2020 · 1 repository · arXiv:2010.09600
-
The RELX Dataset and Matching the Multilingual Blanks for Cross-Lingual Relation Classification 19 Oct 2020 · 1 repository · arXiv:2010.09381Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Explaining and Improving Model Behavior with k Nearest Neighbor Representations 18 Oct 2020 · 0 repositories · arXiv:2010.09030
-
Towards Interpreting BERT for Reading Comprehension Based QA 18 Oct 2020 · 1 repository · arXiv:2010.08983Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Answer-checking in Context: A Multi-modal FullyAttention Network for Visual Question Answering 17 Oct 2020 · 0 repositories · arXiv:2010.08708
-
HABERTOR: An Efficient and Effective Deep Hatespeech Detector 17 Oct 2020 · 0 repositories · arXiv:2010.08865
-
Hierarchical Multitask Learning Approach for BERT 17 Oct 2020 · 0 repositories · arXiv:2011.04451
-
Question Answering over Knowledge Base using Language Model Embeddings 17 Oct 2020 · 0 repositories · arXiv:2010.08883
-
TweetBERT: A Pretrained Language Representation Model for Twitter Text Analysis 17 Oct 2020 · 1 repository · arXiv:2010.11091
-
Coarse-to-Fine Pre-training for Named Entity Recognition 16 Oct 2020 · 1 repository · arXiv:2010.08210Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Delaying Interaction Layers in Transformer-based Encoders for Efficient Open Domain Question Answering 16 Oct 2020 · 1 repository · arXiv:2010.08422
-
It's not Greek to mBERT: Inducing Word-Level Translations from Multilingual BERT 16 Oct 2020 · 1 repository · arXiv:2010.08275Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Linguistically-Informed Transformations (LIT): A Method for Automatically Generating Contrast Sets 16 Oct 2020 · 0 repositories · arXiv:2010.08580
-
Context-Guided BERT for Targeted Aspect-Based Sentiment Analysis 15 Oct 2020 · 1 repository · arXiv:2010.07523
-
Does Chinese BERT Encode Word Structure? 15 Oct 2020 · 1 repository · arXiv:2010.07711
-
Neural Deepfake Detection with Factual Structure of Text 15 Oct 2020 · 1 repository · arXiv:2010.07475
-
NUIG-Shubhanker@Dravidian-CodeMix-FIRE2020: Sentiment Analysis of Code-Mixed Dravidian text using XLNet 15 Oct 2020 · 0 repositories · arXiv:2010.07773
-
Response Selection for Multi-Party Conversations withDynamic Topic Tracking 15 Oct 2020 · 0 repositories · arXiv:2010.07785
-
Unsupervised Bitext Mining and Translation via Self-trained Contextual Embeddings 15 Oct 2020 · 0 repositories · arXiv:2010.07761
-
An Investigation on Different Underlying Quantization Schemes for Pre-trained Language Models 14 Oct 2020 · 0 repositories · arXiv:2010.07109
-
DA-Transformer: Distance-aware Transformer 14 Oct 2020 · 0 repositories · arXiv:2010.06925
-
Geometry matters: Exploring language examples at the decision boundary 14 Oct 2020 · 0 repositories · arXiv:2010.07212
-
No Rumours Please! A Multi-Indic-Lingual Approach for COVID Fake-Tweet Detection 14 Oct 2020 · 1 repository · arXiv:2010.06906
-
Aspect-based Document Similarity for Research Papers 13 Oct 2020 · 1 repository · arXiv:2010.06395Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
BERT-EMD: Many-to-Many Layer Mapping for BERT Compression with Earth Mover's Distance 13 Oct 2020 · 1 repository · arXiv:2010.06133
-
CAPT: Contrastive Pre-Training for Learning Denoised Sequence Representations 13 Oct 2020 · 0 repositories · arXiv:2010.06351
-
Improving Text Generation Evaluation with Batch Centering and Tempered Word Mover Distance 13 Oct 2020 · 0 repositories · arXiv:2010.06150
-
Incorporating BERT into Parallel Sequence Decoding with Adapters 13 Oct 2020 · 1 repository · arXiv:2010.06138Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
Interpreting Attention Models with Human Visual Attention in Machine Reading Comprehension 13 Oct 2020 · 0 repositories · arXiv:2010.06396
-
Multilingual Argument Mining: Datasets and Analysis 13 Oct 2020 · 0 repositories · arXiv:2010.06432
-
Pretrained Transformers for Text Ranking: BERT and Beyond 13 Oct 2020 · 1 repository · arXiv:2010.06467
-
Probing for Multilingual Numerical Understanding in Transformer-Based Language Models 13 Oct 2020 · 1 repository · arXiv:2010.06666
-
Chatbot Interaction with Artificial Intelligence: Human Data Augmentation with T5 and Language Transformer Ensemble for Text Classification 12 Oct 2020 · 0 repositories · arXiv:2010.05990
-
Counterfactual Variable Control for Robust and Interpretable Question Answering 12 Oct 2020 · 1 repository · arXiv:2010.05581
-
Cross-Modal BERT for Text-Audio Sentiment Analysis 12 Oct 2020 · 1 repository
-
EFSG: Evolutionary Fooling Sentences Generator 12 Oct 2020 · 0 repositories · arXiv:2010.05736
-
From Hero to Zéroe: A Benchmark of Low-Level Adversarial Attacks 12 Oct 2020 · 1 repository · arXiv:2010.05648Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
HUJI-KU at MRP 2020: Two Transition-based Neural Parsers 12 Oct 2020 · 0 repositories · arXiv:2010.05710
-
Improving Compositional Generalization in Semantic Parsing 12 Oct 2020 · 1 repository · arXiv:2010.05647Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Layer-wise Guided Training for BERT: Learning Incrementally Refined Document Representations 12 Oct 2020 · 0 repositories · arXiv:2010.05763
-
Load What You Need: Smaller Versions of Multilingual BERT 12 Oct 2020 · 2 repositories · arXiv:2010.05609
-
Probing Pretrained Language Models for Lexical Semantics 12 Oct 2020 · 0 repositories · arXiv:2010.05731
-
Zero-shot Entity Linking with Efficient Long Range Sequence Modeling 12 Oct 2020 · 1 repository · arXiv:2010.06065
-
Connecting the Dots Between Fact Verification and Fake News Detection 11 Oct 2020 · 0 repositories · arXiv:2010.05202
-
Data Agnostic RoBERTa-based Natural Language to SQL Query Generation 11 Oct 2020 · 1 repository · arXiv:2010.05243
-
Detecting Foodborne Illness Complaints in Multiple Languages Using English Annotations Only 11 Oct 2020 · 0 repositories · arXiv:2010.05194
-
Incremental Processing in the Age of Non-Incremental Encoders: An Empirical Assessment of Bidirectional Models for Incremental NLU 11 Oct 2020 · 1 repository · arXiv:2010.05330
-
Learning Which Features Matter: RoBERTa Acquires a Preference for Linguistic Generalizations (Eventually) 11 Oct 2020 · 1 repository · arXiv:2010.05358