Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 61
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 61 of 71: papers 6,001 to 6,100 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Contextual and Non-Contextual Word Embeddings: an in-depth Linguistic Investigation 1 Jul 2020 · 0 repositories
-
CopyBERT: A Unified Approach to Question Generation with Self-Attention 1 Jul 2020 · 0 repositories
-
Cross-Lingual Disaster-related Multi-label Tweet Classification with Manifold Mixup 1 Jul 2020 · 1 repository
-
Detecting Sarcasm in Conversation Context Using Transformer-Based Models 1 Jul 2020 · 0 repositories
-
Evaluating the Utility of Model Configurations and Data Augmentation on Clinical Semantic Textual Similarity 1 Jul 2020 · 0 repositories
-
Exploring the Limits of Simple Learners in Knowledge Distillation for Document Classification with DocBERT 1 Jul 2020 · 0 repositories
-
Feature Projection for Improved Text Classification 1 Jul 2020 · 0 repositories
-
GAN-BERT: Generative Adversarial Learning for Robust Text Classification with a Bunch of Labeled Examples 1 Jul 2020 · 2 repositories
-
Getting the ##life out of living: How Adequate Are Word-Pieces for Modelling Complex Morphology? 1 Jul 2020 · 0 repositories
-
Go Figure! Multi-task transformer-based architecture for metaphor detection using idioms: ETS team in 2020 metaphor shared task 1 Jul 2020 · 0 repositories
-
Go Wide, Then Narrow: Efficient Training of Deep Thin Networks 1 Jul 2020 · 0 repositories · arXiv:2007.00811
-
How does BERT's attention change when you fine-tune? An analysis methodology and a case study in negation scope 1 Jul 2020 · 0 repositories
-
IlliniMet: Illinois System for Metaphor Detection with Contextual and Linguistic Information 1 Jul 2020 · 0 repositories
-
Improving Multimodal Named Entity Recognition via Entity Span Detection with Unified Multimodal Transformer 1 Jul 2020 · 1 repository
-
Information Retrieval and Extraction on COVID-19 Clinical Articles Using Graph Community Detection and Bio-BERT Embeddings 1 Jul 2020 · 0 repositories
-
Intermediate-Task Transfer Learning with Pretrained Language Models: When and Why Does It Work? 1 Jul 2020 · 0 repositories
-
Investigating the effect of auxiliary objectives for the automated grading of learner English speech transcriptions 1 Jul 2020 · 0 repositories
-
Item-based Collaborative Filtering with BERT 1 Jul 2020 · 0 repositories
-
Joint Training with Semantic Role Labeling for Better Generalization in Natural Language Inference 1 Jul 2020 · 0 repositories
-
K\opsala: Transition-Based Graph Parsing via Efficient Training and Effective Encoding 1 Jul 2020 · 0 repositories
-
Metaphor Detection Using Contextual Word Embeddings From Transformers 1 Jul 2020 · 0 repositories
-
Modelling Context and Syntactical Features for Aspect-based Sentiment Analysis 1 Jul 2020 · 1 repository
-
Neural Sarcasm Detection using Conversation Context 1 Jul 2020 · 0 repositories
-
Revisiting Higher-Order Dependency Parsers 1 Jul 2020 · 0 repositories
-
RobertNLP at the IWPT 2020 Shared Task: Surprisingly Simple Enhanced UD Parsing for English 1 Jul 2020 · 0 repositories
-
Roles and Utilization of Attention Heads in Transformer-based Neural Language Models 1 Jul 2020 · 1 repository
-
Sarcasm Identification and Detection in Conversion Context using BERT 1 Jul 2020 · 0 repositories
-
Self-supervised context-aware COVID-19 document exploration through atlas grounding 1 Jul 2020 · 1 repository
-
SentiTel: TABSA for Twitter reviews on Uganda Telecoms 1 Jul 2020 · 0 repositories
-
Should You Fine-Tune BERT for Automated Essay Scoring? 1 Jul 2020 · 0 repositories
-
tBERT: Topic Models and BERT Joining Forces for Semantic Similarity Detection 1 Jul 2020 · 1 repository
-
The HW-TSC Video Speech Translation System at IWSLT 2020 1 Jul 2020 · 0 repositories
-
Transformers on Sarcasm Detection with Context 1 Jul 2020 · 0 repositories
-
Transition-based Semantic Dependency Parsing with Pointer Networks 1 Jul 2020 · 0 repositories
-
Turku Enhanced Parser Pipeline: From Raw Text to Enhanced Graphs in the IWPT 2020 Shared Task 1 Jul 2020 · 0 repositories
-
Understanding Advertisements with BERT 1 Jul 2020 · 0 repositories
-
Unsupervised FAQ Retrieval with Question Generation and BERT 1 Jul 2020 · 0 repositories
-
Why is penguin more similar to polar bear than to sea gull? Analyzing conceptual knowledge in distributional models 1 Jul 2020 · 0 repositories
-
Would you Rather? A New Benchmark for Learning Machine Alignment with Cultural Values and Social Preferences 1 Jul 2020 · 0 repositories
-
Data Movement Is All You Need: A Case Study on Optimizing Transformers 30 Jun 2020 · 1 repository · arXiv:2007.00072Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
SE3M: A Model for Software Effort Estimation Using Pre-trained Embedding Models 30 Jun 2020 · 0 repositories · arXiv:2006.16831
-
Segmentation Approach for Coreference Resolution Task 30 Jun 2020 · 0 repositories · arXiv:2007.04301
-
Improving Sequence Tagging for Vietnamese Text Using Transformer-based Neural Models 29 Jun 2020 · 2 repositories · arXiv:2006.15994
-
Building Interpretable Interaction Trees for Deep NLP Models 29 Jun 2020 · 0 repositories · arXiv:2007.04298
-
Want to Identify, Extract and Normalize Adverse Drug Reactions in Tweets? Use RoBERTa 29 Jun 2020 · 0 repositories · arXiv:2006.16146
-
BOND: BERT-Assisted Open-Domain Named Entity Recognition with Distant Supervision 28 Jun 2020 · 1 repository · arXiv:2006.15509Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Rethinking Positional Encoding in Language Pre-training 28 Jun 2020 · 3 repositories · arXiv:2006.15595Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
FastSpec: Scalable Generation and Detection of Spectre Gadgets Using Neural Embeddings 25 Jun 2020 · 1 repository · arXiv:2006.14147
-
LSBert: A Simple Framework for Lexical Simplification 25 Jun 2020 · 1 repository · arXiv:2006.14939
-
Normalizing Text using Language Modelling based on Phonetics and String Similarity 25 Jun 2020 · 0 repositories · arXiv:2006.14116
-
Accelerated Large Batch Optimization of BERT Pretraining in 54 minutes 24 Jun 2020 · 1 repository · arXiv:2006.13484
-
Efficient Constituency Parsing by Pointing 24 Jun 2020 · 0 repositories · arXiv:2006.13557
-
ReCO: A Large Scale Chinese Reading Comprehension Dataset on Opinion 22 Jun 2020 · 1 repository · arXiv:2006.12146
-
Students Need More Attention: BERT-based AttentionModel for Small Data with Application to AutomaticPatient Message Triage 22 Jun 2020 · 1 repository · arXiv:2006.11991
-
MaxVA: Fast Adaptation of Step Sizes by Maximizing Observed Variance of Gradients 21 Jun 2020 · 1 repository · arXiv:2006.11918
-
Sarcasm Detection in Tweets with BERT and GloVe Embeddings 20 Jun 2020 · 0 repositories · arXiv:2006.11512
-
A Qualitative Evaluation of Language Models on Automatic Question-Answering for COVID-19 19 Jun 2020 · 1 repository · arXiv:2006.10964
-
New Vietnamese Corpus for Machine Reading Comprehension of Health News Articles 19 Jun 2020 · 0 repositories · arXiv:2006.11138
-
SqueezeBERT: What can computer vision teach NLP about efficient neural networks? 19 Jun 2020 · 6 repositories · arXiv:2006.11316Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Exploring the BERT Cross-Lingual Transferability: a Case Study in Reading Comprehension 17 Jun 2020 · 0 repositories
-
Tagging and parsing of multidomain collections 17 Jun 2020 · 1 repository
-
End-to-End Code Switching Language Models for Automatic Speech Recognition 16 Jun 2020 · 0 repositories · arXiv:2006.08870
-
Improving accuracy and speeding up Document Image Classification through parallel systems 16 Jun 2020 · 1 repository · arXiv:2006.09141
-
Memory-Efficient Pipeline-Parallel DNN Training 16 Jun 2020 · 1 repository · arXiv:2006.09503Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 10 harvested samples)
-
PERL: Pivot-based Domain Adaptation for Pre-trained Deep Contextualized Embedding Models 16 Jun 2020 · 1 repository · arXiv:2006.09075
-
Scalable Cross Lingual Pivots to Model Pronoun Gender for Translation 16 Jun 2020 · 0 repositories · arXiv:2006.08881
-
The SPPD System for Schema Guided Dialogue State Tracking Challenge 16 Jun 2020 · 0 repositories · arXiv:2006.09035
-
Cooking Is All About People: Comment Classification On Cookery Channels Using BERT and Classification Models (Malayalam-English Mix-Code) 15 Jun 2020 · 0 repositories · arXiv:2007.04249
-
Document Classification for COVID-19 Literature 15 Jun 2020 · 1 repository · arXiv:2006.13816
-
FinBERT: A Pretrained Language Model for Financial Communications 15 Jun 2020 · 2 repositories · arXiv:2006.08097Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
FinEst BERT and CroSloEngual BERT: less is more in multilingual models 14 Jun 2020 · 0 repositories · arXiv:2006.07890
-
Transferring Monolingual Model to Low-Resource Language: The Case of Tigrinya 13 Jun 2020 · 0 repositories · arXiv:2006.07698
-
A Monolingual Approach to Contextualized Word Embeddings for Mid-Resource Languages 11 Jun 2020 · 0 repositories · arXiv:2006.06202
-
MC-BERT: Efficient Language Pre-Training via a Meta Controller 10 Jun 2020 · 1 repository · arXiv:2006.05744Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Revisiting Few-sample BERT Fine-tuning 10 Jun 2020 · 1 repository · arXiv:2006.05987
-
On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong Baselines 8 Jun 2020 · 2 repositories · arXiv:2006.04884Syntology official (archive's flag): 8 ran · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 0 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 20 harvested samples) · 3 pointer-only (licence)
-
Medical Concept Normalization in User Generated Texts by Learning Target Concept Embeddings 7 Jun 2020 · 0 repositories · arXiv:2006.04014
-
Pre-training Polish Transformer-based Language Models at Scale 7 Jun 2020 · 1 repository · arXiv:2006.04229
-
Accelerating Natural Language Understanding in Task-Oriented Dialog 5 Jun 2020 · 1 repository · arXiv:2006.03701
-
DeBERTa: Decoding-enhanced BERT with Disentangled Attention 5 Jun 2020 · 14 repositories · arXiv:2006.03654Syntology official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
UDPipe at EvaLatin 2020: Contextualized Embeddings and Treebank Embeddings 5 Jun 2020 · 0 repositories · arXiv:2006.03687
-
The SOFC-Exp Corpus and Neural Approaches to Information Extraction in the Materials Science Domain 4 Jun 2020 · 1 repository · arXiv:2006.03039
-
Automatic Text Summarization of COVID-19 Medical Research Articles using BERT and GPT-2 3 Jun 2020 · 1 repository · arXiv:2006.01997
-
A Pairwise Probe for Understanding BERT Fine-Tuning on Machine Reading Comprehension 2 Jun 2020 · 0 repositories · arXiv:2006.01346
-
BERT Based Multilingual Machine Comprehension in English and Hindi 2 Jun 2020 · 2 repositories · arXiv:2006.01432
-
Exploring Cross-sentence Contexts for Named Entity Recognition with BERT 2 Jun 2020 · 1 repository · arXiv:2006.01563Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Position Masking for Language Models 2 Jun 2020 · 0 repositories · arXiv:2006.05676
-
Question Answering on Scholarly Knowledge Graphs 2 Jun 2020 · 0 repositories · arXiv:2006.01527
-
WikiBERT models: deep transfer learning for many languages 2 Jun 2020 · 0 repositories · arXiv:2006.01538
-
An Effective Contextual Language Modeling Framework for Speech Summarization with Augmented Features 1 Jun 2020 · 0 repositories · arXiv:2006.01189
-
BERT-based Ensembles for Modeling Disclosure and Support in Conversational Social Media Text 1 Jun 2020 · 0 repositories · arXiv:2006.01222
-
Conversational Machine Comprehension: a Literature Review 1 Jun 2020 · 0 repositories · arXiv:2006.00671
-
Emergence of Separable Manifolds in Deep Language Representations 1 Jun 2020 · 1 repository · arXiv:2006.01095
-
Étude des variations sémantiques à travers plusieurs dimensions (Studying semantic variations through several dimensions ) 1 Jun 2020 · 0 repositories
-
Introduction d'informations sémantiques dans un système de reconnaissance de la parole (Despite spectacular advances in recent years, the Automatic Speech Recognition (ASR) systems still make mistakes, especially in noisy environments) 1 Jun 2020 · 0 repositories
-
Les modèles de langue contextuels Camembert pour le français : impact de la taille et de l'hétérogénéité des données d'entrainement (C AMEM BERT Contextual Language Models for French: Impact of Training Data Size and Heterogeneity ) 1 Jun 2020 · 0 repositories
-
Qu'apporte BERT à l'analyse syntaxique en constituants discontinus ? Une suite de tests pour évaluer les prédictions de structures syntaxiques discontinues en anglais (What does BERT contribute to discontinuous constituency parsing ? A test suite to evaluate discontinuous constituency structure predictions in English) 1 Jun 2020 · 0 repositories
-
Ré-entraîner ou entraîner soi-même ? Stratégies de pré-entraînement de BERT en domaine médical (Re-train or train from scratch ? Pre-training strategies for BERT in the medical domain ) 1 Jun 2020 · 0 repositories
-
Amnesic Probing: Behavioral Explanation with Amnesic Counterfactuals 1 Jun 2020 · 0 repositories · arXiv:2006.00995
-
BPGC at SemEval-2020 Task 11: Propaganda Detection in News Articles with Multi-Granularity Knowledge Sharing and Linguistic Features based Ensemble Learning 31 May 2020 · 0 repositories · arXiv:2006.00593