Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 39
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 39 of 71: papers 3,801 to 3,900 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Memory-Efficient Backpropagation through Large Linear Layers 31 Jan 2022 · 2 repositories · arXiv:2201.13195
-
Electra: Conditional Generative Model based Predicate-Aware Query Approximation 28 Jan 2022 · 0 repositories · arXiv:2201.12420
-
Clinical-Longformer and Clinical-BigBird: Transformers for long clinical sequences 27 Jan 2022 · 1 repository · arXiv:2201.11838
-
Going Extreme: Comparative Analysis of Hate Speech in Parler and Gab 27 Jan 2022 · 1 repository · arXiv:2201.11770
-
Grad2Task: Improved Few-shot Text Classification Using Gradients for Task Representation 27 Jan 2022 · 1 repository · arXiv:2201.11576
-
DiscoScore: Evaluating Text Generation with BERT and Discourse Coherence 26 Jan 2022 · 1 repository · arXiv:2201.11176
-
FiNCAT: Financial Numeral Claim Analysis Tool 26 Jan 2022 · 1 repository · arXiv:2202.00631
-
Neural Grapheme-to-Phoneme Conversion with Pre-trained Grapheme Models 26 Jan 2022 · 1 repository · arXiv:2201.10716
-
Self-supervised 3D Semantic Representation Learning for Vision-and-Language Navigation 26 Jan 2022 · 0 repositories · arXiv:2201.10788
-
BERTHA: Video Captioning Evaluation Via Transfer-Learned Human Assessment 25 Jan 2022 · 1 repository · arXiv:2201.10243
-
Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models 25 Jan 2022 · 0 repositories · arXiv:2201.10103
-
Emotion-based Modeling of Mental Disorders on Social Media 24 Jan 2022 · 0 repositories · arXiv:2201.09451
-
Polyphone disambiguation and accent prediction using pre-trained language models in Japanese TTS front-end 24 Jan 2022 · 0 repositories · arXiv:2201.09427
-
Unified Multimodal Punctuation Restoration Framework for Mixed-Modality Corpus 24 Jan 2022 · 1 repository · arXiv:2202.00468
-
A Large and Diverse Arabic Corpus for Language Modeling 23 Jan 2022 · 0 repositories · arXiv:2201.09227
-
An Application of Pseudo-Log-Likelihoods to Natural Language Scoring 23 Jan 2022 · 0 repositories · arXiv:2201.09377
-
Black-box Prompt Learning for Pre-trained Language Models 21 Jan 2022 · 1 repository · arXiv:2201.08531Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Dual Contrastive Learning: Text Classification via Label-Aware Data Augmentation 21 Jan 2022 · 2 repositories · arXiv:2201.08702
-
Less is Less: When Are Snippets Insufficient for Human vs Machine Relevance Estimation? 21 Jan 2022 · 0 repositories · arXiv:2201.08721
-
Cheating Automatic Short Answer Grading: On the Adversarial Usage of Adjectives and Adverbs 20 Jan 2022 · 1 repository · arXiv:2201.08318
-
Sentiment Analysis: Predicting Yelp Scores 20 Jan 2022 · 0 repositories · arXiv:2201.07999
-
Transfer Learning Approaches for Building Cross-Language Dense Retrieval Models 20 Jan 2022 · 1 repository · arXiv:2201.08471Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Near-Optimal Sparse Allreduce for Distributed Deep Learning 19 Jan 2022 · 1 repository · arXiv:2201.07598
-
TourBERT: A pretrained language model for the tourism industry 19 Jan 2022 · 0 repositories · arXiv:2201.07449
-
Hierarchical Neural Network Approaches for Long Document Classification 18 Jan 2022 · 0 repositories · arXiv:2201.06774
-
BERT vs ALBERT explained 17 Jan 2022 · 0 repositories
-
MuLVE, A Multi-Language Vocabulary Evaluation Data Set 17 Jan 2022 · 0 repositories · arXiv:2201.06286
-
Unintended Bias in Language Model-driven Conversational Recommendation 17 Jan 2022 · 0 repositories · arXiv:2201.06224
-
A Balanced Data Approach for Evaluating Cross-Lingual Transfer: Mapping the Linguistic Blood Bank 16 Jan 2022 · 0 repositories
-
A Multi-Granularity Opinion Summarization Method 16 Jan 2022 · 0 repositories
-
A Study of the Attention Abnormality in Trojaned BERTs 16 Jan 2022 · 0 repositories
-
An Exploitation of Heterogeneous Graph Neural Network for Extractive Long Document Summarization 16 Jan 2022 · 0 repositories
-
Applying SoftTriple Loss for Supervised Language Model Fine Tuning 16 Jan 2022 · 0 repositories
-
AutoAttention: Automatic Attention Head Selection Through Differentiable Pruning 16 Jan 2022 · 0 repositories
-
Bridge the Gap Between CV and NLP! A Gradient-based Textual Adversarial Attack Framework 16 Jan 2022 · 0 repositories
-
Can BERT Conduct Logical Reasoning? On the Difficulty of Learning to Reason from Data 16 Jan 2022 · 0 repositories
-
Context-Aware Prompt: Customize A Unique Prompt For Each Input 16 Jan 2022 · 0 repositories
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 16 Jan 2022 · 0 repositories
-
Divide and Conquer: Text Semantic Matching with Disentangled Keywords and Intents 16 Jan 2022 · 0 repositories
-
Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs 16 Jan 2022 · 0 repositories
-
EiCi: A New Method of Dynamic Embedding Incorporating Contextual Information in Chinese NER 16 Jan 2022 · 0 repositories
-
Event Detection via Derangement Reading Comprehension 16 Jan 2022 · 0 repositories
-
Experiments with adversarial attacks on text genres 16 Jan 2022 · 0 repositories
-
Feasibility of BERT Embeddings For Domain-Specific Knowledge Mining 16 Jan 2022 · 0 repositories
-
FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks 16 Jan 2022 · 0 repositories
-
Global Entity Disambiguation with BERT 16 Jan 2022 · 0 repositories
-
Identifying the Source of Vulnerability in Fragile Interpretations: A Case Study in Neural Text Classification 16 Jan 2022 · 0 repositories
-
IMPLI: Investigatng NLI Models' Performance on Figurative Language 16 Jan 2022 · 0 repositories
-
Improving Contextual Representation with Gloss Regularized Pre-training 16 Jan 2022 · 0 repositories
-
Investigating and Explaining Feature and Representation Learning in Translationese Classification 16 Jan 2022 · 0 repositories
-
Investigating the saliency of sentiment expressions in aspect-based sentiment analysis 16 Jan 2022 · 0 repositories
-
Language Models for Code-switch Detection of te reo Māori and English in a Low-resource Setting 16 Jan 2022 · 0 repositories
-
Magic Pyramid: Accelerating Inference with Early Exiting and Token Pruning 16 Jan 2022 · 0 repositories
-
Measuring Word-Context Biases in Lexical Semantic Datasets 16 Jan 2022 · 0 repositories
-
Minimally-Supervised Relation Induction from Pre-trained Language Model 16 Jan 2022 · 0 repositories
-
Modeling Multi-Granularity Hierarchical Features for Relation Extraction 16 Jan 2022 · 0 repositories
-
Multi-Stage Pre-Training for Math-Understanding: μ²(AL)BERT 16 Jan 2022 · 0 repositories
-
PCEE-BERT: Accelerating BERT Inference via Patient and Confident Early Exiting 16 Jan 2022 · 0 repositories
-
Probing The Linguistic Capacity of Pre-Trained Vision-Language Models 16 Jan 2022 · 0 repositories
-
Progressive Class Semantic Matching for Semi-supervised Text Classification 16 Jan 2022 · 0 repositories
-
Re2G: Retrieve, Rerank, Generate 16 Jan 2022 · 0 repositories
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 16 Jan 2022 · 0 repositories
-
Revisiting Additive Compositionality: AND, OR, and NOT Operations with Word Embeddings 16 Jan 2022 · 0 repositories
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 16 Jan 2022 · 0 repositories
-
Roof-BERT: Divide Understanding Labour and Join in Work 16 Jan 2022 · 0 repositories
-
SemAttack: Natural Textual Attacks via Different Semantic Spaces 16 Jan 2022 · 0 repositories
-
Seq-GAN-BERT:Sequence Generative Adversarial Learning for Low-resource Name Entity Recognition 16 Jan 2022 · 0 repositories
-
Simple Local Attentions Remain Competitive for Long-Context Tasks 16 Jan 2022 · 0 repositories
-
TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning 16 Jan 2022 · 0 repositories
-
Tapping BERT for Preposition Sense Disambiguation 16 Jan 2022 · 0 repositories
-
TEMPLATE: TempRel Classification Model Trained with Embedded Temporal Relation Knowledge 16 Jan 2022 · 0 repositories
-
That is a good looking car !: Visual Aspect based Sentiment Controlled Personalized Response Generation 16 Jan 2022 · 0 repositories
-
Tree Knowledge Distillation for Compressing Transformer-Based Language Models 16 Jan 2022 · 0 repositories
-
Uncovering Surprising Event Boundaries in Narratives 16 Jan 2022 · 0 repositories
-
Understand before Answer: Improve Temporal Reading Comprehension via Precise Question Understanding 16 Jan 2022 · 0 repositories
-
VEE-BERT: Accelerating BERT Inference for Named Entity Recognition via Vote Early Exiting 16 Jan 2022 · 0 repositories
-
What do tokens know about their characters and how do they know it? 16 Jan 2022 · 1 repository
-
What Role Does BERT Play in the Neural Machine Translation Encoder? 16 Jan 2022 · 0 repositories
-
Automatic Correction of Syntactic Dependency Annotation Differences 15 Jan 2022 · 0 repositories · arXiv:2201.05891
-
Automatic Lexical Simplification for Turkish 15 Jan 2022 · 0 repositories · arXiv:2201.05878
-
Machine Learning for Food Review and Recommendation 15 Jan 2022 · 0 repositories · arXiv:2201.10978
-
Polarity and Subjectivity Detection with Multitask Learning and BERT Embedding 14 Jan 2022 · 0 repositories · arXiv:2201.05363
-
Knowledge Graph Augmented Network Towards Multiview Representation Learning for Aspect-based Sentiment Analysis 13 Jan 2022 · 1 repository · arXiv:2201.04831
-
Multi-task Pre-training Language Model for Semantic Network Completion 13 Jan 2022 · 1 repository · arXiv:2201.04843
-
Towards Automated Error Analysis: Learning to Characterize Errors 13 Jan 2022 · 0 repositories · arXiv:2201.05017
-
Diagnosing BERT with Retrieval Heuristics 12 Jan 2022 · 1 repository · arXiv:2201.04458
-
Generative Adversarial Network for Text-to-Face Synthesis and Manipulation with Pretrained BERT Model 12 Jan 2022 · 0 repositories
-
PromptBERT: Improving BERT Sentence Embeddings with Prompts 12 Jan 2022 · 1 repository · arXiv:2201.04337Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; the one sample that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
A Feature Extraction based Model for Hate Speech Identification 11 Jan 2022 · 0 repositories · arXiv:2201.04227
-
Explaining Predictive Uncertainty by Looking Back at Model Explanations 11 Jan 2022 · 0 repositories · arXiv:2201.03742
-
Quantifying Robustness to Adversarial Word Substitutions 11 Jan 2022 · 0 repositories · arXiv:2201.03829
-
BERT for Sentiment Analysis: Pre-trained and Fine-Tuned Alternatives 10 Jan 2022 · 2 repositories · arXiv:2201.03382
-
Black-Box Tuning for Language-Model-as-a-Service 10 Jan 2022 · 2 repositories · arXiv:2201.03514Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Handwriting recognition and automatic scoring for descriptive answers in Japanese language tests 10 Jan 2022 · 0 repositories · arXiv:2201.03215
-
SCROLLS: Standardized CompaRison Over Long Language Sequences 10 Jan 2022 · 2 repositories · arXiv:2201.03533Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Latency Adjustable Transformer Encoder for Language Understanding 10 Jan 2022 · 0 repositories · arXiv:2201.03327
-
Self-Training Vision Language BERTs with a Unified Conditional Model 6 Jan 2022 · 0 repositories · arXiv:2201.02010
-
Formal Analysis of Art: Proxy Learning of Visual Concepts from Style Through Language Models 5 Jan 2022 · 0 repositories · arXiv:2201.01819
-
Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction 5 Jan 2022 · 2 repositories · arXiv:2201.02184
-
Comparison of biomedical relationship extraction methods and models for knowledge graph creation 5 Jan 2022 · 0 repositories · arXiv:2201.01647