Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 49
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 49 of 71: papers 4,801 to 4,900 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Data Curation and Quality Assurance for Machine Learning-based Cyber Intrusion Detection 20 May 2021 · 1 repository · arXiv:2105.10041
-
Intra-Document Cascading: Learning to Select Passages for Neural Document Ranking 20 May 2021 · 1 repository · arXiv:2105.09816
-
Do Models Learn the Directionality of Relations? A New Evaluation: Relation Direction Recognition 19 May 2021 · 1 repository · arXiv:2105.09045
-
Explainable Tsetlin Machine framework for fake news detection with credibility score assessment 19 May 2021 · 6 repositories · arXiv:2105.09114
-
Improving Adverse Drug Event Extraction with SpanBERT on Different Text Typologies 19 May 2021 · 1 repository · arXiv:2105.08882
-
Methods for Detoxification of Texts for the Russian Language 19 May 2021 · 3 repositories · arXiv:2105.09052
-
DRILL: Dynamic Representations for Imbalanced Lifelong Learning 18 May 2021 · 1 repository · arXiv:2105.08445Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Effective Attention Sheds Light On Interpretability 18 May 2021 · 1 repository · arXiv:2105.08855
-
Compressed Communication for Distributed Training: Adaptive Methods and System 17 May 2021 · 1 repository · arXiv:2105.07829
-
Pay Attention to MLPs 17 May 2021 · 20 repositories · arXiv:2105.08050Syntology 34 ran (of which 15 constructed an object rather than computing a result; 31 with no instrument failure: 1 honoured, 4 violated, 26 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 44 harvested samples) · 11 pointer-only (licence)
-
BdLAN:BERTdoc Label Attention Networks for Multi-label text classification 16 May 2021 · 0 repositories
-
Coming to its senses: Lessons learned from Approximating Retrofitted BERT representations for Word Sense information 16 May 2021 · 0 repositories
-
Fine-Tuned Transformers Show Clusters of Similar Representations Across Layers 16 May 2021 · 0 repositories
-
How is BERT surprised? Layerwise detection of linguistic anomalies 16 May 2021 · 1 repository · arXiv:2105.07452
-
SINA-BERT: A Pre-Trained Language Model for Analysis of Medical Texts in Persian 16 May 2021 · 0 repositories
-
Subtopic Clustering with a Query-Specific Siamese Similarity Metric 16 May 2021 · 0 repositories
-
DirectQE: Direct Pretraining for Machine Translation Quality Estimation 15 May 2021 · 0 repositories · arXiv:2105.07149
-
Lexicon Enhanced Chinese Sequence Labeling Using BERT Adapter 15 May 2021 · 1 repository · arXiv:2105.07148
-
The Low-Dimensional Linear Geometry of Contextualized Word Representations 15 May 2021 · 0 repositories · arXiv:2105.07109
-
BERT Busters: Outlier Dimensions that Disrupt Transformers 14 May 2021 · 0 repositories · arXiv:2105.06990
-
Counterfactual Interventions Reveal the Causal Effect of Relative Clause Representations on Agreement Prediction 14 May 2021 · 0 repositories · arXiv:2105.06965
-
DaLAJ - a dataset for linguistic acceptability judgments for Swedish: Format, baseline, sharing 14 May 2021 · 0 repositories · arXiv:2105.06681
-
Distilling BERT for low complexity network training 13 May 2021 · 0 repositories · arXiv:2105.06514
-
BertGCN: Transductive Text Classification by Combining GCN and BERT 12 May 2021 · 1 repository · arXiv:2105.05727
-
Better than BERT but Worse than Baseline 12 May 2021 · 0 repositories · arXiv:2105.05915
-
Building a Question and Answer System for News Domain 12 May 2021 · 0 repositories · arXiv:2105.05744
-
Evaluating Gender Bias in Natural Language Inference 12 May 2021 · 1 repository · arXiv:2105.05541
-
Go Beyond Plain Fine-tuning: Improving Pretrained Models for Social Commonsense 12 May 2021 · 0 repositories · arXiv:2105.05913
-
Kleister: Key Information Extraction Datasets Involving Long Documents with Complex Layouts 12 May 2021 · 0 repositories · arXiv:2105.05796
-
MATE-KD: Masked Adversarial TExt, a Companion to Knowledge Distillation 12 May 2021 · 1 repository · arXiv:2105.05912
-
OCHADAI-KYOTO at SemEval-2021 Task 1: Enhancing Model Generalization and Robustness for Lexical Complexity Prediction 12 May 2021 · 0 repositories · arXiv:2105.05535
-
Playing Codenames with Language Graphs and Word Embeddings 12 May 2021 · 1 repository · arXiv:2105.05885
-
Priberam at MESINESP Multi-label Classification of Medical Texts Task 12 May 2021 · 1 repository · arXiv:2105.05614
-
Priberam Labs at the NTCIR-15 SHINRA2020-ML: Classification Task 12 May 2021 · 0 repositories · arXiv:2105.05605
-
UIUC_BioNLP at SemEval-2021 Task 11: A Cascade of Neural Models for Structuring Scholarly NLP Contributions 12 May 2021 · 1 repository · arXiv:2105.05435
-
Addressing "Documentation Debt" in Machine Learning Research: A Retrospective Datasheet for BookCorpus 11 May 2021 · 1 repository · arXiv:2105.05241Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
BERT is to NLP what AlexNet is to CV: Can Pre-Trained Language Models Identify Analogies? 11 May 2021 · 1 repository · arXiv:2105.04949
-
Integrating extracted information from bert and multiple embedding methods with the deep neural network for humour detection 11 May 2021 · 0 repositories · arXiv:2105.05112
-
Role of Artificial Intelligence in Detection of Hateful Speech for Hinglish Data on Social Media 11 May 2021 · 0 repositories · arXiv:2105.04913
-
Assessing the Syntactic Capabilities of Transformer-based Multilingual Language Models 10 May 2021 · 0 repositories · arXiv:2105.04688
-
Automatic Classification of Human Translation and Machine Translation: A Study from the Perspective of Lexical Diversity 10 May 2021 · 0 repositories · arXiv:2105.04616
-
SRLF: A Stance-aware Reinforcement Learning Framework for Content-based Rumor Detection on Social Media 10 May 2021 · 0 repositories · arXiv:2105.04098
-
FNet: Mixing Tokens with Fourier Transforms 9 May 2021 · 12 repositories · arXiv:2105.03824Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Improving Named Entity Recognition by External Context Retrieving and Cooperative Learning 8 May 2021 · 3 repositories · arXiv:2105.03654Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
NLP-IIS@UT at SemEval-2021 Task 4: Machine Reading Comprehension using the Long Document Transformer 8 May 2021 · 0 repositories · arXiv:2105.03775
-
Adapting by Pruning: A Case Study on BERT 7 May 2021 · 1 repository · arXiv:2105.03343
-
Empirical Evaluation of Pre-trained Transformers for Human-Level NLP: The Role of Sample Size and Dimensionality 7 May 2021 · 1 repository · arXiv:2105.03484
-
Understanding by Understanding Not: Modeling Negation in Language Models 7 May 2021 · 1 repository · arXiv:2105.03519
-
Adapting Monolingual Models: Data can be Scarce when Language Similarity is High 6 May 2021 · 1 repository · arXiv:2105.02855
-
Aligning Subtitles in Sign Language Videos 6 May 2021 · 0 repositories · arXiv:2105.02877
-
Bird's Eye: Probing for Linguistic Graph Structures with a Simple Information-Theoretic Approach 6 May 2021 · 1 repository · arXiv:2105.02629
-
Introducing Information Retrieval for Biomedical Informatics Students 6 May 2021 · 1 repository · arXiv:2105.02746
-
TABBIE: Pretrained Representations of Tabular Data 6 May 2021 · 2 repositories · arXiv:2105.02584
-
Goldilocks: Just-Right Tuning of BERT for Technology-Assisted Review 3 May 2021 · 0 repositories · arXiv:2105.01044
-
SmoothI: Smooth Rank Indicators for Differentiable IR Metrics 3 May 2021 · 1 repository · arXiv:2105.00942
-
Unreasonable Effectiveness of Rule-Based Heuristics in Solving Russian SuperGLUE Tasks 3 May 2021 · 0 repositories · arXiv:2105.01192
-
MathBERT: A Pre-Trained Model for Mathematical Formula Understanding 2 May 2021 · 0 repositories · arXiv:2105.00377
-
MRCBert: A Machine Reading ComprehensionApproach for Unsupervised Summarization 1 May 2021 · 1 repository · arXiv:2105.00239
-
BERT Meets Relational DB: Contextual Representations of Relational Databases 30 Apr 2021 · 0 repositories · arXiv:2104.14914
-
Word Sense Disambiguation with Transformer Models 30 Apr 2021 · 0 repositories
-
Let's Play Mono-Poly: BERT Can Reveal Words' Polysemy Level and Partitionability into Senses 29 Apr 2021 · 1 repository · arXiv:2104.14694
-
Improving BERT Model Using Contrastive Learning for Biomedical Relation Extraction 28 Apr 2021 · 1 repository · arXiv:2104.13913
-
MelBERT: Metaphor Detection via Contextualized Late Interaction using Metaphorical Identification Theories 28 Apr 2021 · 1 repository · arXiv:2104.13615Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Multi-Task Learning of Query Intent and Named Entities using Transfer Learning 28 Apr 2021 · 0 repositories · arXiv:2105.03316
-
Societal Biases in Retrieved Contents: Measurement Framework and Adversarial Mitigation for BERT Rankers 28 Apr 2021 · 1 repository · arXiv:2104.13640
-
Multi-class Text Classification using BERT-based Active Learning 27 Apr 2021 · 0 repositories · arXiv:2104.14289
-
Semi-supervised Interactive Intent Labeling 27 Apr 2021 · 0 repositories · arXiv:2104.13406
-
UoT-UWF-PartAI at SemEval-2021 Task 5: Self Attention Based Bi-GRU with Multi-Embedding Representation for Toxicity Highlighter 27 Apr 2021 · 0 repositories · arXiv:2104.13164
-
Diverse Image Inpainting with Bidirectional and Autoregressive Transformers 26 Apr 2021 · 0 repositories · arXiv:2104.12335
-
MDETR -- Modulated Detection for End-to-End Multi-Modal Understanding 26 Apr 2021 · 5 repositories · arXiv:2104.12763Syntology community repositories only · 7 ran (of which 4 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
Phrase break prediction with bidirectional encoder representations in Japanese text-to-speech synthesis 26 Apr 2021 · 1 repository · arXiv:2104.12395
-
Potential Idiomatic Expression (PIE)-English: Corpus for Classes of Idioms 25 Apr 2021 · 2 repositories · arXiv:2105.03280
-
Extract then Distill: Efficient and Effective Task-Agnostic BERT Distillation 24 Apr 2021 · 0 repositories · arXiv:2104.11928
-
Learning Passage Impacts for Inverted Indexes 24 Apr 2021 · 1 repository · arXiv:2104.12016Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Analysing Cyberbullying using Natural Language Processing by Understanding Jargon in Social Media 23 Apr 2021 · 0 repositories · arXiv:2107.08902
-
BERT-CoQAC: BERT-based Conversational Question Answering in Context 23 Apr 2021 · 0 repositories · arXiv:2104.11394
-
Comparative Analysis of Machine Learning and Deep Learning Algorithms for Detection of Online Hate Speech 23 Apr 2021 · 0 repositories · arXiv:2108.01063
-
Multimodal Fusion with BERT and Attention Mechanism for Fake News Detection 23 Apr 2021 · 1 repository · arXiv:2104.11476
-
Optimizing small BERTs trained for German NER 23 Apr 2021 · 2 repositories · arXiv:2104.11559
-
Towards Trustworthy Deception Detection: Benchmarking Model Robustness across Domains, Modalities, and Languages 23 Apr 2021 · 0 repositories · arXiv:2104.11761
-
On Geodesic Distances and Contextual Embedding Compression for Text Classification 22 Apr 2021 · 1 repository · arXiv:2104.11295
-
Discriminative Self-training for Punctuation Prediction 21 Apr 2021 · 0 repositories · arXiv:2104.10339
-
Disfluency Detection with Unlabeled Data and Small BERT Models 21 Apr 2021 · 0 repositories · arXiv:2104.10769
-
B-PROP: Bootstrapped Pre-training with Representative Words Prediction for Ad-hoc Retrieval 20 Apr 2021 · 1 repository · arXiv:2104.09791Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples) · 3 pointer-only (licence)
-
Efficient pre-training objectives for Transformers 20 Apr 2021 · 0 repositories · arXiv:2104.09694
-
Measuring Shifts in Attitudes Towards COVID-19 Measures in Belgium Using Multilingual BERT 20 Apr 2021 · 1 repository · arXiv:2104.09947
-
Subsentence Extraction from Text Using Coverage-Based Deep Learning Language Models 20 Apr 2021 · 1 repository · arXiv:2104.09777
-
UIT-ISE-NLP at SemEval-2021 Task 5: Toxic Spans Detection with BiLSTM-CRF and ToxicBERT Comment Classification 20 Apr 2021 · 1 repository · arXiv:2104.10100
-
WASSA@IITK at WASSA 2021: Multi-task Learning and Transformer Finetuning for Emotion Classification and Empathy Prediction 20 Apr 2021 · 0 repositories · arXiv:2104.09827
-
BigGreen at SemEval-2021 Task 1: Lexical Complexity Prediction with Assembly Models 19 Apr 2021 · 1 repository · arXiv:2104.09040
-
ELECTRAMed: a new pre-trained language representation model for biomedical NLP 19 Apr 2021 · 2 repositories · arXiv:2104.09585
-
Modeling "Newsworthiness" for Lead-Generation Across Corpora 19 Apr 2021 · 0 repositories · arXiv:2104.09653
-
Neural Language Models with Distant Supervision to Identify Major Depressive Disorder from Clinical Notes 19 Apr 2021 · 0 repositories · arXiv:2104.09644
-
OCTIS: Comparing and Optimizing Topic models is Simple! 19 Apr 2021 · 1 repository
-
Operationalizing a National Digital Library: The Case for a Norwegian Transformer Model 19 Apr 2021 · 2 repositories · arXiv:2104.09617
-
Probing for Bridging Inference in Transformer Language Models 19 Apr 2021 · 1 repository · arXiv:2104.09400
-
Sentiment Classification in Swahili Language Using Multilingual BERT 19 Apr 2021 · 0 repositories · arXiv:2104.09006
-
TeamUNCC@LT-EDI-EACL2021: Hope Speech Detection using Transfer Learning with Transformers 19 Apr 2021 · 1 repository
-
CEAR: Cross-Entity Aware Reranker for Knowledge Base Completion 18 Apr 2021 · 0 repositories · arXiv:2104.08741
-
Dual-View Distilled BERT for Sentence Embedding 18 Apr 2021 · 0 repositories · arXiv:2104.08675