Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 45
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 45 of 71: papers 4,401 to 4,500 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Effectiveness of Deep Networks in NLP using BiDAF as an example architecture 31 Aug 2021 · 0 repositories · arXiv:2109.00074
-
Enjoy the Salience: Towards Better Transformer-based Faithful Explanations with Word Salience 31 Aug 2021 · 1 repository · arXiv:2108.13759
-
How Does Adversarial Fine-Tuning Benefit BERT? 31 Aug 2021 · 0 repositories · arXiv:2108.13602
-
Monolingual versus Multilingual BERTology for Vietnamese Extractive Multi-Document Summarization 31 Aug 2021 · 0 repositories · arXiv:2108.13741
-
Sense representations for Portuguese: experiments with sense embeddings and deep neural language models 31 Aug 2021 · 0 repositories · arXiv:2109.00025
-
Improving Query Representations for Dense Retrieval with Pseudo Relevance Feedback 30 Aug 2021 · 2 repositories · arXiv:2108.13454Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Knowledge Base Completion Meets Transfer Learning 30 Aug 2021 · 1 repository · arXiv:2108.13073Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Shatter: An Efficient Transformer Encoder with Single-Headed Self-Attention and Relative Sequence Partitioning 30 Aug 2021 · 0 repositories · arXiv:2108.13032
-
Analyzing and Mitigating Interference in Neural Architecture Search 29 Aug 2021 · 0 repositories · arXiv:2108.12821
-
NoiER: An Approach for Training more Reliable Fine-TunedDownstream Task Models 29 Aug 2021 · 0 repositories · arXiv:2110.02054
-
DKM: Differentiable K-Means Clustering Layer for Neural Network Compression 28 Aug 2021 · 0 repositories · arXiv:2108.12659
-
Automatic Text Evaluation through the Lens of Wasserstein Barycenters 27 Aug 2021 · 2 repositories · arXiv:2108.12463
-
Dealing with Typos for BERT-based Passage Retrieval and Ranking 27 Aug 2021 · 2 repositories · arXiv:2108.12139
-
Evaluating the Robustness of Neural Language Models to Input Perturbations 27 Aug 2021 · 1 repository · arXiv:2108.12237Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Query-Focused Extractive Summarisation for Finding Ideal Answers to Biomedical and COVID-19 Questions 27 Aug 2021 · 1 repository · arXiv:2108.12189
-
A Computational Approach to Measure Empathy and Theory-of-Mind from Written Texts 26 Aug 2021 · 1 repository · arXiv:2108.11810
-
A New Sentence Ordering Method Using BERT Pretrained Model 26 Aug 2021 · 0 repositories · arXiv:2108.11994
-
EmoBERTa: Speaker-Aware Emotion Recognition in Conversation with RoBERTa 26 Aug 2021 · 1 repository · arXiv:2108.12009
-
Rethinking Why Intermediate-Task Fine-Tuning Works 26 Aug 2021 · 1 repository · arXiv:2108.11696
-
SLIM: Explicit Slot-Intent Mapping with BERT for Joint Multi-Intent Detection and Slot Filling 26 Aug 2021 · 1 repository · arXiv:2108.11711
-
Multilingual Multi-Aspect Explainability Analyses on Machine Reading Comprehension Models 26 Aug 2021 · 1 repository · arXiv:2108.11574
-
Models In a Spelling Bee: Language Models Implicitly Learn the Character Composition of Tokens 25 Aug 2021 · 1 repository · arXiv:2108.11193Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
On Approximate Nearest Neighbour Selection for Multi-Stage Dense Retrieval 25 Aug 2021 · 1 repository · arXiv:2108.11480
-
Ontology-Enhanced Slot Filling 25 Aug 2021 · 0 repositories · arXiv:2108.11275
-
What do pre-trained code models know about code? 25 Aug 2021 · 1 repository · arXiv:2108.11308
-
sigmoidF1: A Smooth F1 Score Surrogate Loss for Multilabel Classification 24 Aug 2021 · 1 repository · arXiv:2108.10566
-
Towards Offensive Language Identification for Tamil Code-Mixed YouTube Comments and Posts 24 Aug 2021 · 1 repository · arXiv:2108.10939
-
Weakly Supervised Cross-platform Teenager Detection with Adversarial BERT 24 Aug 2021 · 0 repositories · arXiv:2108.10619
-
Deploying a BERT-based Query-Title Relevance Classifier in a Production System: a View from the Trenches 23 Aug 2021 · 0 repositories · arXiv:2108.10197
-
Query Embedding Pruning for Dense Retrieval 23 Aug 2021 · 1 repository · arXiv:2108.10341
-
Regularizing Transformers With Deep Probabilistic Layers 23 Aug 2021 · 0 repositories · arXiv:2108.10764
-
Sarcasm Detection in Twitter -- Performance Impact while using Data Augmentation: Word Embeddings 23 Aug 2021 · 1 repository · arXiv:2108.09924
-
Using Large Pre-Trained Models with Cross-Modal Attention for Multi-Modal Emotion Recognition 22 Aug 2021 · 0 repositories · arXiv:2108.09669
-
UzBERT: pretraining a BERT model for Uzbek 22 Aug 2021 · 0 repositories · arXiv:2108.09814
-
Extracting Radiological Findings With Normalized Anatomical Information Using a Span-Based BERT Relation Extraction Model 20 Aug 2021 · 0 repositories · arXiv:2108.09211
-
Frozen Pretrained Transformers for Neural Sign Language Translation 20 Aug 2021 · 1 repository
-
A Framework for Neural Topic Modeling of Text Corpora 19 Aug 2021 · 1 repository · arXiv:2108.08946
-
Detection of Illicit Drug Trafficking Events on Instagram: A Deep Multimodal Multilabel Learning Approach 19 Aug 2021 · 0 repositories · arXiv:2108.08920
-
Fast Passage Re-ranking with Contextualized Exact Term Matching and Efficient Passage Expansion 19 Aug 2021 · 1 repository · arXiv:2108.08513
-
Fine-Grained Element Identification in Complaint Text of Internet Fraud 19 Aug 2021 · 0 repositories · arXiv:2108.08676
-
How Hateful are Movies? A Study and Prediction on Movie Subtitles 19 Aug 2021 · 1 repository · arXiv:2108.10724
-
Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models 19 Aug 2021 · 2 repositories · arXiv:2108.08877
-
UNIQORN: Unified Question Answering over RDF Knowledge Graphs and Natural Language Text 19 Aug 2021 · 1 repository · arXiv:2108.08614
-
Contributions of Transformer Attention Heads in Multi- and Cross-lingual Tasks 18 Aug 2021 · 0 repositories · arXiv:2108.08375
-
Integrating Dialog History into End-to-End Spoken Language Understanding Systems 18 Aug 2021 · 0 repositories · arXiv:2108.08405
-
SIFN: A Sentiment-aware Interactive Fusion Network for Review-based Item Recommendation 18 Aug 2021 · 0 repositories · arXiv:2108.08022
-
TSI: an Ad Text Strength Indicator using Text-to-CTR and Semantic-Ad-Similarity 18 Aug 2021 · 0 repositories · arXiv:2108.08226
-
Response Ranking with Multi-types of Deep Interactive Representations in Retrieval-based Dialogues 17 Aug 2021 · 1 repository
-
Deep Natural Language Processing for LinkedIn Search 16 Aug 2021 · 0 repositories · arXiv:2108.13300
-
Misleading the Covid-19 vaccination discourse on Twitter: An exploratory study of infodemic around the pandemic 16 Aug 2021 · 1 repository · arXiv:2108.10735
-
On the Opportunities and Risks of Foundation Models 16 Aug 2021 · 2 repositories · arXiv:2108.07258
-
Maps Search Misspelling Detection Leveraging Domain-Augmented Contextual Representations 15 Aug 2021 · 0 repositories · arXiv:2108.06842
-
Towards Structured Dynamic Sparse Pre-Training of BERT 13 Aug 2021 · 0 repositories · arXiv:2108.06277
-
AMMUS : A Survey of Transformer-based Pretrained Models in Natural Language Processing 12 Aug 2021 · 1 repository · arXiv:2108.05542
-
Modeling Relevance Ranking under the Pre-training and Fine-tuning Paradigm 12 Aug 2021 · 0 repositories · arXiv:2108.05652
-
Multimodal analysis of the predictability of hand-gesture properties 12 Aug 2021 · 0 repositories · arXiv:2108.05762
-
Overview of the HASOC track at FIRE 2020: Hate Speech and Offensive Content Identification in Indo-European Languages 12 Aug 2021 · 0 repositories · arXiv:2108.05927
-
Medical-VLBERT: Medical Visual Language BERT for COVID-19 CT Report Generation With Alternate Learning 11 Aug 2021 · 0 repositories · arXiv:2108.05067
-
NoFake at CheckThat! 2021: Fake News Detection Using BERT 11 Aug 2021 · 0 repositories · arXiv:2108.05419
-
Variable-Length Music Score Infilling via XLNet and Musically Specialized Positional Encoding 11 Aug 2021 · 1 repository · arXiv:2108.05064
-
A Study of Social and Behavioral Determinants of Health in Lung Cancer Patients Using Transformers-based Natural Language Processing Models 10 Aug 2021 · 0 repositories · arXiv:2108.04949
-
BROS: A Pre-trained Language Model Focusing on Text and Layout for Better Key Information Extraction from Documents 10 Aug 2021 · 2 repositories · arXiv:2108.04539Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples)
-
SynCoBERT: Syntax-Guided Multi-Modal Contrastive Pre-Training for Code Representation 10 Aug 2021 · 0 repositories · arXiv:2108.04556
-
Embodied BERT: A Transformer Model for Embodied, Language-guided Visual Task Completion 10 Aug 2021 · 1 repository · arXiv:2108.04927
-
TrUMAn: Trope Understanding in Movies and Animations 10 Aug 2021 · 0 repositories · arXiv:2108.04542
-
DoSSIER@COLIEE 2021: Leveraging dense retrieval and summarization-based re-ranking for case law retrieval 9 Aug 2021 · 1 repository · arXiv:2108.03937
-
FiLMing Multimodal Sarcasm Detection with Attention 9 Aug 2021 · 1 repository · arXiv:2110.00416
-
Efficacy of BERT embeddings on predicting disaster from Twitter data 8 Aug 2021 · 1 repository · arXiv:2108.10698
-
Leveraging Commonsense Knowledge on Classifying False News and Determining Checkworthiness of Claims 8 Aug 2021 · 0 repositories · arXiv:2108.03731
-
Deriving Disinformation Insights from Geolocalized Twitter Callouts 6 Aug 2021 · 1 repository · arXiv:2108.03067
-
Offensive Language and Hate Speech Detection with Deep Learning and Transfer Learning 6 Aug 2021 · 0 repositories · arXiv:2108.03305
-
Adaptive Residue-wise Profile Fusion for Low Homologous Protein SecondaryStructure Prediction Using External Knowledge 5 Aug 2021 · 0 repositories · arXiv:2108.04176
-
Decoupled Transformer for Scalable Inference in Open-domain Question Answering 5 Aug 2021 · 0 repositories · arXiv:2108.02765
-
Robust Transfer Learning with Pretrained Language Models through Adapters 5 Aug 2021 · 0 repositories · arXiv:2108.02340
-
Sentiment Analysis on the News to Improve Mental Health 5 Aug 2021 · 0 repositories · arXiv:2108.07706
-
Curriculum learning for language modeling 4 Aug 2021 · 1 repository · arXiv:2108.02170Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
ExBERT: An External Knowledge Enhanced BERT for Natural Language Inference 3 Aug 2021 · 0 repositories · arXiv:2108.01589
-
HTTP2vec: Embedding of HTTP Requests for Detection of Anomalous Traffic 3 Aug 2021 · 0 repositories · arXiv:2108.01763
-
Large-Scale Differentially Private BERT 3 Aug 2021 · 0 repositories · arXiv:2108.01624
-
Changes in European Solidarity Before and During COVID-19: Evidence from a Large Crowd- and Expert-Annotated Twitter Dataset 2 Aug 2021 · 1 repository · arXiv:2108.01042
-
LICHEE: Improving Language Model Pre-training with Multi-grained Tokenization 2 Aug 2021 · 1 repository · arXiv:2108.00801
-
Relation Aware Semi-autoregressive Semantic Parsing for NL2SQL 2 Aug 2021 · 0 repositories · arXiv:2108.00804
-
Transfer Learning for Mining Feature Requests and Bug Reports from Tweets and App Store Reviews 2 Aug 2021 · 1 repository · arXiv:2108.00663
-
1213Li at SemEval-2021 Task 6: Detection of Propaganda with Multi-modal Attention and Pre-trained Models 1 Aug 2021 · 0 repositories
-
ADEPT: An Adjective-Dependent Plausibility Task 1 Aug 2021 · 0 repositories
-
AStarTwice at SemEval-2021 Task 5: Toxic Span Detection Using RoBERTa-CRF, Domain Specific Pre-Training and Self-Training 1 Aug 2021 · 0 repositories
-
AttesTable at SemEval-2021 Task 9: Extending Statement Verification with Tables for Unknown Class, and Semantic Evidence Finding 1 Aug 2021 · 1 repository
-
BennettNLP at SemEval-2021 Task 5: Toxic Spans Detection using Stacked Embedding Powered Toxic Entity Recognizer 1 Aug 2021 · 0 repositories
-
BERTAC: Enhancing Transformer-based Language Models with Adversarially Pretrained Convolutional Neural Networks 1 Aug 2021 · 1 repository
-
Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection Incremental 1 Aug 2021 · 0 repositories
-
Can Transformer Models Measure Coherence In Text: Re-Thinking the Shuffle Test 1 Aug 2021 · 0 repositories
-
CMTA: COVID-19 Misinformation Multilingual Analysis on Twitter 1 Aug 2021 · 0 repositories
-
Cross-lingual Evidence Improves Monolingual Fake News Detection 1 Aug 2021 · 1 repository
-
CSECU-DSG at SemEval-2021 Task 1: Fusion of Transformer Models for Lexical Complexity Prediction 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 5: Leveraging Ensemble of Sequence Tagging Models for Toxic Spans Detection 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 6: Orchestrating Multimodal Neural Architectures for Identifying Persuasion Techniques in Texts and Images 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 7: Detecting and Rating Humor and Offense Employing Transformers 1 Aug 2021 · 0 repositories
-
DeepBlueAI at SemEval-2021 Task 7: Detecting and Rating Humor and Offense with Stacking Diverse Language Model-Based Methods 1 Aug 2021 · 0 repositories
-
DLJUST at SemEval-2021 Task 7: Hahackathon: Linking Humor and Offense 1 Aug 2021 · 0 repositories
-
Early Detection of Sexual Predators in Chats 1 Aug 2021 · 1 repository