Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 34
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 34 of 71: papers 3,301 to 3,400 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
AIR-JPMC@SMM4H'22: Classifying Self-Reported Intimate Partner Violence in Tweets with Multiple BERT-based Models 22 Sep 2022 · 0 repositories · arXiv:2209.10763
-
Bias at a Second Glance: A Deep Dive into Bias for German Educational Peer-Review Data Modeling 21 Sep 2022 · 2 repositories · arXiv:2209.10335Syntology official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
CAE: Mechanism to Diminish the Class Imbalanced in SLU Slot Filling Task 21 Sep 2022 · 1 repository
-
Representing Affect Information in Word Embeddings 21 Sep 2022 · 0 repositories · arXiv:2209.10583
-
Subject Verb Agreement Error Patterns in Meaningless Sentences: Humans vs. BERT 21 Sep 2022 · 0 repositories · arXiv:2209.10538
-
Towards Fine-tuning Pre-trained Language Models with Integer Forward and Backward Propagation 20 Sep 2022 · 0 repositories · arXiv:2209.09815
-
One-to-Many Semantic Communication Systems: Design, Implementation, Performance Evaluation 20 Sep 2022 · 0 repositories · arXiv:2209.09425
-
Unsupervised Early Exit in DNNs with Multiple Exits 20 Sep 2022 · 1 repository · arXiv:2209.09480
-
Detecting Generated Scientific Papers using an Ensemble of Transformer Models 17 Sep 2022 · 1 repository · arXiv:2209.08283
-
CodeQueries: A Dataset of Semantic Queries over Code 17 Sep 2022 · 1 repository · arXiv:2209.08372
-
Changing the Representation: Examining Language Representation for Neural Sign Language Production 16 Sep 2022 · 0 repositories · arXiv:2210.06312
-
Machine Reading, Fast and Slow: When Do Models "Understand" Language? 15 Sep 2022 · 0 repositories · arXiv:2209.07430
-
uChecker: Masked Pretrained Language Models as Unsupervised Chinese Spelling Checkers 15 Sep 2022 · 0 repositories · arXiv:2209.07068
-
Automated Fidelity Assessment for Strategy Training in Inpatient Rehabilitation using Natural Language Processing 14 Sep 2022 · 0 repositories · arXiv:2209.06727
-
BERT-based Ensemble Approaches for Hate Speech Detection 14 Sep 2022 · 0 repositories · arXiv:2209.06505
-
Pre-training for Information Retrieval: Are Hyperlinks Fully Explored? 14 Sep 2022 · 0 repositories · arXiv:2209.06583
-
CNN-Trans-Enc: A CNN-Enhanced Transformer-Encoder On Top Of Static BERT representations for Document Classification 13 Sep 2022 · 0 repositories · arXiv:2209.06344
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 13 Sep 2022 · 0 repositories · arXiv:2209.05707
-
SkIn: Skimming-Intensive Long-Text Classification Using BERT for Medical Corpus 13 Sep 2022 · 0 repositories · arXiv:2209.05741
-
A new hazard event classification model via deep learning and multifractal 12 Sep 2022 · 0 repositories · arXiv:2209.05263
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 12 Sep 2022 · 0 repositories · arXiv:2209.05286
-
Probing for Understanding of English Verb Classes and Alternations in Large Pre-trained Language Models 11 Sep 2022 · 0 repositories · arXiv:2209.04811
-
Yes, DLGM! A novel hierarchical model for hazard classification 10 Sep 2022 · 0 repositories · arXiv:2209.04576
-
EchoCoTr: Estimation of the Left Ventricular Ejection Fraction from Spatiotemporal Echocardiography 9 Sep 2022 · 1 repository · arXiv:2209.04242
-
Trigger Warnings: Bootstrapping a Violence Detector for FanFiction 9 Sep 2022 · 0 repositories · arXiv:2209.04409
-
CLaCLab at SocialDisNER: Using Medical Gazetteers for Named-Entity Recognition of Disease Mentions in Spanish Tweets 8 Sep 2022 · 1 repository · arXiv:2209.03528
-
5q032e@SMM4H'22: Transformer-based classification of premise in tweets related to COVID-19 8 Sep 2022 · 0 repositories · arXiv:2209.03851
-
Multilingual Bidirectional Unsupervised Translation Through Multilingual Finetuning and Back-Translation 6 Sep 2022 · 1 repository · arXiv:2209.02821
-
Distilling the Knowledge of BERT for CTC-based ASR 5 Sep 2022 · 0 repositories · arXiv:2209.02030
-
Generalization in Neural Networks: A Broad Survey 4 Sep 2022 · 0 repositories · arXiv:2209.01610
-
GReS: Graphical Cross-domain Recommendation for Supply Chain Platform 2 Sep 2022 · 0 repositories · arXiv:2209.01031
-
Enhancing Semantic Understanding with Self-supervised Methods for Abstractive Dialogue Summarization 1 Sep 2022 · 0 repositories · arXiv:2209.00278
-
Isotropic Representation Can Improve Dense Retrieval 1 Sep 2022 · 1 repository · arXiv:2209.00218
-
Distilling Multi-Scale Knowledge for Event Temporal Relation Extraction 1 Sep 2022 · 0 repositories · arXiv:2209.00568
-
Negation detection in Dutch clinical texts: an evaluation of rule-based and machine learning methods 1 Sep 2022 · 1 repository · arXiv:2209.00470
-
Few-Shot Learning for Clinical Natural Language Processing Using Siamese Neural Networks 31 Aug 2022 · 0 repositories · arXiv:2208.14923
-
Large-scale Multi-granular Concept Extraction Based on Machine Reading Comprehension 30 Aug 2022 · 1 repository · arXiv:2208.14139
-
No means ‘No’; a non-im-proper modeling approach, with embedded speculative context 30 Aug 2022 · 0 repositories
-
SwiftPruner: Reinforced Evolutionary Pruning for Efficient Ad Relevance 30 Aug 2022 · 0 repositories · arXiv:2209.00625
-
Transformers with Learnable Activation Functions 30 Aug 2022 · 2 repositories · arXiv:2208.14111Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Multi-dimensional Racism Classification during COVID-19: Stigmatization, Offensiveness, Blame, and Exclusion 29 Aug 2022 · 0 repositories · arXiv:2208.13318
-
Building the Intent Landscape of Real-World Conversational Corpora with Extractive Question-Answering Transformers 26 Aug 2022 · 0 repositories · arXiv:2208.12886
-
Task-specific Pre-training and Prompt Decomposition for Knowledge Graph Population with Language Models 26 Aug 2022 · 1 repository · arXiv:2208.12539
-
Addressing Token Uniformity in Transformers via Singular Value Transformation 24 Aug 2022 · 1 repository · arXiv:2208.11790Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Evaluate Confidence Instead of Perplexity for Zero-shot Commonsense Reasoning 23 Aug 2022 · 0 repositories · arXiv:2208.11007
-
A Syntax Aware BERT for Identifying Well-Formed Queries in a Curriculum Framework 21 Aug 2022 · 0 repositories · arXiv:2208.09912
-
CMSBERT-CLR: Context-driven Modality Shifting BERT with Contrastive Learning for linguistic, visual, acoustic Representations 21 Aug 2022 · 0 repositories · arXiv:2209.07424
-
BSpell: A CNN-Blended BERT Based Bangla Spell Checker 20 Aug 2022 · 1 repository · arXiv:2208.09709
-
Combining Compressions for Multiplicative Size Scaling on Natural Language Tasks 20 Aug 2022 · 0 repositories · arXiv:2208.09684
-
Pretrained Language Encoders are Natural Tagging Frameworks for Aspect Sentiment Triplet Extraction 20 Aug 2022 · 0 repositories · arXiv:2208.09617
-
SPOT: Knowledge-Enhanced Language Representations for Information Extraction 20 Aug 2022 · 0 repositories · arXiv:2208.09625
-
Graph-Augmented Cyclic Learning Framework for Similarity Estimation of Medical Clinical Notes 19 Aug 2022 · 0 repositories · arXiv:2208.09437
-
UniCausal: Unified Benchmark and Repository for Causal Text Mining 19 Aug 2022 · 1 repository · arXiv:2208.09163
-
VAuLT: Augmenting the Vision-and-Language Transformer for Sentiment Classification on Social Media 18 Aug 2022 · 1 repository · arXiv:2208.09021Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
EmoMent: An Emotion Annotated Mental Health Corpus from two South Asian Countries 17 Aug 2022 · 0 repositories · arXiv:2208.08486
-
Transformer Encoder for Social Science 17 Aug 2022 · 1 repository · arXiv:2208.08005
-
Continuous Active Learning Using Pretrained Transformers 15 Aug 2022 · 0 repositories · arXiv:2208.06955
-
Text Difficulty Study: Do machines behave the same as humans regarding text difficulty? 14 Aug 2022 · 0 repositories · arXiv:2208.14509
-
Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models 13 Aug 2022 · 9 repositories · arXiv:2208.06677Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Interpreting BERT-based Text Similarity via Activation and Saliency Maps 13 Aug 2022 · 0 repositories · arXiv:2208.06612
-
Is Your Model Sensitive? SPeDaC: A New Benchmark for Detecting and Classifying Sensitive Personal Data 12 Aug 2022 · 0 repositories · arXiv:2208.06216
-
Pre-training Tasks for User Intent Detection and Embedding Retrieval in E-commerce Search 12 Aug 2022 · 1 repository · arXiv:2208.06150
-
A Model of Anaphoric Ambiguities using Sheaf Theoretic Quantum-like Contextuality and BERT 11 Aug 2022 · 0 repositories · arXiv:2208.05720
-
A Twitter-Driven Deep Learning Mechanism for the Determination of Vehicle Hijacking Spots in Cities 11 Aug 2022 · 0 repositories · arXiv:2208.10280
-
Searching for chromate replacements using natural language processing and machine learning algorithms 11 Aug 2022 · 0 repositories · arXiv:2208.05672
-
A Multimodal Transformer: Fusing Clinical Notes with Structured EHR Data for Interpretable In-Hospital Mortality Prediction 9 Aug 2022 · 0 repositories · arXiv:2208.10240
-
E2EG: End-to-End Node Classification Using Graph Topology and Text-based Node Attributes 9 Aug 2022 · 1 repository · arXiv:2208.04609Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 4 pointer-only (licence)
-
Emotion Detection From Tweets Using a BERT and SVM Ensemble Model 9 Aug 2022 · 1 repository · arXiv:2208.04547
-
Exploring Hate Speech Detection with HateXplain and BERT 9 Aug 2022 · 1 repository · arXiv:2208.04489
-
Efficient Fine-Tuning of Compressed Language Models with Learners 3 Aug 2022 · 0 repositories · arXiv:2208.02070
-
A Comparative Study on COVID-19 Fake News Detection Using Different Transformer Based Models 2 Aug 2022 · 0 repositories · arXiv:2208.01355
-
Automatic Classification of Bug Reports Based on Multiple Text Information and Reports' Intention 2 Aug 2022 · 0 repositories · arXiv:2208.01274
-
Debiasing Gender Bias in Information Retrieval Models 2 Aug 2022 · 0 repositories · arXiv:2208.01755
-
giMLPs: Gate with Inhibition Mechanism in MLPs 1 Aug 2022 · 1 repository · arXiv:2208.00929
-
Aggretriever: A Simple Approach to Aggregate Textual Representations for Robust Dense Passage Retrieval 31 Jul 2022 · 1 repository · arXiv:2208.00511
-
A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond 30 Jul 2022 · 0 repositories · arXiv:2208.00173
-
Code Comment Inconsistency Detection with BERT and Longformer 29 Jul 2022 · 1 repository · arXiv:2207.14444
-
Curriculum Learning for Data-Efficient Vision-Language Alignment 29 Jul 2022 · 0 repositories · arXiv:2207.14525
-
SERCNN: Stacked Embedding Recurrent Convolutional Neural Network in Detecting Depression on Twitter 29 Jul 2022 · 0 repositories · arXiv:2207.14535
-
CrAM: A Compression-Aware Minimizer 28 Jul 2022 · 1 repository · arXiv:2207.14200Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
SDBERT: SparseDistilBERT, a faster and smaller BERT model 28 Jul 2022 · 0 repositories · arXiv:2208.10246
-
Sequence to sequence pretraining for a less-resourced Slovenian language 28 Jul 2022 · 1 repository · arXiv:2207.13988
-
SoundChoice: Grapheme-to-Phoneme Models with Semantic Disambiguation 27 Jul 2022 · 1 repository · arXiv:2207.13703
-
Bundle MCR: Towards Conversational Bundle Recommendation 26 Jul 2022 · 1 repository · arXiv:2207.12628
-
Fine-Tuning BERT for Automatic ADME Semantic Labeling in FDA Drug Labeling to Enhance Product-Specific Guidance Assessment 25 Jul 2022 · 0 repositories · arXiv:2207.12376
-
A Cognitive Study on Semantic Similarity Analysis of Large Corpora: A Transformer-based Approach 24 Jul 2022 · 0 repositories · arXiv:2207.11716
-
Better Reasoning Behind Classification Predictions with BERT for Fake News Detection 23 Jul 2022 · 0 repositories · arXiv:2207.11562
-
Efficient model compression with Random Operation Access Specific Tile (ROAST) hashing 21 Jul 2022 · 1 repository · arXiv:2207.10702
-
Enhancing Collaborative Filtering Recommender with Prompt-Based Sentiment Analysis 19 Jul 2022 · 1 repository · arXiv:2207.12883
-
PiC: A Phrase-in-Context Dataset for Phrase Understanding and Semantic Search 19 Jul 2022 · 1 repository · arXiv:2207.09068
-
Pre-trained language models with domain knowledge for biomedical extractive summarization 19 Jul 2022 · 1 repository
-
Revealing Secrets From Pre-trained Models 19 Jul 2022 · 0 repositories · arXiv:2207.09539
-
Selection Bias Induced Spurious Correlations in Large Language Models 18 Jul 2022 · 1 repository · arXiv:2207.08982
-
Aspect-specific Context Modeling for Aspect-based Sentiment Analysis 17 Jul 2022 · 1 repository · arXiv:2207.08099
-
ELECTRA is a Zero-Shot Learner, Too 17 Jul 2022 · 1 repository · arXiv:2207.08141Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Representation Learning of Image Schema 17 Jul 2022 · 0 repositories · arXiv:2207.08256
-
Robust Action Governor for Uncertain Piecewise Affine Systems with Non-convex Constraints and Safe Reinforcement Learning 17 Jul 2022 · 0 repositories · arXiv:2207.08240
-
A Context-Sensitive Word Embedding Approach for The Detection of Troll Tweets 17 Jul 2022 · 0 repositories · arXiv:2207.08230
-
POET: Training Neural Networks on Tiny Devices with Integrated Rematerialization and Paging 15 Jul 2022 · 1 repository · arXiv:2207.07697
-
Position Prediction as an Effective Pretraining Strategy 15 Jul 2022 · 1 repository · arXiv:2207.07611