Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 28
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 28 of 71: papers 2,701 to 2,800 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Skill over Scale: The Case for Medium, Domain-Specific Models for SE 5 Jun 2023 · 0 repositories · arXiv:2306.03268
-
Using Sequences of Life-events to Predict Human Lives 5 Jun 2023 · 2 repositories · arXiv:2306.03009
-
SpellMapper: A non-autoregressive neural spellchecker for ASR customization with candidate retrieval based on n-gram mappings 4 Jun 2023 · 1 repository · arXiv:2306.02317
-
Financial sentiment analysis using FinBERT with application in predicting stock movement 3 Jun 2023 · 0 repositories · arXiv:2306.02136
-
MultiLegalPile: A 689GB Multilingual Legal Corpus 3 Jun 2023 · 0 repositories · arXiv:2306.02069
-
Concurrent Classifier Error Detection (CCED) in Large Scale Machine Learning Systems 2 Jun 2023 · 0 repositories · arXiv:2306.01820
-
Establishment of NLP-Based Greenwashing Pattern Detection Service 2 Jun 2023 · 0 repositories
-
Context selectivity with dynamic availability enables lifelong continual learning 2 Jun 2023 · 1 repository · arXiv:2306.01690
-
Word Embeddings for Banking Industry 2 Jun 2023 · 0 repositories · arXiv:2306.01807
-
Adapting Pre-trained Language Models to Vision-Language Tasks via Dynamic Visual Prompting 1 Jun 2023 · 1 repository · arXiv:2306.00409
-
Boosting the Performance of Transformer Architectures for Semantic Textual Similarity 1 Jun 2023 · 0 repositories · arXiv:2306.00708
-
Column Type Annotation using ChatGPT 1 Jun 2023 · 1 repository · arXiv:2306.00745
-
Feature Engineering-Based Detection of Buffer Overflow Vulnerability in Source Code Using Neural Networks 1 Jun 2023 · 0 repositories · arXiv:2306.07981
-
Make Pre-trained Model Reversible: From Parameter to Memory Efficient Fine-Tuning 1 Jun 2023 · 1 repository · arXiv:2306.00477Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Training-free Neural Architecture Search for RNNs and Transformers 1 Jun 2023 · 1 repository · arXiv:2306.00288Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
UCAS-IIE-NLP at SemEval-2023 Task 12: Enhancing Generalization of Multilingual BERT for Low-resource Sentiment Analysis 1 Jun 2023 · 1 repository · arXiv:2306.01093
-
Supplementary Features of BiLSTM for Enhanced Sequence Labeling 31 May 2023 · 1 repository · arXiv:2305.19928
-
Building Extractive Question Answering System to Support Human-AI Health Coaching Model for Sleep Domain 31 May 2023 · 0 repositories · arXiv:2305.19707
-
Catalysis distillation neural network for the few shot open catalyst challenge 31 May 2023 · 0 repositories · arXiv:2305.19545
-
DeepMerge: Deep-Learning-Based Region-Merging for Image Segmentation 31 May 2023 · 1 repository · arXiv:2305.19787
-
XPhoneBERT: A Pre-trained Multilingual Model for Phoneme Representations for Text-to-Speech 31 May 2023 · 2 repositories · arXiv:2305.19709Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples) · 10 pointer-only (licence)
-
Explaining Hate Speech Classification with Model Agnostic Methods 30 May 2023 · 0 repositories · arXiv:2306.00021
-
GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation 30 May 2023 · 0 repositories · arXiv:2305.18997
-
Multitask learning for recognizing stress and depression in social media 30 May 2023 · 0 repositories · arXiv:2305.18907
-
PreQuant: A Task-agnostic Quantization Approach for Pre-trained Language Models 30 May 2023 · 0 repositories · arXiv:2306.00014
-
Research on Multilingual News Clustering Based on Cross-Language Word Embeddings 30 May 2023 · 0 repositories · arXiv:2305.18880
-
ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning 30 May 2023 · 1 repository · arXiv:2305.19426
-
Abstractive Summarization as Augmentation for Document-Level Event Detection 29 May 2023 · 0 repositories · arXiv:2305.18023
-
From Adversarial Arms Race to Model-centric Evaluation: Motivating a Unified Automatic Robustness Evaluation Framework 29 May 2023 · 1 repository · arXiv:2305.18503Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
SlimFit: Memory-Efficient Fine-Tuning of Transformer-based Models Using Training Dynamics 29 May 2023 · 0 repositories · arXiv:2305.18513
-
Transformer Language Models Handle Word Frequency in Prediction Head 29 May 2023 · 0 repositories · arXiv:2305.18294
-
Rethinking Masked Language Modeling for Chinese Spelling Correction 28 May 2023 · 1 repository · arXiv:2305.17721
-
Transfer Learning for Power Outage Detection Task with Limited Training Data 28 May 2023 · 0 repositories · arXiv:2305.17817
-
Diagnosing Transformers: Illuminating Feature Spaces for Clinical Decision-Making 27 May 2023 · 1 repository · arXiv:2305.17588
-
Complementary and Integrative Health Lexicon (CIHLex) and Entity Recognition in the Literature 27 May 2023 · 0 repositories · arXiv:2305.17353
-
Modeling Adversarial Attack on Pre-trained Language Models as Sequential Decision Making 27 May 2023 · 1 repository · arXiv:2305.17440
-
Calibration of Transformer-based Models for Identifying Stress and Depression in Social Media 26 May 2023 · 0 repositories · arXiv:2305.16797
-
GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot Attention for Vision-and-Language Navigation 26 May 2023 · 1 repository · arXiv:2305.17102
-
Incorporating Distributions of Discourse Structure for Long Document Abstractive Summarization 26 May 2023 · 0 repositories · arXiv:2305.16784
-
KNSE: A Knowledge-aware Natural Language Inference Framework for Dialogue Symptom Status Recognition 26 May 2023 · 0 repositories · arXiv:2305.16833
-
Theoretical and Practical Perspectives on what Influence Functions Do 26 May 2023 · 0 repositories · arXiv:2305.16971
-
Zero is Not Hero Yet: Benchmarking Zero-Shot Performance of LLMs for Financial Tasks 26 May 2023 · 1 repository · arXiv:2305.16633
-
Comparative Study of Pre-Trained BERT Models for Code-Mixed Hindi-English Data 25 May 2023 · 0 repositories · arXiv:2305.15722
-
Context-aware attention layers coupled with optimal transport domain adaptation and multimodal fusion methods for recognizing dementia from spontaneous speech 25 May 2023 · 0 repositories · arXiv:2305.16406
-
Not wacky vs. definitely wacky: A study of scalar adverbs in pretrained language models 25 May 2023 · 0 repositories · arXiv:2305.16426
-
Text-to-Motion Retrieval: Towards Joint Understanding of Human Motion Data and Natural Language 25 May 2023 · 1 repository · arXiv:2305.15842
-
A Causal View of Entity Bias in (Large) Language Models 24 May 2023 · 1 repository · arXiv:2305.14695Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Complex Mathematical Symbol Definition Structures: A Dataset and Model for Coordination Resolution in Definition Extraction 24 May 2023 · 1 repository · arXiv:2305.14660
-
Context-Aware Transformer Pre-Training for Answer Sentence Selection 24 May 2023 · 0 repositories · arXiv:2305.15358
-
Dynamic Masking Rate Schedules for MLM Pretraining 24 May 2023 · 0 repositories · arXiv:2305.15096
-
Extracting Psychological Indicators Using Question Answering 24 May 2023 · 0 repositories · arXiv:2305.14891
-
Ghostbuster: Detecting Text Ghostwritten by Large Language Models 24 May 2023 · 2 repositories · arXiv:2305.15047
-
How to Distill your BERT: An Empirical Study on the Impact of Weight Initialisation and Distillation Objectives 24 May 2023 · 1 repository · arXiv:2305.15032Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Neural Summarization of Electronic Health Records 24 May 2023 · 0 repositories · arXiv:2305.15222
-
Revisiting Token Dropping Strategy in Efficient BERT Pretraining 24 May 2023 · 1 repository · arXiv:2305.15273
-
Few-shot Adaptation to Distribution Shifts By Mixing Source and Target Embeddings 23 May 2023 · 0 repositories · arXiv:2305.14521
-
When your Cousin has the Right Connections: Unsupervised Bilingual Lexicon Induction for Related Data-Imbalanced Languages 23 May 2023 · 1 repository · arXiv:2305.14012
-
All Roads Lead to Rome? Exploring the Invariance of Transformers' Representations 23 May 2023 · 1 repository · arXiv:2305.14555Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Assessing Linguistic Generalisation in Language Models: A Dataset for Brazilian Portuguese 23 May 2023 · 0 repositories · arXiv:2305.14070
-
AxomiyaBERTa: A Phonologically-aware Transformer Model for Assamese 23 May 2023 · 1 repository · arXiv:2305.13641
-
Connecting the Dots: What Graph-Based Text Representations Work Best for Text Classification Using Graph Neural Networks? 23 May 2023 · 1 repository · arXiv:2305.14578
-
Exploring Large Language Models for Classical Philology 23 May 2023 · 1 repository · arXiv:2305.13698
-
Handling Realistic Label Noise in BERT Text Classification 23 May 2023 · 0 repositories · arXiv:2305.16337
-
On Robustness of Finetuned Transformer-based NLP Models 23 May 2023 · 1 repository · arXiv:2305.14453
-
Text Is All You Need: Learning Language Representations for Sequential Recommendation 23 May 2023 · 1 repository · arXiv:2305.13731Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Training Transitive and Commutative Multimodal Transformers with LoReTTa 23 May 2023 · 0 repositories · arXiv:2305.14243
-
Exploring Energy-based Language Models with Different Architectures and Training Methods for Speech Recognition 22 May 2023 · 2 repositories · arXiv:2305.12676
-
GATology for Linguistics: What Syntactic Dependencies It Knows 22 May 2023 · 0 repositories · arXiv:2305.13403
-
SimCSE++: Improving Contrastive Learning for Sentence Embeddings from Two Perspectives 22 May 2023 · 0 repositories · arXiv:2305.13192
-
Language-Agnostic Bias Detection in Language Models with Bias Probing 22 May 2023 · 1 repository · arXiv:2305.13302
-
Atomic Inference for NLI with Generated Facts as Atoms 22 May 2023 · 1 repository · arXiv:2305.13214
-
Stock and market index prediction using Informer network 22 May 2023 · 0 repositories · arXiv:2305.14382
-
Syntactic Knowledge via Graph Attention with BERT in Machine Translation 22 May 2023 · 0 repositories · arXiv:2305.13413
-
A Deeper (Autoregressive) Approach to Non-Convergent Discourse Parsing 21 May 2023 · 0 repositories · arXiv:2305.12510
-
A Symbolic Framework for Evaluating Mathematical Reasoning and Generalisation with Transformers 21 May 2023 · 0 repositories · arXiv:2305.12563
-
BertRLFuzzer: A BERT and Reinforcement Learning Based Fuzzer 21 May 2023 · 1 repository · arXiv:2305.12534
-
F-PABEE: Flexible-patience-based Early Exiting for Single-label and Multi-label text Classification Tasks 21 May 2023 · 0 repositories · arXiv:2305.11916
-
Infor-Coef: Information Bottleneck-based Dynamic Token Downsampling for Compact and Efficient language model 21 May 2023 · 0 repositories · arXiv:2305.12458
-
IR Models and the COVID-19 Pandemic: A Comparative Study of Performance and Challenges 21 May 2023 · 0 repositories · arXiv:2305.12528
-
CDJUR-BR -- A Golden Collection of Legal Document from Brazilian Justice with Fine-Grained Named Entities 20 May 2023 · 0 repositories · arXiv:2305.18315
-
SEntFiN 1.0: Entity-Aware Sentiment Analysis for Financial News 20 May 2023 · 0 repositories · arXiv:2305.12257
-
A Sequence-to-Sequence Approach for Arabic Pronoun Resolution 19 May 2023 · 0 repositories · arXiv:2305.11529
-
Eye-SpatialNet: Spatial Information Extraction from Ophthalmology Notes 19 May 2023 · 0 repositories · arXiv:2305.11948
-
Federated Foundation Models: Privacy-Preserving and Collaborative Learning for Large Models 19 May 2023 · 0 repositories · arXiv:2305.11414
-
Language-Universal Phonetic Representation in Multilingual Speech Pretraining for Low-Resource Speech Recognition 19 May 2023 · 0 repositories · arXiv:2305.11569
-
Ahead-of-Time P-Tuning 18 May 2023 · 0 repositories · arXiv:2305.10835
-
Ditto: A Simple and Efficient Approach to Improve Sentence Embeddings 18 May 2023 · 1 repository · arXiv:2305.10786
-
PDP: Parameter-free Differentiable Pruning is All You Need 18 May 2023 · 0 repositories · arXiv:2305.11203
-
Trading Syntax Trees for Wordpieces: Target-oriented Opinion Words Extraction with Wordpieces and Aspect Enhancement 18 May 2023 · 0 repositories · arXiv:2305.11034
-
A quantitative study of NLP approaches to question difficulty estimation 17 May 2023 · 1 repository · arXiv:2305.10236
-
AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression 17 May 2023 · 1 repository · arXiv:2305.10010Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Explaining black box text modules in natural language with language models 17 May 2023 · 2 repositories · arXiv:2305.09863Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Solving Cosine Similarity Underestimation between High Frequency Words by L2 Norm Discounting 17 May 2023 · 1 repository · arXiv:2305.10610
-
CWTM: Leveraging Contextualized Word Embeddings from BERT for Neural Topic Modeling 16 May 2023 · 1 repository · arXiv:2305.09329
-
Measuring Dimensions of Self-Presentation in Twitter Bios and their Links to Misinformation Sharing 16 May 2023 · 1 repository · arXiv:2305.09548
-
Weight-Inherited Distillation for Task-Agnostic BERT Compression 16 May 2023 · 1 repository · arXiv:2305.09098
-
Coreference-aware Double-channel Attention Network for Multi-party Dialogue Reading Comprehension 15 May 2023 · 1 repository · arXiv:2305.08348
-
Knowledge Rumination for Pre-trained Language Models 15 May 2023 · 1 repository · arXiv:2305.08732
-
Private Training Set Inspection in MLaaS 15 May 2023 · 0 repositories · arXiv:2305.09058
-
Text2Gender: A Deep Learning Architecture for Analysis of Blogger's Age and Gender 15 May 2023 · 0 repositories · arXiv:2305.08633