Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 33
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 33 of 71: papers 3,201 to 3,300 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Probing for targeted syntactic knowledge through grammatical error detection 28 Oct 2022 · 1 repository · arXiv:2210.16228
-
BERT-Flow-VAE: A Weakly-supervised Model for Multi-Label Text Classification 27 Oct 2022 · 0 repositories · arXiv:2210.15225
-
COCO-DR: Combating Distribution Shifts in Zero-Shot Dense Retrieval with Contrastive and Distributionally Robust Learning 27 Oct 2022 · 1 repository · arXiv:2210.15212Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
COST-EFF: Collaborative Optimization of Spatial and Temporal Efficiency with Slenderized Multi-exit Language Models 27 Oct 2022 · 1 repository · arXiv:2210.15523
-
Fast DistilBERT on CPUs 27 Oct 2022 · 1 repository · arXiv:2211.07715
-
FCTalker: Fine and Coarse Grained Context Modeling for Expressive Conversational Speech Synthesis 27 Oct 2022 · 1 repository · arXiv:2210.15360
-
Masked Vision-Language Transformer in Fashion 27 Oct 2022 · 1 repository · arXiv:2210.15110
-
Unsupervised Boundary-Aware Language Model Pretraining for Chinese Sequence Labeling 27 Oct 2022 · 2 repositories · arXiv:2210.15231
-
Automatic extraction of materials and properties from superconductors scientific literature 26 Oct 2022 · 2 repositories · arXiv:2210.15600
-
Beyond English-Centric Bitexts for Better Multilingual Language Representation Learning 26 Oct 2022 · 0 repositories · arXiv:2210.14867
-
Bi-Link: Bridging Inductive Link Predictions from Text via Contrastive Learning of Transformers and Prompts 26 Oct 2022 · 0 repositories · arXiv:2210.14463
-
Don't Prompt, Search! Mining-based Zero-Shot Learning with Language Models 26 Oct 2022 · 0 repositories · arXiv:2210.14803
-
Exploring Robustness of Prefix Tuning in Noisy Data: A Case Study in Financial Sentiment Analysis 26 Oct 2022 · 0 repositories · arXiv:2211.05584
-
How Long Is Enough? Exploring the Optimal Intervals of Long-Range Clinical Note Language Modeling 25 Oct 2022 · 1 repository · arXiv:2211.07713
-
IELM: An Open Information Extraction Benchmark for Pre-Trained Language Models 25 Oct 2022 · 0 repositories · arXiv:2210.14128
-
Effective Pre-Training Objectives for Transformer-based Autoencoders 24 Oct 2022 · 0 repositories · arXiv:2210.13536
-
Entity-level Sentiment Analysis in Contact Center Telephone Conversations 24 Oct 2022 · 0 repositories · arXiv:2210.13401
-
Explaining Translationese: why are Neural Classifiers Better and what do they Learn? 24 Oct 2022 · 0 repositories · arXiv:2210.13391
-
Exploring Euphemism Detection in Few-Shot and Zero-Shot Settings 24 Oct 2022 · 1 repository · arXiv:2210.12926
-
The Better Your Syntax, the Better Your Semantics? Probing Pretrained Language Models for the English Comparative Correlative 24 Oct 2022 · 0 repositories · arXiv:2210.13181
-
A BERT-based Deep Learning Approach for Reputation Analysis in Social Media 23 Oct 2022 · 0 repositories · arXiv:2211.01954
-
Data Augmentation for Automated Essay Scoring using Transformer Models 23 Oct 2022 · 0 repositories · arXiv:2210.12809
-
Discriminative Language Model as Semantic Consistency Scorer for Prompt-based Few-Shot Text Classification 23 Oct 2022 · 0 repositories · arXiv:2210.12763
-
Meta-learning Pathologies from Radiology Reports using Variance Aware Prototypical Networks 22 Oct 2022 · 0 repositories · arXiv:2210.13979
-
Amos: An Adam-style Optimizer with Adaptive Weight Decay towards Model-Oriented Scale 21 Oct 2022 · 1 repository · arXiv:2210.11693Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples)
-
Discovering Differences in the Representation of People using Contextualized Semantic Axes 21 Oct 2022 · 1 repository · arXiv:2210.12170
-
LittleBird: Efficient Faster & Longer Transformer for Question Answering 21 Oct 2022 · 0 repositories · arXiv:2210.11870
-
Probing with Noise: Unpicking the Warp and Weft of Embeddings 21 Oct 2022 · 1 repository · arXiv:2210.12206
-
SpaBERT: A Pretrained Language Model from Geographic Data for Geo-Entity Representation 21 Oct 2022 · 0 repositories · arXiv:2210.12213
-
A Unified Neural Network Model for Readability Assessment with Feature Projection and Length-Balanced Loss 19 Oct 2022 · 1 repository · arXiv:2210.10305
-
BioGPT: Generative Pre-trained Transformer for Biomedical Text Generation and Mining 19 Oct 2022 · 4 repositories · arXiv:2210.10341
-
Language Model Decomposition: Quantifying the Dependency and Correlation of Language Models 19 Oct 2022 · 1 repository · arXiv:2210.10289
-
Tempo: Accelerating Transformer-Based Model Training through Memory Footprint Reduction 19 Oct 2022 · 1 repository · arXiv:2210.10246Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
ELASTIC: Numerical Reasoning with Adaptive Symbolic Compiler 18 Oct 2022 · 1 repository · arXiv:2210.10105Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
CAN-BERT do it? Controller Area Network Intrusion Detection System based on BERT Language Model 17 Oct 2022 · 0 repositories · arXiv:2210.09439
-
iDNA-ABF: multi-scale deep biological language learning model for the interpretable prediction of DNA methylations 17 Oct 2022 · 2 repositories
-
Multi-granularity Argument Mining in Legal Texts 17 Oct 2022 · 0 repositories · arXiv:2210.09472
-
Using Bottleneck Adapters to Identify Cancer in Clinical Notes under Low-Resource Constraints 17 Oct 2022 · 1 repository · arXiv:2210.09440
-
Zero-Shot Ranking Socio-Political Texts with Transformer Language Models to Reduce Close Reading Time 17 Oct 2022 · 0 repositories · arXiv:2210.09179
-
Acoustic-aware Non-autoregressive Spell Correction with Mask Sample Decoding 16 Oct 2022 · 0 repositories · arXiv:2210.08665
-
CTCBERT: Advancing Hidden-unit BERT with CTC Objectives 16 Oct 2022 · 0 repositories · arXiv:2210.08603
-
Improving Semantic Matching through Dependency-Enhanced Pre-trained Model with Adaptive Fusion 16 Oct 2022 · 0 repositories · arXiv:2210.08471
-
AraLegal-BERT: A pretrained language model for Arabic Legal text 15 Oct 2022 · 0 repositories · arXiv:2210.08284
-
DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation 14 Oct 2022 · 2 repositories · arXiv:2210.07558Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Kernel-Whitening: Overcome Dataset Bias with Isotropic Sentence Embedding 14 Oct 2022 · 2 repositories · arXiv:2210.07547
-
Overlooked Video Classification in Weakly Supervised Video Anomaly Detection 13 Oct 2022 · 1 repository · arXiv:2210.06688
-
SQuAT: Sharpness- and Quantization-Aware Training for BERT 13 Oct 2022 · 0 repositories · arXiv:2210.07171
-
Tone prediction and orthographic conversion for Basaa 13 Oct 2022 · 0 repositories · arXiv:2210.06986
-
Foundation Transformers 12 Oct 2022 · 4 repositories · arXiv:2210.06423Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
GMP*: Well-Tuned Gradual Magnitude Pruning Can Outperform Most BERT-Pruning Methods 12 Oct 2022 · 0 repositories · arXiv:2210.06384
-
On Text Style Transfer via Style Masked Language Models 12 Oct 2022 · 0 repositories · arXiv:2210.06394
-
Probing Commonsense Knowledge in Pre-trained Language Models with Sense-level Precision and Expanded Vocabulary 12 Oct 2022 · 1 repository · arXiv:2210.06376
-
RankT5: Fine-Tuning T5 for Text Ranking with Ranking Losses 12 Oct 2022 · 0 repositories · arXiv:2210.10634
-
A Win-win Deal: Towards Sparse and Robust Pre-trained Language Models 11 Oct 2022 · 1 repository · arXiv:2210.05211Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
An Exploration of Hierarchical Attention Transformers for Efficient Long Document Classification 11 Oct 2022 · 0 repositories · arXiv:2210.05529
-
CLIP also Understands Text: Prompting CLIP for Phrase Understanding 11 Oct 2022 · 0 repositories · arXiv:2210.05836
-
Multilingual BERT has an accent: Evaluating English influences on fluency in multilingual models 11 Oct 2022 · 0 repositories · arXiv:2210.05619
-
On the Interpolation of Contextualized Term-based Ranking with BM25 for Query-by-Example Retrieval 11 Oct 2022 · 1 repository · arXiv:2210.05512
-
On the Use of Semantically-Aligned Speech Representations for Spoken Language Understanding 11 Oct 2022 · 0 repositories · arXiv:2210.05291
-
Vote'n'Rank: Revision of Benchmarking with Social Choice Theory 11 Oct 2022 · 1 repository · arXiv:2210.05769Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 17 harvested samples)
-
DEPTWEET: A Typology for Social Media Texts to Detect Depression Severities 10 Oct 2022 · 1 repository · arXiv:2210.05372
-
Empowering the Fact-checkers! Automatic Identification of Claim Spans on Twitter 10 Oct 2022 · 1 repository · arXiv:2210.04710
-
Multi-CLS BERT: An Efficient Alternative to Traditional Ensembling 10 Oct 2022 · 1 repository · arXiv:2210.05043
-
Uncertainty Quantification with Pre-trained Language Models: A Large-Scale Empirical Analysis 10 Oct 2022 · 1 repository · arXiv:2210.04714Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Fine-Grained Detection of Solidarity for Women and Migrants in 155 Years of German Parliamentary Debates 9 Oct 2022 · 2 repositories · arXiv:2210.04359
-
Better Pre-Training by Reducing Representation Confusion 9 Oct 2022 · 0 repositories · arXiv:2210.04246
-
Spread Love Not Hate: Undermining the Importance of Hateful Pre-training for Hate Speech Detection 9 Oct 2022 · 1 repository · arXiv:2210.04267
-
KG-MTT-BERT: Knowledge Graph Enhanced BERT for Multi-Type Medical Text Classification 8 Oct 2022 · 0 repositories · arXiv:2210.03970
-
On Task-Adaptive Pretraining for Dialogue Response Selection 8 Oct 2022 · 0 repositories · arXiv:2210.04073
-
DABERT: Dual Attention Enhanced BERT for Semantic Matching 7 Oct 2022 · 0 repositories · arXiv:2210.03454
-
Generating Quizzes to Support Training on Quality Management and Assurance in Space Science and Engineering 7 Oct 2022 · 0 repositories · arXiv:2210.03427
-
Knowledge Injected Prompt Based Fine-tuning for Multi-label Few-shot ICD Coding 7 Oct 2022 · 1 repository · arXiv:2210.03304Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
UU-Tax at SemEval-2022 Task 3: Improving the generalizability of language models for taxonomy classification through data augmentation 7 Oct 2022 · 1 repository · arXiv:2210.03378
-
ByteTransformer: A High-Performance Transformer Boosted for Variable-Length Inputs 6 Oct 2022 · 1 repository · arXiv:2210.03052
-
Explainable Verbal Deception Detection using Transformers 6 Oct 2022 · 0 repositories · arXiv:2210.03080
-
Improving the Domain Adaptation of Retrieval Augmented Generation (RAG) Models for Open Domain Question Answering 6 Oct 2022 · 1 repository · arXiv:2210.02627
-
Join-Chain Network: A Logical Reasoning View of the Multi-head Attention in Transformer 6 Oct 2022 · 0 repositories · arXiv:2210.02729
-
Matching Text and Audio Embeddings: Exploring Transfer-learning Strategies for Language-based Audio Retrieval 6 Oct 2022 · 0 repositories · arXiv:2210.02833
-
MuRAG: Multimodal Retrieval-Augmented Generator for Open Question Answering over Images and Text 6 Oct 2022 · 0 repositories · arXiv:2210.02928
-
Privacy-Preserving Text Classification on BERT Embeddings with Homomorphic Encryption 5 Oct 2022 · 0 repositories · arXiv:2210.02574
-
Revisiting Structured Dropout 5 Oct 2022 · 0 repositories · arXiv:2210.02570
-
The (In)Effectiveness of Intermediate Task Training For Domain Adaptation and Cross-Lingual Transfer Learning 3 Oct 2022 · 0 repositories · arXiv:2210.01091
-
Probing of Quantitative Values in Abstractive Summarization Models 3 Oct 2022 · 0 repositories · arXiv:2210.00667
-
Construction and Evaluation of a Self-Attention Model for Semantic Understanding of Sentence-Final Particles 1 Oct 2022 · 0 repositories · arXiv:2210.00282
-
LambdaKG: A Library for Pre-trained Language Model-Based Knowledge Graph Embeddings 1 Oct 2022 · 2 repositories · arXiv:2210.00305
-
CEFER: A Four Facets Framework based on Context and Emotion embedded features for Implicit and Explicit Emotion Recognition 28 Sep 2022 · 0 repositories · arXiv:2209.13999
-
Downstream Datasets Make Surprisingly Good Pretraining Corpora 28 Sep 2022 · 1 repository · arXiv:2209.14389
-
Supervised Contrastive Learning as Multi-Objective Optimization for Fine-Tuning Large Pre-trained Language Models 28 Sep 2022 · 0 repositories · arXiv:2209.14161
-
YATO: Yet Another deep learning based Text analysis Open toolkit 28 Sep 2022 · 1 repository · arXiv:2209.13877
-
Extractive Question Answering on Queries in Hindi and Tamil 27 Sep 2022 · 0 repositories · arXiv:2210.06356
-
Outlier Suppression: Pushing the Limit of Low-bit Transformer Language Models 27 Sep 2022 · 1 repository · arXiv:2209.13325Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
WikiDes: A Wikipedia-Based Dataset for Generating Short Descriptions from Paragraphs 27 Sep 2022 · 1 repository · arXiv:2209.13101
-
Do ever larger octopi still amplify reporting biases? Evidence from judgments of typical colour 26 Sep 2022 · 0 repositories · arXiv:2209.12786
-
Towards Simple and Efficient Task-Adaptive Pre-training for Text Classification 26 Sep 2022 · 0 repositories · arXiv:2209.12943
-
SpeedLimit: Neural Architecture Search for Quantized Transformer Models 25 Sep 2022 · 0 repositories · arXiv:2209.12127
-
Sentiment Analysis on Inflation after Covid-19 25 Sep 2022 · 0 repositories · arXiv:2209.14737
-
Can Transformer Models Effectively Detect Software Aspects in StackOverflow Discussion? 24 Sep 2022 · 0 repositories · arXiv:2209.12065
-
Learning Chess With Language Models and Transformers 24 Sep 2022 · 0 repositories · arXiv:2209.11902
-
IDEA: Interactive DoublE Attentions from Label Embedding for Text Classification 23 Sep 2022 · 0 repositories · arXiv:2209.11407
-
Adaptation of domain-specific transformer models with text oversampling for sentiment analysis of social media posts on Covid-19 vaccines 22 Sep 2022 · 1 repository · arXiv:2209.10966