Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 27
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 27 of 71: papers 2,601 to 2,700 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Can Model Fusing Help Transformers in Long Document Classification? An Empirical Study 18 Jul 2023 · 1 repository · arXiv:2307.09532
-
KATIE: A System for Key Attributes Identification in Product Knowledge Graph Construction 18 Jul 2023 · 0 repositories
-
Cross-Lingual NER for Financial Transaction Data in Low-Resource Languages 16 Jul 2023 · 0 repositories · arXiv:2307.08714
-
Recognition of Mental Adjectives in An Efficient and Automatic Style 16 Jul 2023 · 0 repositories · arXiv:2307.11767
-
Do not Mask Randomly: Effective Domain-adaptive Pre-training by Masking In-domain Keywords 14 Jul 2023 · 0 repositories · arXiv:2307.07160
-
Improving BERT with Hybrid Pooling Network and Drop Mask 14 Jul 2023 · 0 repositories · arXiv:2307.07258
-
Sensi-BERT: Towards Sensitivity Driven Fine-Tuning for Parameter-Efficient BERT 14 Jul 2023 · 0 repositories · arXiv:2307.11764
-
Towards spoken dialect identification of Irish 14 Jul 2023 · 0 repositories · arXiv:2307.07436
-
TVPR: Text-to-Video Person Retrieval and a New Benchmark 14 Jul 2023 · 0 repositories · arXiv:2307.07184
-
Convolutional Neural Networks for Sentiment Analysis on Weibo Data: A Natural Language Processing Approach 13 Jul 2023 · 0 repositories · arXiv:2307.06540
-
SecureFalcon: Are We There Yet in Automated Software Vulnerability Detection with LLMs? 13 Jul 2023 · 0 repositories · arXiv:2307.06616
-
Tackling Fake News in Bengali: Unraveling the Impact of Summarization vs. Augmentation on Pre-trained Language Models 13 Jul 2023 · 1 repository · arXiv:2307.06979
-
Retrieval Augmented Generation using Engineering Design Knowledge 13 Jul 2023 · 2 repositories · arXiv:2307.06985
-
Detecting the Presence of COVID-19 Vaccination Hesitancy from South African Twitter Data Using Machine Learning 12 Jul 2023 · 0 repositories · arXiv:2307.15072
-
No Train No Gain: Revisiting Efficient Training Algorithms For Transformer-based Language Models 12 Jul 2023 · 1 repository · arXiv:2307.06440Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 3 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Prompt Generate Train (PGT): Few-shot Domain Adaption of Retrieval Augmented Generation Models for Open Book Question-Answering 12 Jul 2023 · 0 repositories · arXiv:2307.05915
-
Vacaspati: A Diverse Corpus of Bangla Literature 11 Jul 2023 · 0 repositories · arXiv:2307.05083
-
ChatGPT for Digital Forensic Investigation: The Good, The Bad, and The Unknown 10 Jul 2023 · 1 repository · arXiv:2307.10195
-
Is ChatGPT a Good Personality Recognizer? A Preliminary Study 8 Jul 2023 · 0 repositories · arXiv:2307.03952
-
Goal-Conditioned Predictive Coding for Offline Reinforcement Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03406
-
TRAQ: Trustworthy Retrieval Augmented Question Answering via Conformal Prediction 7 Jul 2023 · 1 repository · arXiv:2307.04642
-
A Novel Site-Agnostic Multimodal Deep Learning Model to Identify Pro-Eating Disorder Content on Social Media 6 Jul 2023 · 0 repositories · arXiv:2307.06775
-
Can ChatGPT's Responses Boost Traditional Natural Language Processing? 6 Jul 2023 · 1 repository · arXiv:2307.04648
-
Text Alignment Is An Efficient Unified Model for Massive NLP Tasks 6 Jul 2023 · 1 repository · arXiv:2307.02729
-
UIT-Saviors at MEDVQA-GI 2023: Improving Multimodal Learning with Image Enhancement for Gastrointestinal Visual Question Answering 6 Jul 2023 · 0 repositories · arXiv:2307.02783
-
CAME: Confidence-guided Adaptive Memory Efficient Optimization 5 Jul 2023 · 2 repositories · arXiv:2307.02047Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Emoji Prediction in Tweets using BERT 5 Jul 2023 · 1 repository · arXiv:2307.02054
-
Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations 5 Jul 2023 · 0 repositories · arXiv:2307.03678
-
Named Entity Inclusion in Abstractive Text Summarization 5 Jul 2023 · 0 repositories · arXiv:2307.02570
-
KDSTM: Neural Semi-supervised Topic Modeling with Knowledge Distillation 4 Jul 2023 · 0 repositories · arXiv:2307.01878Syntology 6 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
ALBERTI, a Multilingual Domain Specific Language Model for Poetry Analysis 3 Jul 2023 · 0 repositories · arXiv:2307.01387
-
Improving Language Plasticity via Pretraining with Active Forgetting 3 Jul 2023 · 1 repository · arXiv:2307.01163
-
Interpretability and Transparency-Driven Detection and Transformation of Textual Adversarial Examples (IT-DT) 3 Jul 2023 · 0 repositories · arXiv:2307.01225
-
How far is Language Model from 100% Few-shot Named Entity Recognition in Medical Domain 1 Jul 2023 · 1 repository · arXiv:2307.00186
-
Ticket-BERT: Labeling Incident Management Tickets with Language Models 30 Jun 2023 · 0 repositories · arXiv:2307.00108
-
Classifying Crime Types using Judgment Documents from Social Media 29 Jun 2023 · 0 repositories · arXiv:2306.17020
-
Harnessing the Power of Hugging Face Transformers for Predicting Mental Health Disorders in Social Networks 29 Jun 2023 · 0 repositories · arXiv:2306.16891
-
An Efficient Sparse Inference Software Accelerator for Transformer-based Language Models on CPUs 28 Jun 2023 · 1 repository · arXiv:2306.16601
-
Beyond the Hype: Assessing the Performance, Trustworthiness, and Clinical Suitability of GPT3.5 28 Jun 2023 · 0 repositories · arXiv:2306.15887
-
Multi-Site Clinical Federated Learning using Recursive and Attentive Models and NVFlare 28 Jun 2023 · 0 repositories · arXiv:2306.16367
-
Gender Bias in BERT -- Measuring and Analysing Biases through Sentiment Rating in a Realistic Downstream Classification Task 27 Jun 2023 · 0 repositories · arXiv:2306.15298
-
Investigating Cross-Domain Behaviors of BERT in Review Understanding 27 Jun 2023 · 0 repositories · arXiv:2306.15123
-
MAT: Mixed-Strategy Game of Adversarial Training in Fine-tuning 27 Jun 2023 · 0 repositories · arXiv:2306.15826
-
SparseOptimizer: Sparsify Language Models through Moreau-Yosida Regularization and Accelerate via Compiler Co-design 27 Jun 2023 · 0 repositories · arXiv:2306.15656
-
Unleashing the Power of User Reviews: Exploring Airline Choices at Catania Airport, Italy 27 Jun 2023 · 0 repositories · arXiv:2306.15541
-
Constraint-aware and Ranking-distilled Token Pruning for Efficient Transformer Inference 26 Jun 2023 · 1 repository · arXiv:2306.14393
-
Addressing Cold Start Problem for End-to-end Automatic Speech Scoring 25 Jun 2023 · 0 repositories · arXiv:2306.14310
-
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices 25 Jun 2023 · 0 repositories · arXiv:2306.14263
-
Switch-BERT: Learning to Model Multimodal Interactions by Switching Attention and Input 25 Jun 2023 · 0 repositories · arXiv:2306.14182
-
Comparison of Pre-trained Language Models for Turkish Address Parsing 24 Jun 2023 · 0 repositories · arXiv:2306.13947
-
IERL: Interpretable Ensemble Representation Learning -- Combining CrowdSourced Knowledge and Distributed Semantic Representations 24 Jun 2023 · 0 repositories · arXiv:2306.13865
-
L3Cube-MahaSent-MD: A Multi-domain Marathi Sentiment Analysis Dataset and Transformer Models 24 Jun 2023 · 1 repository · arXiv:2306.13888
-
Math Word Problem Solving by Generating Linguistic Variants of Problem Statements 24 Jun 2023 · 1 repository · arXiv:2306.13899
-
My Boli: Code-mixed Marathi-English Corpora, Pretrained Language Models and Evaluation Benchmarks 24 Jun 2023 · 1 repository · arXiv:2306.14030
-
Partitioning-Guided K-Means: Extreme Empty Cluster Resolution for Extreme Model Compression 24 Jun 2023 · 0 repositories · arXiv:2306.14031
-
Resume Information Extraction via Post-OCR Text Processing 23 Jun 2023 · 0 repositories · arXiv:2306.13775
-
Named entity recognition in resumes 22 Jun 2023 · 0 repositories · arXiv:2306.13062
-
Investigating Pre-trained Language Models on Cross-Domain Datasets, a Step Closer to General AI 21 Jun 2023 · 0 repositories · arXiv:2306.12205
-
Fine-Tuning Language Models for Scientific Writing Support 19 Jun 2023 · 1 repository · arXiv:2306.10974
-
Instant Soup: Cheap Pruning Ensembles in A Single Pass Can Draw Lottery Tickets from Large Models 18 Jun 2023 · 1 repository · arXiv:2306.10460Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Investigating Masking-based Data Generation in Language Models 16 Jun 2023 · 0 repositories · arXiv:2307.00008
-
Revealing the impact of social circumstances on the selection of cancer therapy through natural language processing of social work notes 16 Jun 2023 · 0 repositories · arXiv:2306.09877
-
BED: Bi-Encoder-Based Detectors for Out-of-Distribution Detection 15 Jun 2023 · 1 repository · arXiv:2306.08852
-
Distillation Strategies for Discriminative Speech Recognition Rescoring 15 Jun 2023 · 0 repositories · arXiv:2306.09452
-
Mapping Researcher Activity based on Publication Data by means of Transformers 15 Jun 2023 · 0 repositories · arXiv:2306.09049
-
SLAMB: Accelerated Large Batch Training with Sparse Communication 15 Jun 2023 · 1 repository
-
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization 15 Jun 2023 · 0 repositories · arXiv:2306.09222
-
A semantically enhanced dual encoder for aspect sentiment triplet extraction 14 Jun 2023 · 1 repository · arXiv:2306.08373
-
Building a Corpus for Biomedical Relation Extraction of Species Mentions 14 Jun 2023 · 0 repositories · arXiv:2306.08403
-
Language models are not naysayers: An analysis of language models on negation benchmarks 14 Jun 2023 · 1 repository · arXiv:2306.08189
-
Research on Named Entity Recognition in Improved transformer with R-Drop structure 14 Jun 2023 · 0 repositories · arXiv:2306.08315
-
World-to-Words: Grounded Open Vocabulary Acquisition through Fast Mapping in Vision-Language Models 14 Jun 2023 · 1 repository · arXiv:2306.08685Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition 13 Jun 2023 · 0 repositories · arXiv:2306.07848
-
Improving Zero-Shot Detection of Low Prevalence Chest Pathologies using Domain Pre-trained Language Models 13 Jun 2023 · 1 repository · arXiv:2306.08000
-
Monolingual and Cross-Lingual Knowledge Transfer for Topic Classification 13 Jun 2023 · 0 repositories · arXiv:2306.07797
-
A Survey of Vision-Language Pre-training from the Lens of Multimodal Machine Translation 12 Jun 2023 · 0 repositories · arXiv:2306.07198
-
Imbalanced Multi-label Classification for Business-related Text with Moderately Large Label Spaces 12 Jun 2023 · 0 repositories · arXiv:2306.07046
-
Linear Classifier: An Often-Forgotten Baseline for Text Classification 12 Jun 2023 · 1 repository · arXiv:2306.07111Syntology official: no sample here; runs from other or unrecorded repositories · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Multimodal Audio-textual Architecture for Robust Spoken Language Understanding 12 Jun 2023 · 0 repositories · arXiv:2306.06819
-
EaSyGuide : ESG Issue Identification Framework leveraging Abilities of Generative Large Language Models 11 Jun 2023 · 1 repository · arXiv:2306.06662
-
RoBERTweet: A BERT Language Model for Romanian Tweets 11 Jun 2023 · 0 repositories · arXiv:2306.06598
-
Enhancing Low Resource NER Using Assisting Language And Transfer Learning 10 Jun 2023 · 0 repositories · arXiv:2306.06477
-
Medical Data Augmentation via ChatGPT: A Case Study on Medication Identification and Medication Event Classification 10 Jun 2023 · 0 repositories · arXiv:2306.07297
-
COVER: A Heuristic Greedy Adversarial Attack on Prompt-based Learning in Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.05659
-
End-to-End Neural Network Compression via ℓ₁/ℓ₂ Regularized Latency Surrogates 9 Jun 2023 · 0 repositories · arXiv:2306.05785
-
Implementing BERT and fine-tuned RobertA to detect AI generated news by ChatGPT 9 Jun 2023 · 0 repositories · arXiv:2306.07401
-
Prodigy: An Expeditiously Adaptive Parameter-Free Learner 9 Jun 2023 · 1 repository · arXiv:2306.06101
-
Understanding Telecom Language Through Large Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.07933
-
Augmenting Hessians with Inter-Layer Dependencies for Mixed-Precision Post-Training Quantization 8 Jun 2023 · 0 repositories · arXiv:2306.04879
-
Bias Against 93 Stigmatized Groups in Masked Language Models and Downstream Sentiment Classification Tasks 8 Jun 2023 · 1 repository · arXiv:2306.05550Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Extensive Evaluation of Transformer-based Architectures for Adverse Drug Events Extraction 8 Jun 2023 · 1 repository · arXiv:2306.05276
-
Leveraging Language Identification to Enhance Code-Mixed Text Classification 8 Jun 2023 · 0 repositories · arXiv:2306.04964
-
Mixture-of-Supernets: Improving Weight-Sharing Supernet Training with Architecture-Routed Mixture-of-Experts 8 Jun 2023 · 1 repository · arXiv:2306.04845
-
NOWJ at COLIEE 2023 -- Multi-Task and Ensemble Approaches in Legal Information Processing 8 Jun 2023 · 0 repositories · arXiv:2306.04903
-
An Empirical Analysis of Parameter-Efficient Methods for Debiasing Pre-Trained Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04067
-
Detecting Human Rights Violations on Social Media during Russia-Ukraine War 6 Jun 2023 · 0 repositories · arXiv:2306.05370
-
LEACE: Perfect linear concept erasure in closed form 6 Jun 2023 · 2 repositories · arXiv:2306.03819Syntology official (archive's flag): 3 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
On the Difference of BERT-style and CLIP-style Text Encoders 6 Jun 2023 · 1 repository · arXiv:2306.03678
-
COMET: Learning Cardinality Constrained Mixture of Experts with Trees and Local Search 5 Jun 2023 · 2 repositories · arXiv:2306.02824
-
On "Scientific Debt" in NLP: A Case for More Rigour in Language Model Pre-Training Research 5 Jun 2023 · 0 repositories · arXiv:2306.02870