Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 29
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 29 of 71: papers 2,801 to 2,900 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
MatSci-NLP: Evaluating Scientific Language Models on Materials Science Language Tasks Using Text-to-Schema Modeling 14 May 2023 · 1 repository · arXiv:2305.08264Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
GPT-Sentinel: Distinguishing Human and ChatGPT Generated Content 13 May 2023 · 2 repositories · arXiv:2305.07969
-
PESTS: Persian_English Cross Lingual Corpus for Semantic Textual Similarity 13 May 2023 · 0 repositories · arXiv:2305.07893
-
A General-Purpose Multilingual Document Encoder 11 May 2023 · 1 repository · arXiv:2305.07016
-
A Method to Automate the Discharge Summary Hospital Course for Neurology Patients 10 May 2023 · 0 repositories · arXiv:2305.06416
-
Enriching language models with graph-based context information to better understand textual data 10 May 2023 · 1 repository · arXiv:2305.11070
-
A Review of Vision-Language Models and their Performance on the Hateful Memes Challenge 9 May 2023 · 1 repository · arXiv:2305.06159
-
Alleviating Over-smoothing for Unsupervised Sentence Representation 9 May 2023 · 1 repository · arXiv:2305.06154Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples)
-
Attack Named Entity Recognition by Entity Boundary Interference 9 May 2023 · 0 repositories · arXiv:2305.05253
-
Detection of depression on social networks using transformers and ensembles 9 May 2023 · 1 repository · arXiv:2305.05325
-
Effects of sub-word segmentation on performance of transformer language models 9 May 2023 · 0 repositories · arXiv:2305.05480
-
StrAE: Autoencoding for Pre-Trained Embeddings using Explicit Structure 9 May 2023 · 0 repositories · arXiv:2305.05588
-
GersteinLab at MEDIQA-Chat 2023: Clinical Note Summarization from Doctor-Patient Conversations through Fine-tuning and In-context Learning 8 May 2023 · 0 repositories · arXiv:2305.05001
-
PreCog: Exploring the Relation between Memorization and Performance in Pre-trained Language Models 8 May 2023 · 0 repositories · arXiv:2305.04673
-
Vulnerability Detection Using Two-Stage Deep Learning Models 8 May 2023 · 0 repositories · arXiv:2305.09673
-
Stanford MLab at SemEval-2023 Task 10: Exploring GloVe- and Transformer-Based Methods for the Explainable Detection of Online Sexism 7 May 2023 · 0 repositories · arXiv:2305.04356
-
On the Usage of Continual Learning for Out-of-Distribution Generalization in Pre-trained Language Models of Code 6 May 2023 · 0 repositories · arXiv:2305.04106
-
Pre-training Language Model as a Multi-perspective Course Learner 6 May 2023 · 0 repositories · arXiv:2305.03981
-
Rhetorical Role Labeling of Legal Documents using Transformers and Graph Neural Networks 6 May 2023 · 0 repositories · arXiv:2305.04100
-
Block the Label and Noise: An N-Gram Masked Speller for Chinese Spell Checking 5 May 2023 · 0 repositories · arXiv:2305.03314
-
CLaC at SemEval-2023 Task 2: Comparing Span-Prediction and Sequence-Labeling approaches for NER 5 May 2023 · 0 repositories · arXiv:2305.03845
-
Harnessing the Power of BERT in the Turkish Clinical Domain: Pretraining Approaches for Limited Data Scenarios 5 May 2023 · 0 repositories · arXiv:2305.03788
-
Predicting COVID-19 and pneumonia complications from admission texts 5 May 2023 · 0 repositories · arXiv:2305.03661
-
Using ChatGPT for Entity Matching 5 May 2023 · 1 repository · arXiv:2305.03423
-
Enhancing Pashto Text Classification using Language Processing Techniques for Single And Multi-Label Analysis 4 May 2023 · 0 repositories · arXiv:2305.03201
-
Improving Code Example Recommendations on Informal Documentation Using BERT and Query-Aware LSH: A Comparative Study 4 May 2023 · 1 repository · arXiv:2305.03017
-
Leveraging BERT Language Model for Arabic Long Document Classification 4 May 2023 · 0 repositories · arXiv:2305.03519
-
A Novel Plagiarism Detection Approach Combining BERT-based Word Embedding, Attention-based LSTMs and an Improved Differential Evolution Algorithm 3 May 2023 · 0 repositories · arXiv:2305.02374
-
evaluating bert and parsbert for analyzing persian advertisement data 3 May 2023 · 0 repositories · arXiv:2305.02426
-
Evaluating BERT-based Scientific Relation Classifiers for Scholarly Knowledge Graph Construction on Digital Library Collections 3 May 2023 · 0 repositories · arXiv:2305.02291
-
Exploring Linguistic Properties of Monolingual BERTs with Typological Classification among Languages 3 May 2023 · 0 repositories · arXiv:2305.02215
-
Improving Cancer Hallmark Classification with BERT-based Deep Learning Approach 2 May 2023 · 0 repositories · arXiv:2305.03501
-
Unlimiformer: Long-Range Transformers with Unlimited Length Input 2 May 2023 · 2 repositories · arXiv:2305.01625Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Logion: Machine Learning for Greek Philology 1 May 2023 · 0 repositories · arXiv:2305.01099
-
Retrieving Comparative Arguments using Ensemble Methods and Neural Information Retrieval 1 May 2023 · 0 repositories · arXiv:2305.01513
-
SafeWebUH at SemEval-2023 Task 11: Learning Annotator Disagreement in Derogatory Text: Comparison of Direct Training vs Aggregation 1 May 2023 · 1 repository · arXiv:2305.01050
-
Are the Best Multilingual Document Embeddings simply Based on Sentence Embeddings? 28 Apr 2023 · 1 repository · arXiv:2304.14796
-
FlowTransformer: A Transformer Framework for Flow-based Network Intrusion Detection Systems 28 Apr 2023 · 1 repository · arXiv:2304.14746
-
Towards Better Domain Adaptation for Self-supervised Models: A Case Study of Child ASR 28 Apr 2023 · 1 repository · arXiv:2305.00115
-
Assessing Text Mining and Technical Analyses on Forecasting Financial Time Series 27 Apr 2023 · 0 repositories · arXiv:2304.14544
-
pyBibX -- A Python Library for Bibliometric and Scientometric Analysis Powered with Artificial Intelligence Tools 27 Apr 2023 · 1 repository · arXiv:2304.14516
-
Fine Tuning with Abnormal Examples 26 Apr 2023 · 0 repositories · arXiv:2304.13783
-
HausaNLP at SemEval-2023 Task 12: Leveraging African Low Resource TweetData for Sentiment Analysis 26 Apr 2023 · 1 repository · arXiv:2304.13634
-
Technical Report: Impact of Position Bias on Language Models in Token Classification 26 Apr 2023 · 2 repositories · arXiv:2304.13567
-
NLP-LTU at SemEval-2023 Task 10: The Impact of Data Augmentation and Semi-Supervised Learning Techniques on Text Classification Performance on an Imbalanced Dataset 25 Apr 2023 · 0 repositories · arXiv:2304.12847
-
What does BERT learn about prosody? 25 Apr 2023 · 0 repositories · arXiv:2304.12706
-
PARAGRAPH2GRAPH: A GNN-based framework for layout paragraph analysis 24 Apr 2023 · 1 repository · arXiv:2304.11810
-
Pre-trained Embeddings for Entity Resolution: An Experimental Analysis [Experiment, Analysis & Benchmark] 24 Apr 2023 · 1 repository · arXiv:2304.12329
-
SocialDial: A Benchmark for Socially-Aware Dialogue Systems 24 Apr 2023 · 1 repository · arXiv:2304.12026
-
Processing Natural Language on Embedded Devices: How Well Do Transformer Models Perform? 23 Apr 2023 · 2 repositories · arXiv:2304.11520
-
L3Cube-IndicSBERT: A simple approach for learning cross-lingual sentence representations using multilingual BERT 22 Apr 2023 · 0 repositories · arXiv:2304.11434
-
A Group-Specific Approach to NLP for Hate Speech Detection 21 Apr 2023 · 1 repository · arXiv:2304.11223
-
BERT Based Clinical Knowledge Extraction for Biomedical Knowledge Graph Construction and Analysis 21 Apr 2023 · 0 repositories · arXiv:2304.10996
-
Building Multimodal AI Chatbots 21 Apr 2023 · 1 repository · arXiv:2305.03512
-
Multi-Modal Deep Learning for Credit Rating Prediction Using Text and Numerical Data Streams 21 Apr 2023 · 1 repository · arXiv:2304.10740
-
Text2Time: Transformer-based Article Time Period Prediction 21 Apr 2023 · 0 repositories · arXiv:2304.10859
-
Domain-specific Continued Pretraining of Language Models for Capturing Long Context in Mental Health 20 Apr 2023 · 0 repositories · arXiv:2304.10447
-
Is Cross-modal Information Retrieval Possible without Training? 20 Apr 2023 · 0 repositories · arXiv:2304.11095
-
Movie Box Office Prediction With Self-Supervised and Visually Grounded Pretraining 20 Apr 2023 · 0 repositories · arXiv:2304.10311
-
Word Sense Induction with Knowledge Distillation from BERT 20 Apr 2023 · 0 repositories · arXiv:2304.10642
-
Scaling Transformer to 1M tokens and beyond with RMT 19 Apr 2023 · 3 repositories · arXiv:2304.11062Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
TieFake: Title-Text Similarity and Emotion-Aware Fake News Detection 19 Apr 2023 · 2 repositories · arXiv:2304.09421
-
Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling 18 Apr 2023 · 1 repository · arXiv:2304.09145Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Context-Dependent Embedding Utterance Representations for Emotion Recognition in Conversations 17 Apr 2023 · 1 repository · arXiv:2304.08216
-
InstructUIE: Multi-task Instruction Tuning for Unified Information Extraction 17 Apr 2023 · 1 repository · arXiv:2304.08085
-
Multimodal Short Video Rumor Detection System Based on Contrastive Learning 17 Apr 2023 · 0 repositories · arXiv:2304.08401
-
New Product Development (NPD) through Social Media-based Analysis by Comparing Word2Vec and BERT Word Embeddings 17 Apr 2023 · 0 repositories · arXiv:2304.08369
-
The MiniPile Challenge for Data-Efficient Language Models 17 Apr 2023 · 1 repository · arXiv:2304.08442
-
A Virtual Simulation-Pilot Agent for Training of Air Traffic Controllers 16 Apr 2023 · 0 repositories · arXiv:2304.07842
-
ArguGPT: evaluating, understanding and identifying argumentative essays generated by GPT models 16 Apr 2023 · 2 repositories · arXiv:2304.07666
-
Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models 15 Apr 2023 · 0 repositories · arXiv:2304.07619
-
SimpLex: a lexical text simplification architecture 14 Apr 2023 · 1 repository · arXiv:2304.07002
-
Automated Mapping of CVE Vulnerability Records to MITRE CWE Weaknesses 13 Apr 2023 · 0 repositories · arXiv:2304.11130
-
Evaluation of Social Biases in Recent Large Pre-Trained Models 13 Apr 2023 · 0 repositories · arXiv:2304.06861
-
Exploring the Use of Foundation Models for Named Entity Recognition and Lemmatization Tasks in Slavic Languages 11 Apr 2023 · 0 repositories · arXiv:2304.05336
-
Towards preserving word order importance through Forced Invalidation 11 Apr 2023 · 1 repository · arXiv:2304.05221
-
Incorporating Structured Sentences with Time-enhanced BERT for Fully-inductive Temporal Relation Prediction 10 Apr 2023 · 0 repositories · arXiv:2304.04717
-
Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study 10 Apr 2023 · 1 repository · arXiv:2304.04339Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Learning to Tokenize for Generative Retrieval 9 Apr 2023 · 1 repository · arXiv:2304.04171
-
Factify 2: A Multimodal Fake News and Satire News Dataset 8 Apr 2023 · 1 repository · arXiv:2304.03897
-
FlexMoE: Scaling Large-scale Sparse Pre-trained Model Training via Dynamic Device Placement 8 Apr 2023 · 0 repositories · arXiv:2304.03946
-
Interpretable Multi Labeled Bengali Toxic Comments Classification using Deep Learning 8 Apr 2023 · 1 repository · arXiv:2304.04087
-
Multi-class Categorization of Reasons behind Mental Disturbance in Long Texts 8 Apr 2023 · 0 repositories · arXiv:2304.04118
-
tmn at SemEval-2023 Task 9: Multilingual Tweet Intimacy Detection using XLM-T, Google Translate, and Ensemble Learning 8 Apr 2023 · 1 repository · arXiv:2304.04054
-
Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4 7 Apr 2023 · 1 repository · arXiv:2304.03439
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
Deep Learning for Opinion Mining and Topic Classification of Course Reviews 6 Apr 2023 · 0 repositories · arXiv:2304.03394
-
Micron-BERT: BERT-based Facial Micro-Expression Recognition 6 Apr 2023 · 1 repository · arXiv:2304.03195Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Multi-label classification of open-ended questions with BERT 6 Apr 2023 · 0 repositories · arXiv:2304.02945
-
Bengali Fake Review Detection using Semi-supervised Generative Adversarial Networks 5 Apr 2023 · 0 repositories · arXiv:2304.02739
-
Context-Aware Classification of Legal Document Pages 5 Apr 2023 · 0 repositories · arXiv:2304.02787
-
Improved Visual Fine-tuning with Natural Language Supervision 4 Apr 2023 · 1 repository · arXiv:2304.01489Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
San-BERT: Extractive Summarization for Sanskrit Documents using BERT and it's variants 4 Apr 2023 · 0 repositories · arXiv:2304.01894
-
Detection of Homophobia & Transphobia in Dravidian Languages: Exploring Deep Learning Methods 3 Apr 2023 · 0 repositories · arXiv:2304.01241
-
GreekBART: The First Pretrained Greek Sequence-to-Sequence Model 3 Apr 2023 · 2 repositories · arXiv:2304.00869
-
Hate Speech Targets Detection in Parler using BERT 3 Apr 2023 · 1 repository · arXiv:2304.01179
-
MiniRBT: A Two-stage Distilled Small Chinese Pre-trained Model 3 Apr 2023 · 1 repository · arXiv:2304.00717
-
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT 3 Apr 2023 · 2 repositories · arXiv:2304.01246
-
Classifying COVID-19 Related Tweets for Fake News Detection and Sentiment Analysis with BERT-based Models 2 Apr 2023 · 0 repositories · arXiv:2304.00636
-
The Other Side of Compression: Measuring Bias in Pruned Transformers 1 Apr 2023 · 1 repository