Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 30
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 30 of 71: papers 2,901 to 3,000 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BERTino: an Italian DistilBERT model 31 Mar 2023 · 1 repository · arXiv:2303.18121
-
Extracting Thyroid Nodules Characteristics from Ultrasound Reports Using Transformer-based Natural Language Processing Methods 31 Mar 2023 · 0 repositories · arXiv:2304.00115
-
JobHam-place with smart recommend job options and candidate filtering options 31 Mar 2023 · 0 repositories · arXiv:2303.17930
-
Quick Dense Retrievers Consume KALE: Post Training Kullback Leibler Alignment of Embeddings for Asymmetrical dual encoders 31 Mar 2023 · 0 repositories · arXiv:2304.01016
-
Evaluation of GPT and BERT-based models on identifying protein-protein interactions in biomedical text 30 Mar 2023 · 0 repositories · arXiv:2303.17728
-
Fine-Tuning BERT with Character-Level Noise for Zero-Shot Transfer to Dialects and Closely-Related Languages 30 Mar 2023 · 0 repositories · arXiv:2303.17683
-
oBERTa: Improving Sparse Transfer Learning via improved initialization, distillation, and pruning regimes 30 Mar 2023 · 0 repositories · arXiv:2303.17612
-
BERT4ETH: A Pre-trained Transformer for Ethereum Fraud Detection 29 Mar 2023 · 1 repository · arXiv:2303.18138
-
Larger Probes Tell a Different Story: Extending Psycholinguistic Datasets Via In-Context Learning 29 Mar 2023 · 1 repository · arXiv:2303.16445
-
TextMI: Textualize Multimodal Information for Integrating Non-verbal Cues in Pre-trained Language Models 27 Mar 2023 · 0 repositories · arXiv:2303.15430
-
Exploring Multimodal Sentiment Analysis via CBAM Attention and Double-layer BiLSTM Architecture 26 Mar 2023 · 0 repositories · arXiv:2303.14708
-
Indonesian Text-to-Image Synthesis with Sentence-BERT and FastGAN 25 Mar 2023 · 1 repository · arXiv:2303.14517
-
Spatio-Temporal driven Attention Graph Neural Network with Block Adjacency matrix (STAG-NN-BA) 25 Mar 2023 · 0 repositories · arXiv:2303.14322
-
Depression detection in social media posts using affective and social norm features 24 Mar 2023 · 0 repositories · arXiv:2303.14279
-
SIGMORPHON 2023 Shared Task of Interlinear Glossing: Baseline Model 24 Mar 2023 · 1 repository · arXiv:2303.14234
-
Toward Open-domain Slot Filling via Self-supervised Co-training 24 Mar 2023 · 0 repositories · arXiv:2303.13801
-
Where to Go Next for Recommender Systems? ID- vs. Modality-based Recommender Models Revisited 24 Mar 2023 · 1 repository · arXiv:2303.13835
-
A Novel Patent Similarity Measurement Methodology: Semantic Distance and Technological Distance 23 Mar 2023 · 1 repository · arXiv:2303.16767
-
Retrieval-Augmented Classification with Decoupled Representation 23 Mar 2023 · 1 repository · arXiv:2303.13065
-
Analyzing the Generalizability of Deep Contextualized Language Representations For Text Classification 22 Mar 2023 · 0 repositories · arXiv:2303.12936
-
Generate labeled training data using Prompt Programming and GPT-3. An example of Big Five Personality Classification 22 Mar 2023 · 0 repositories · arXiv:2303.12279
-
TRON: Transformer Neural Network Acceleration with Non-Coherent Silicon Photonics 22 Mar 2023 · 0 repositories · arXiv:2303.12914
-
Fine-tuning ClimateBert transformer with ClimaText for the disclosure analysis of climate-related financial risks 21 Mar 2023 · 0 repositories · arXiv:2303.13373
-
Is BERT Blind? Exploring the Effect of Vision-and-Language Pretraining on Visual Language Understanding 21 Mar 2023 · 1 repository · arXiv:2303.12513
-
Multimodal Pre-training Framework for Sequential Recommendation via Contrastive Learning 21 Mar 2023 · 0 repositories · arXiv:2303.11879
-
Character, Word, or Both? Revisiting the Segmentation Granularity for Chinese Pre-trained Language Models 20 Mar 2023 · 1 repository · arXiv:2303.10893
-
CTRAN: CNN-Transformer-based Network for Natural Language Understanding 19 Mar 2023 · 1 repository · arXiv:2303.10606
-
PACO: Provocation Involving Action, Culture, and Oppression 19 Mar 2023 · 0 repositories · arXiv:2303.12808
-
An Empirical Study of Pre-trained Language Models in Simple Knowledge Graph Question Answering 18 Mar 2023 · 1 repository · arXiv:2303.10368
-
NoisyHate: Mining Online Human-Written Perturbations for Realistic Robustness Benchmarking of Content Moderation Models 18 Mar 2023 · 0 repositories · arXiv:2303.10430
-
GADformer: A Transparent Transformer Model for Group Anomaly Detection on Trajectories 17 Mar 2023 · 1 repository · arXiv:2303.09841
-
Trained on 100 million words and still in shape: BERT meets British National Corpus 17 Mar 2023 · 2 repositories · arXiv:2303.09859Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Block-wise Bit-Compression of Transformer-based Models 16 Mar 2023 · 0 repositories · arXiv:2303.09184
-
Jump to Conclusions: Short-Cutting Transformers With Linear Transformations 16 Mar 2023 · 2 repositories · arXiv:2303.09435Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Measuring Improvement of F₁-Scores in Detection of Self-Admitted Technical Debt 16 Mar 2023 · 0 repositories · arXiv:2303.09617
-
SmartBERT: A Promotion of Dynamic Early Exiting Mechanism for Accelerating BERT Inference 16 Mar 2023 · 0 repositories · arXiv:2303.09266Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
Efficient Uncertainty Estimation with Gaussian Process for Reliable Dialog Response Retrieval 15 Mar 2023 · 0 repositories · arXiv:2303.08599
-
Do Transformers Parse while Predicting the Masked Word? 14 Mar 2023 · 0 repositories · arXiv:2303.08117
-
Features matching using natural language processing 14 Mar 2023 · 0 repositories · arXiv:2303.12804
-
Finding the Needle in a Haystack: Unsupervised Rationale Extraction from Long Text Classifiers 14 Mar 2023 · 0 repositories · arXiv:2303.07991
-
MEDBERT.de: A Comprehensive German BERT Model for the Medical Domain 14 Mar 2023 · 0 repositories · arXiv:2303.08179
-
Neuro-symbolic Commonsense Social Reasoning 14 Mar 2023 · 3 repositories · arXiv:2303.08264
-
Deep Learning Approach for Classifying the Aggressive Comments on Social Media: Machine Translated Data Vs Real Life Data 13 Mar 2023 · 0 repositories · arXiv:2303.07484
-
Transformer-based approaches to Sentiment Detection 13 Mar 2023 · 0 repositories · arXiv:2303.07292
-
LUKE-Graph: A Transformer-based Approach with Gated Relational Graph Attention for Cloze-style Reading Comprehension 12 Mar 2023 · 0 repositories · arXiv:2303.06675
-
Is In-hospital Meta-information Useful for Abstractive Discharge Summary Generation? 10 Mar 2023 · 0 repositories · arXiv:2303.06002
-
Research on CPI Prediction Based on Natural Language Processing 10 Mar 2023 · 0 repositories · arXiv:2303.05666
-
A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT 7 Mar 2023 · 1 repository · arXiv:2303.04226
-
ADELT: Transpilation Between Deep Learning Frameworks 7 Mar 2023 · 0 repositories · arXiv:2303.03593
-
Classifying Text-Based Conspiracy Tweets related to COVID-19 using Contextualized Word Embeddings 7 Mar 2023 · 0 repositories · arXiv:2303.03706
-
German BERT Model for Legal Named Entity Recognition 7 Mar 2023 · 0 repositories · arXiv:2303.05388
-
Gradient-Free Structured Pruning with Unlabeled Data 7 Mar 2023 · 0 repositories · arXiv:2303.04185
-
Video Question Answering Using CLIP-Guided Visual-Text Attention 6 Mar 2023 · 0 repositories · arXiv:2303.03131
-
Robust affine point matching via quadratic assignment on Grassmannians 5 Mar 2023 · 3 repositories · arXiv:2303.02698
-
Early Warning Signals of Social Instabilities in Twitter Data 3 Mar 2023 · 0 repositories · arXiv:2303.05401
-
Exploring Data Augmentation Methods on Social Media Corpora 3 Mar 2023 · 0 repositories · arXiv:2303.02198
-
Multi label classification of Artificial Intelligence related patents using Modified D2SBERT and Sentence Attention mechanism 3 Mar 2023 · 0 repositories · arXiv:2303.03165
-
Pre-trained Model Representations and their Robustness against Noise for Speech Emotion Analysis 3 Mar 2023 · 0 repositories · arXiv:2303.03177
-
TrojText: Test-time Invisible Textual Trojan Insertion 3 Mar 2023 · 1 repository · arXiv:2303.02242Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Will Affective Computing Emerge from Foundation Models and General AI? A First Evaluation on ChatGPT 3 Mar 2023 · 0 repositories · arXiv:2303.03186
-
Adopting the Multi-answer Questioning Task with an Auxiliary Metric for Extreme Multi-label Text Classification Utilizing the Label Hierarchy 2 Mar 2023 · 0 repositories · arXiv:2303.01064
-
Can BERT Refrain from Forgetting on Sequential Tasks? A Probing Study 2 Mar 2023 · 1 repository · arXiv:2303.01081
-
Evaluating Parameter-Efficient Transfer Learning Approaches on SURE Benchmark for Speech Understanding 2 Mar 2023 · 1 repository · arXiv:2303.03267
-
INO at Factify 2: Structure Coherence based Multi-Modal Fact Verification 2 Mar 2023 · 1 repository · arXiv:2303.01510
-
Sparse MoE as the New Dropout: Scaling Dense and Self-Slimmable Transformers 2 Mar 2023 · 1 repository · arXiv:2303.01610Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Competence-Based Analysis of Language Models 1 Mar 2023 · 0 repositories · arXiv:2303.00333
-
Domain-adapted large language models for classifying nuclear medicine reports 1 Mar 2023 · 0 repositories · arXiv:2303.01258
-
ToxVis: Enabling Interpretability of Implicit vs. Explicit Toxicity Detection Models with Interactive Visualization 1 Mar 2023 · 0 repositories · arXiv:2303.09402
-
Automatically Classifying Emotions based on Text: A Comparative Exploration of Different Datasets 28 Feb 2023 · 0 repositories · arXiv:2302.14727
-
Text classification dataset and analysis for Uzbek language 28 Feb 2023 · 1 repository · arXiv:2302.14494
-
Weighted Sampling for Masked Language Modeling 28 Feb 2023 · 0 repositories · arXiv:2302.14225
-
Elementwise Language Representation 27 Feb 2023 · 0 repositories · arXiv:2302.13475
-
Using Auxiliary Tasks In Multimodal Fusion Of Wav2vec 2.0 And BERT For Multimodal Emotion Recognition 27 Feb 2023 · 0 repositories · arXiv:2302.13661
-
Efficient Ensemble for Multimodal Punctuation Restoration using Time-Delay Neural Network 26 Feb 2023 · 1 repository · arXiv:2302.13376
-
Fast Attention Requires Bounded Entries 26 Feb 2023 · 0 repositories · arXiv:2302.13214
-
HULAT at SemEval-2023 Task 10: Data augmentation for pre-trained transformers applied to the detection of sexism in social media 24 Feb 2023 · 1 repository · arXiv:2302.12840
-
MUX-PLMs: Data Multiplexing for High-throughput Language Models 24 Feb 2023 · 1 repository · arXiv:2302.12441
-
Window transformer for dialogue document: a joint framework for causal emotion entailment 24 Feb 2023 · 0 repositories
-
Teacher Intervention: Improving Convergence of Quantization Aware Training for Ultra-Low Precision Transformers 23 Feb 2023 · 1 repository · arXiv:2302.11812Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Solution for the EPO CodeFest on Green Plastics: Hierarchical multi-label classification of patents relating to green plastics using deep learning 22 Feb 2023 · 0 repositories · arXiv:2302.13784
-
Boosting classification reliability of NLP transformer models in the long run 20 Feb 2023 · 0 repositories · arXiv:2302.10016
-
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey 20 Feb 2023 · 1 repository · arXiv:2302.10035
-
Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT 19 Feb 2023 · 1 repository · arXiv:2302.10198
-
Evaluating the Effectiveness of Pre-trained Language Models in Predicting the Helpfulness of Online Product Reviews 19 Feb 2023 · 1 repository · arXiv:2302.10199
-
Text Classification in the Wild: a Large-scale Long-tailed Name Normalization Dataset 19 Feb 2023 · 1 repository · arXiv:2302.09509
-
A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT 18 Feb 2023 · 0 repositories · arXiv:2302.09419
-
Hate Speech and Offensive Language Detection using an Emotion-aware Shared Encoder 17 Feb 2023 · 0 repositories · arXiv:2302.08777
-
ViTA: A Vision Transformer Inference Accelerator for Edge Applications 17 Feb 2023 · 0 repositories · arXiv:2302.09108
-
Foundation Models for Natural Language Processing -- Pre-trained Language Models Integrating Media 16 Feb 2023 · 0 repositories · arXiv:2302.08575
-
Marich: A Query-efficient Distributionally Equivalent Model Extraction Attack using Public Data 16 Feb 2023 · 1 repository · arXiv:2302.08466Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Retrieval-augmented Image Captioning 16 Feb 2023 · 1 repository · arXiv:2302.08268
-
Syntactic Structure Processing in the Brain while Listening 16 Feb 2023 · 0 repositories · arXiv:2302.08589
-
Commonsense Reasoning for Conversational AI: A Survey of the State of the Art 15 Feb 2023 · 0 repositories · arXiv:2302.07926
-
Towards Optimal Compression: Joint Pruning and Quantization 15 Feb 2023 · 0 repositories · arXiv:2302.07612
-
A Modern Look at the Relationship between Sharpness and Generalization 14 Feb 2023 · 1 repository · arXiv:2302.07011
-
A Psycholinguistic Analysis of BERT's Representations of Compounds 14 Feb 2023 · 1 repository · arXiv:2302.07232
-
Exploring Category Structure with Contextual Language Models and Lexical Semantic Networks 14 Feb 2023 · 0 repositories · arXiv:2302.06942
-
Few-shot learning approaches for classifying low resource domain specific software requirements 14 Feb 2023 · 0 repositories · arXiv:2302.06951
-
Reveal the Unknown: Out-of-Knowledge-Base Mention Discovery with Entity Linking 14 Feb 2023 · 3 repositories · arXiv:2302.07189
-
Linguistic ambiguity analysis in ChatGPT 13 Feb 2023 · 0 repositories · arXiv:2302.06426