Methods › General › Learning Rate Schedules › Linear Warmup With Linear Decay › Papers, page 25
Linear Warmup With Linear Decay
Papers archive 2025-07-28
archive papers tagged: 7,076 · with a code link: 2,913 · where Syntology ran a sample: 650 (531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (650 of 7,076 tagged: 531 with a run with no instrument failure, 119 where every run was a failure of Syntology's instrument)
Page 25 of 71: papers 2,401 to 2,500 of 7,076, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fine-tuning ChatGPT for Automatic Scoring 16 Oct 2023 · 0 repositories · arXiv:2310.10072
-
Investigating Bias in Multilingual Language Models: Cross-Lingual Transfer of Debiasing Techniques 16 Oct 2023 · 1 repository · arXiv:2310.10310
-
Learning to Rank Context for Named Entity Recognition Using a Synthetic Dataset 16 Oct 2023 · 1 repository · arXiv:2310.10118
-
Domain-Specific Language Model Post-Training for Indonesian Financial NLP 15 Oct 2023 · 1 repository · arXiv:2310.09736
-
DPZero: Private Fine-Tuning of Language Models without Backpropagation 14 Oct 2023 · 1 repository · arXiv:2310.09639Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 9 unverified (of 17 harvested samples) · 3 pointer-only (licence)
-
Leveraging Generative AI: Improving Software Metadata Classification with Generated Code-Comment Pairs 14 Oct 2023 · 0 repositories · arXiv:2311.03365
-
Enhancing BERT-Based Visual Question Answering through Keyword-Driven Sentence Selection 13 Oct 2023 · 0 repositories · arXiv:2310.09432
-
Analyzing Textual Data for Fatality Classification in Afghanistan's Armed Conflicts: A BERT Approach 12 Oct 2023 · 0 repositories · arXiv:2310.08653
-
Detection and prediction of clopidogrel treatment failures using longitudinal structured electronic health records 12 Oct 2023 · 0 repositories · arXiv:2310.08757
-
Evaluating The Effectiveness of Capsule Neural Network in Toxic Comment Classification using Pre-trained BERT Embeddings 12 Oct 2023 · 1 repository
-
Expanding the Vocabulary of BERT for Knowledge Base Construction 12 Oct 2023 · 1 repository · arXiv:2310.08291
-
LEMON: Lossless model expansion 12 Oct 2023 · 1 repository · arXiv:2310.07999Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LLM-augmented Preference Learning from Natural Language 12 Oct 2023 · 0 repositories · arXiv:2310.08523
-
The Uncertainty-based Retrieval Framework for Ancient Chinese CWS and POS 12 Oct 2023 · 1 repository · arXiv:2310.08496
-
Fast-ELECTRA for Efficient Pre-training 11 Oct 2023 · 0 repositories · arXiv:2310.07347
-
Jaeger: A Concatenation-Based Multi-Transformer VQA Model 11 Oct 2023 · 0 repositories · arXiv:2310.07091
-
Toward Understanding BERT-Like Pre-Training for DNA Foundation Models 11 Oct 2023 · 0 repositories · arXiv:2310.07644
-
A Comparative Study of Transformer-based Neural Text Representation Techniques on Bug Triaging 10 Oct 2023 · 0 repositories · arXiv:2310.06913
-
GeoLLM: Extracting Geospatial Knowledge from Large Language Models 10 Oct 2023 · 1 repository · arXiv:2310.06213Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
GPT-4 as an Agronomist Assistant? Answering Agriculture Exams Using Large Language Models 10 Oct 2023 · 0 repositories · arXiv:2310.06225
-
Large Language Models for Propaganda Detection 10 Oct 2023 · 2 repositories · arXiv:2310.06422
-
Auditing Gender Analyzers on Text Data 9 Oct 2023 · 0 repositories · arXiv:2310.06061
-
Cabbage Sweeter than Cake? Analysing the Potential of Large Language Models for Learning Conceptual Spaces 9 Oct 2023 · 0 repositories · arXiv:2310.05481
-
Foundation Models Meet Visualizations: Challenges and Opportunities 9 Oct 2023 · 0 repositories · arXiv:2310.05771
-
Transformer Fusion with Optimal Transport 9 Oct 2023 · 1 repository · arXiv:2310.05719Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Breaking Down Word Semantics from Pre-trained Language Models through Layer-wise Dimension Selection 8 Oct 2023 · 0 repositories · arXiv:2310.05115
-
Enhancing Pre-Trained Language Models with Sentence Position Embeddings for Rhetorical Roles Recognition in Legal Opinions 8 Oct 2023 · 0 repositories · arXiv:2310.05276
-
LLM4VV: Developing LLM-Driven Testsuite for Compiler Validation 8 Oct 2023 · 1 repository · arXiv:2310.04963
-
RAC-BERT: Character Radical Enhanced BERT for Ancient Chinese 8 Oct 2023 · 0 repositories
-
A Process for Topic Modelling Via Word Embeddings 6 Oct 2023 · 0 repositories · arXiv:2312.03705
-
Automatic Aspect Extraction from Scientific Texts 6 Oct 2023 · 1 repository · arXiv:2310.04074
-
Quantized Transformer Language Model Implementations on Edge Devices 6 Oct 2023 · 0 repositories · arXiv:2310.03971
-
Segmented Harmonic Loss: Handling Class-Imbalanced Multi-Label Clinical Data for Medical Coding with Large Language Models 6 Oct 2023 · 0 repositories · arXiv:2310.04595
-
COVID-19 South African Vaccine Hesitancy Models Show Boost in Performance Upon Fine-Tuning on M-pox Tweets 4 Oct 2023 · 0 repositories · arXiv:2310.04453
-
Memoria: Resolving Fateful Forgetting Problem through Human-Inspired Memory Architecture 4 Oct 2023 · 1 repository · arXiv:2310.03052Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 10 harvested samples)
-
Retrieval-augmented Generation to Improve Math Question-Answering: Trade-offs Between Groundedness and Human Preference 4 Oct 2023 · 2 repositories · arXiv:2310.03184Syntology official (archive's flag): 7 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 20 harvested samples)
-
Harnessing Pre-Trained Sentence Transformers for Offensive Language Detection in Indian Languages 3 Oct 2023 · 0 repositories · arXiv:2310.02249
-
Label Supervised LLaMA Finetuning 2 Oct 2023 · 2 repositories · arXiv:2310.01208
-
Natural Language Models for Data Visualization Utilizing nvBench Dataset 2 Oct 2023 · 0 repositories · arXiv:2310.00832
-
Target-Aware Contextual Political Bias Detection in News 2 Oct 2023 · 0 repositories · arXiv:2310.01138
-
Question-Answering Model for Schizophrenia Symptoms and Their Impact on Daily Life using Mental Health Forums Data 30 Sep 2023 · 0 repositories · arXiv:2310.00448
-
RelBERT: Embedding Relations with Language Models 30 Sep 2023 · 1 repository · arXiv:2310.00299
-
Intuitive or Dependent? Investigating LLMs' Behavior Style to Conflicting Prompts 29 Sep 2023 · 0 repositories · arXiv:2309.17415
-
Hallucination Reduction in Long Input Text Summarization 28 Sep 2023 · 1 repository · arXiv:2309.16781
-
MKRAG: Medical Knowledge Retrieval Augmented Generation for Medical Question Answering 27 Sep 2023 · 0 repositories · arXiv:2309.16035
-
An NLP Benchmark Dataset for Assessing Corporate Climate Policy Engagement 26 Sep 2023 · 0 repositories
-
CAPP-130: A Corpus of Chinese Application Privacy Policy Summarization and Interpretation 26 Sep 2023 · 1 repository
-
Low-rank Adaptation of Large Language Model Rescoring for Parameter-Efficient Speech Recognition 26 Sep 2023 · 0 repositories · arXiv:2309.15223
-
RAGAS: Automated Evaluation of Retrieval Augmented Generation 26 Sep 2023 · 3 repositories · arXiv:2309.15217Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Comprehensive Overview of Named Entity Recognition: Models, Domain-Specific Applications and Challenges 25 Sep 2023 · 0 repositories · arXiv:2309.14084
-
Accelerating Large Batch Training via Gradient Signal to Noise Ratio (GSNR) 24 Sep 2023 · 0 repositories · arXiv:2309.13681
-
Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models 24 Sep 2023 · 1 repository · arXiv:2309.13579
-
Lexical Squad@Multimodal Hate Speech Event Detection 2023: Multimodal Hate Speech Detection using Fused Ensemble Approach 23 Sep 2023 · 1 repository · arXiv:2309.13354
-
AMPLIFY:Attention-based Mixup for Performance Improvement and Label Smoothing in Transformer 22 Sep 2023 · 1 repository · arXiv:2309.12689
-
TOPFORMER: Topology-Aware Authorship Attribution of Deepfake Texts with Diverse Writing Styles 22 Sep 2023 · 1 repository · arXiv:2309.12934
-
Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection 21 Sep 2023 · 1 repository · arXiv:2309.12247
-
BayesTune: Bayesian Sparse Deep Model Fine-tuning 21 Sep 2023 · 1 repository
-
Implicit Differentiable Outlier Detection Enable Robust Deep Multimodal Analysis 21 Sep 2023 · 1 repository
-
Making Scalable Meta Learning Practical 21 Sep 2023 · 1 repository
-
Marich: A Query-efficient Distributionally Equivalent Model Extraction Attack 21 Sep 2023 · 1 repository
-
On the Relationship between Skill Neurons and Robustness in Prompt Tuning 21 Sep 2023 · 1 repository · arXiv:2309.12263
-
[Re] Exploring the Role of Grammar and Word Choice in Bias Toward African American English (AAE) in Hate Speech Classification 21 Sep 2023 · 0 repositories
-
SLHCat: Mapping Wikipedia Categories and Lists to DBpedia by Leveraging Semantic, Lexical, and Hierarchical Features 21 Sep 2023 · 0 repositories · arXiv:2309.11791
-
SPICED: News Similarity Detection Dataset with Multiple Topics and Complexity Levels 21 Sep 2023 · 0 repositories · arXiv:2309.13080
-
Stock Market Sentiment Classification and Backtesting via Fine-tuned BERT 21 Sep 2023 · 0 repositories · arXiv:2309.11979
-
The Cambridge Law Corpus: A Dataset for Legal AI Research 21 Sep 2023 · 0 repositories · arXiv:2309.12269
-
Towards Efficient Pre-Trained Language Model via Feature Correlation Distillation 21 Sep 2023 · 0 repositories
-
AttentionMix: Data augmentation method that relies on BERT attention mechanism 20 Sep 2023 · 0 repositories · arXiv:2309.11104
-
CoT-BERT: Enhancing Unsupervised Sentence Representation through Chain-of-Thought 20 Sep 2023 · 2 repositories · arXiv:2309.11143
-
GPT-MolBERTa: GPT Molecular Features Language Model for molecular property prediction 20 Sep 2023 · 0 repositories · arXiv:2310.03030
-
Sequence-to-Sequence Spanish Pre-trained Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11259
-
Mixed-Distil-BERT: Code-mixed Language Modeling for Bangla, English, and Hindi 19 Sep 2023 · 0 repositories · arXiv:2309.10272
-
Facilitating NSFW Text Detection in Open-Domain Dialogue Systems via Knowledge Distillation 18 Sep 2023 · 1 repository · arXiv:2309.09749
-
Proposition from the Perspective of Chinese Language: A Chinese Proposition Classification Evaluation Benchmark 18 Sep 2023 · 0 repositories · arXiv:2309.09602
-
Detecting covariate drift in text data using document embeddings and dimensionality reduction 17 Sep 2023 · 1 repository · arXiv:2309.10000
-
SplitEE: Early Exit in Deep Neural Networks with Split Computing 17 Sep 2023 · 1 repository · arXiv:2309.09195
-
Has Sentiment Returned to the Pre-pandemic Level? A Sentiment Analysis Using U.S. College Subreddit Data from 2019 to 2022 16 Sep 2023 · 1 repository · arXiv:2309.08845
-
AlbNER: A Corpus for Named Entity Recognition in Albanian 15 Sep 2023 · 0 repositories · arXiv:2309.08741
-
Detecting Relevant Information in High-Volume Chat Logs: Keyphrase Extraction for Grooming and Drug Dealing Forensic Analysis 15 Sep 2023 · 0 repositories · arXiv:2311.04905
-
Structural Self-Supervised Objectives for Transformers 15 Sep 2023 · 1 repository · arXiv:2309.08272
-
Transformer Based Punctuation Restoration for Turkish 15 Sep 2023 · 1 repository
-
VulnSense: Efficient Vulnerability Detection in Ethereum Smart Contracts by Multimodal Learning with Graph Neural Network and Language Model 15 Sep 2023 · 0 repositories · arXiv:2309.08474
-
Automatic Data Visualization Generation from Chinese Natural Language Questions 14 Sep 2023 · 0 repositories · arXiv:2309.07650
-
DebCSE: Rethinking Unsupervised Contrastive Sentence Embedding Learning in the Debiasing Perspective 14 Sep 2023 · 0 repositories · arXiv:2309.07396
-
EnCodecMAE: Leveraging neural codecs for universal audio representation learning 14 Sep 2023 · 2 repositories · arXiv:2309.07391Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Text Classification of Cancer Clinical Trial Eligibility Criteria 14 Sep 2023 · 0 repositories · arXiv:2309.07812
-
Balanced and Explainable Social Media Analysis for Public Health with Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05951
-
PRESTI: Predicting Repayment Effort of Self-Admitted Technical Debt Using Textual Information 12 Sep 2023 · 0 repositories · arXiv:2309.06020
-
Overview of Memotion 3: Sentiment and Emotion Analysis of Codemixed Hinglish Memes 12 Sep 2023 · 0 repositories · arXiv:2309.06517
-
Applying BioBERT to Extract Germline Gene-Disease Associations for Building a Knowledge Graph from the Biomedical Literature 11 Sep 2023 · 1 repository · arXiv:2309.13061
-
CrisisTransformers: Pre-trained language models and sentence encoders for crisis-related social media texts 11 Sep 2023 · 0 repositories · arXiv:2309.05494
-
Detecting Natural Language Biases with Prompt-based Learning 11 Sep 2023 · 0 repositories · arXiv:2309.05227
-
Learning Personalized User Preference from Cold Start in Multi-turn Conversations 10 Sep 2023 · 0 repositories · arXiv:2309.05127
-
Neural-Hidden-CRF: A Robust Weakly-Supervised Sequence Labeler 10 Sep 2023 · 1 repository · arXiv:2309.05086
-
RGAT: A Deeper Look into Syntactic Dependency Information for Coreference Resolution 10 Sep 2023 · 0 repositories · arXiv:2309.04977
-
Encoding Multi-Domain Scientific Papers by Ensembling Multiple CLS Tokens 8 Sep 2023 · 1 repository · arXiv:2309.04333Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Fuzzy Fingerprinting Transformer Language-Models for Emotion Recognition in Conversations 8 Sep 2023 · 0 repositories · arXiv:2309.04292
-
Leveraging Pretrained Image-text Models for Improving Audio-Visual Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04628
-
UQ at #SMM4H 2023: ALEX for Public Health Analysis with Social Media 8 Sep 2023 · 1 repository · arXiv:2309.04213
-
Certifying LLM Safety against Adversarial Prompting 6 Sep 2023 · 1 repository · arXiv:2309.02705