Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 14
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 14 of 38: papers 1,301 to 1,400 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs 9 Apr 2024 · 0 repositories · arXiv:2404.07242
-
Guiding Large Language Models to Generate Computer-Parsable Content 8 Apr 2024 · 0 repositories · arXiv:2404.05499
-
Evaluating Interventional Reasoning Capabilities of Large Language Models 8 Apr 2024 · 0 repositories · arXiv:2404.05545
-
LTNER: Large Language Model Tagging for Named Entity Recognition with Contextualized Entity Marking 8 Apr 2024 · 0 repositories · arXiv:2404.05624
-
PetKaz at SemEval-2024 Task 3: Advancing Emotion Classification with an LLM for Emotion-Cause Pair Extraction in Conversations 8 Apr 2024 · 1 repository · arXiv:2404.05502
-
Physics of Language Models: Part 3.3, Knowledge Capacity Scaling Laws 8 Apr 2024 · 0 repositories · arXiv:2404.05405
-
Relation Extraction Using Large Language Models: A Case Study on Acupuncture Point Locations 8 Apr 2024 · 0 repositories · arXiv:2404.05415
-
A Multi-Level Framework for Accelerating Training Transformer Models 7 Apr 2024 · 1 repository · arXiv:2404.07999Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials 6 Apr 2024 · 1 repository · arXiv:2404.04510
-
RecGPT: Generative Personalized Prompts for Sequential Recommendation via ChatGPT Training Paradigm 6 Apr 2024 · 0 repositories · arXiv:2404.08675
-
Scope Ambiguities in Large Language Models 5 Apr 2024 · 1 repository · arXiv:2404.04332
-
Do Large Language Models Rank Fairly? An Empirical Study on the Fairness of LLMs as Rankers 4 Apr 2024 · 0 repositories · arXiv:2404.03192
-
NLP at UC Santa Cruz at SemEval-2024 Task 5: Legal Answer Validation using Few-Shot Multi-Choice QA 4 Apr 2024 · 1 repository · arXiv:2404.03150
-
AI-Tutoring in Software Engineering Education 3 Apr 2024 · 0 repositories · arXiv:2404.02548
-
An Incomplete Loop: Deductive, Inductive, and Abductive Learning in Large Language Models 3 Apr 2024 · 0 repositories · arXiv:2404.03028
-
Benchmarking Large Language Models for Persian: A Preliminary Study Focusing on ChatGPT 3 Apr 2024 · 1 repository · arXiv:2404.02403
-
GPT-DETOX: An In-Context Learning-Based Paraphraser for Text Detoxification 3 Apr 2024 · 0 repositories · arXiv:2404.03052
-
uTeBC-NLP at SemEval-2024 Task 9: Can LLMs be Lateral Thinkers? 3 Apr 2024 · 1 repository · arXiv:2404.02474Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Advancing LLM Reasoning Generalists with Preference Trees 2 Apr 2024 · 1 repository · arXiv:2404.02078Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 1 violated, 12 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 20 harvested samples) · 2 pointer-only (licence)
-
Stereotype Detection in LLMs: A Multiclass, Explainable, and Benchmark-Driven Approach 2 Apr 2024 · 0 repositories · arXiv:2404.01768
-
CMAT: A Multi-Agent Collaboration Tuning Framework for Enhancing Small Language Models 2 Apr 2024 · 1 repository · arXiv:2404.01663
-
Collapse of Self-trained Language Models 2 Apr 2024 · 1 repository · arXiv:2404.02305Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Comparative Study of Domain Driven Terms Extraction Using Large Language Models 2 Apr 2024 · 0 repositories · arXiv:2404.02330
-
Deconstructing In-Context Learning: Understanding Prompts via Corruption 2 Apr 2024 · 1 repository · arXiv:2404.02054
-
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks 2 Apr 2024 · 1 repository · arXiv:2404.02151Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
METAL: Towards Multilingual Meta-Evaluation 2 Apr 2024 · 0 repositories · arXiv:2404.01667
-
Scene Adaptive Sparse Transformer for Event-based Object Detection 2 Apr 2024 · 1 repository · arXiv:2404.01882Syntology official (archive's flag): 19 ran · 19 ran (of which 0 constructed an object rather than computing a result; 19 with no instrument failure: 0 honoured, 0 violated, 19 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 23 harvested samples)
-
SGSH: Stimulate Large Language Models with Skeleton Heuristics for Knowledge Base Question Generation 2 Apr 2024 · 1 repository · arXiv:2404.01923
-
Toward Informal Language Processing: Knowledge of Slang in Large Language Models 2 Apr 2024 · 1 repository · arXiv:2404.02323
-
Artificial Intelligence and the Spatial Documentation of Languages 1 Apr 2024 · 0 repositories · arXiv:2404.01263
-
BERT-Enhanced Retrieval Tool for Homework Plagiarism Detection System 1 Apr 2024 · 0 repositories · arXiv:2404.01582
-
FABLES: Evaluating faithfulness and content selection in book-length summarization 1 Apr 2024 · 3 repositories · arXiv:2404.01261Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Unveiling Divergent Inductive Biases of LLMs on Temporal Data 1 Apr 2024 · 1 repository · arXiv:2404.01453
-
CHOPS: CHat with custOmer Profile Systems for Customer Service with LLMs 31 Mar 2024 · 1 repository · arXiv:2404.01343
-
CoUDA: Coherence Evaluation via Unified Data Augmentation 31 Mar 2024 · 1 repository · arXiv:2404.00681
-
Training-Free Semantic Segmentation via LLM-Supervision 31 Mar 2024 · 0 repositories · arXiv:2404.00701
-
A Comprehensive Study on NLP Data Augmentation for Hate Speech Detection: Legacy Methods, BERT, and LLMs 30 Mar 2024 · 0 repositories · arXiv:2404.00303
-
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks 30 Mar 2024 · 0 repositories · arXiv:2404.00376
-
ChatGPT v.s. Media Bias: A Comparative Study of GPT-3.5 and Fine-tuned Language Models 29 Mar 2024 · 0 repositories · arXiv:2403.20158
-
DataAgent: Evaluating Large Language Models' Ability to Answer Zero-Shot, Natural Language Queries 29 Mar 2024 · 0 repositories · arXiv:2404.00188
-
ReALM: Reference Resolution As Language Modeling 29 Mar 2024 · 0 repositories · arXiv:2403.20329
-
A Review of Multi-Modal Large Language and Vision Models 28 Mar 2024 · 0 repositories · arXiv:2404.01322
-
FACTOID: FACtual enTailment fOr hallucInation Detection 28 Mar 2024 · 0 repositories · arXiv:2403.19113
-
Generating Multi-Aspect Queries for Conversational Search 28 Mar 2024 · 0 repositories · arXiv:2403.19302
-
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models 28 Mar 2024 · 1 repository · arXiv:2403.19521Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Just-DNA-Seq, open-source personal genomics platform: longevity science for everyone 28 Mar 2024 · 0 repositories · arXiv:2403.19087
-
A Survey on Large Language Models from Concept to Implementation 27 Mar 2024 · 0 repositories · arXiv:2403.18969
-
Evaluating Large Language Models for Health-Related Text Classification Tasks with Public Social Media Data 27 Mar 2024 · 0 repositories · arXiv:2403.19031
-
LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research Practices 27 Mar 2024 · 0 repositories · arXiv:2403.18173
-
Long-form factuality in large language models 27 Mar 2024 · 3 repositories · arXiv:2403.18802Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
ParCo: Part-Coordinating Text-to-Motion Synthesis 27 Mar 2024 · 1 repository · arXiv:2403.18512
-
Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers 27 Mar 2024 · 1 repository · arXiv:2403.18938
-
Vulnerability Detection with Code Language Models: How Far Are We? 27 Mar 2024 · 1 repository · arXiv:2403.18624Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Decoding Probing: Revealing Internal Linguistic Structures in Neural Language Models using Minimal Pairs 26 Mar 2024 · 0 repositories · arXiv:2403.17299
-
Disambiguate Entity Matching using Large Language Models through Relation Discovery 26 Mar 2024 · 0 repositories · arXiv:2403.17344
-
Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization 26 Mar 2024 · 1 repository · arXiv:2403.18120Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
ELLEN: Extremely Lightly Supervised Learning For Efficient Named Entity Recognition 26 Mar 2024 · 1 repository · arXiv:2403.17385
-
Enhancing Legal Document Retrieval: A Multi-Phase Approach with Large Language Models 26 Mar 2024 · 0 repositories · arXiv:2403.18093
-
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution 26 Mar 2024 · 0 repositories · arXiv:2403.17927
-
Not All Similarities Are Created Equal: Leveraging Data-Driven Biases to Inform GenAI Copyright Disputes 26 Mar 2024 · 0 repositories · arXiv:2403.17691
-
OmniVid: A Generative Framework for Universal Video Understanding 26 Mar 2024 · 1 repository · arXiv:2403.17935Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Targeted Visualization of the Backbone of Encoder LLMs 26 Mar 2024 · 1 repository · arXiv:2403.18872
-
Verbing Weirds Language (Models): Evaluation of English Zero-Derivation in Five LLMs 26 Mar 2024 · 0 repositories · arXiv:2403.17856
-
A comparison of Human, GPT-3.5, and GPT-4 Performance in a University-Level Coding Course 25 Mar 2024 · 1 repository · arXiv:2403.16977
-
Iterative Refinement of Project-Level Code Context for Precise Code Generation with Compiler Feedback 25 Mar 2024 · 1 repository · arXiv:2403.16792
-
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair 25 Mar 2024 · 1 repository · arXiv:2403.17134Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LinkPrompt: Natural and Universal Adversarial Attacks on Prompt-based Language Models 25 Mar 2024 · 1 repository · arXiv:2403.16432Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data 25 Mar 2024 · 1 repository · arXiv:2403.16909
-
SMTF: Sparse transformer with multiscale contextual fusion for medical image segmentation 24 Mar 2024 · 1 repository
-
SQL-Encoder: Improving NL2SQL In-Context Learning Through a Context-Aware Encoder 24 Mar 2024 · 0 repositories · arXiv:2403.16204
-
CodeShell Technical Report 23 Mar 2024 · 0 repositories · arXiv:2403.15747
-
Contact-aware Human Motion Generation from Textual Descriptions 23 Mar 2024 · 0 repositories · arXiv:2403.15709
-
Using Large Language Models for OntoClean-based Ontology Refinement 23 Mar 2024 · 0 repositories · arXiv:2403.15864
-
SOEN-101: Code Generation by Emulating Software Process Models Using Large Language Model Agents 23 Mar 2024 · 0 repositories · arXiv:2403.15852
-
Adapprox: Adaptive Approximation in Adam Optimization via Randomized Low-Rank Matrices 22 Mar 2024 · 0 repositories · arXiv:2403.14958
-
Can large language models explore in-context? 22 Mar 2024 · 0 repositories · arXiv:2403.15371
-
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation 22 Mar 2024 · 1 repository · arXiv:2403.14965
-
Measuring Gender and Racial Biases in Large Language Models 22 Mar 2024 · 0 repositories · arXiv:2403.15281
-
On Zero-Shot Counterspeech Generation by LLMs 22 Mar 2024 · 1 repository · arXiv:2403.14938
-
Optimal path for Biomedical Text Summarization Using Pointer GPT 22 Mar 2024 · 0 repositories · arXiv:2404.08654
-
Selecting Query-bag as Pseudo Relevance Feedback for Information-seeking Conversations 22 Mar 2024 · 0 repositories · arXiv:2404.04272
-
Text Clustering with Large Language Model Embeddings 22 Mar 2024 · 0 repositories · arXiv:2403.15112
-
Emergent World Models and Latent Variable Estimation in Chess-Playing Language Models 21 Mar 2024 · 1 repository · arXiv:2403.15498Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model 21 Mar 2024 · 1 repository · arXiv:2403.14598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding 21 Mar 2024 · 1 repository · arXiv:2403.14743
-
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving 20 Mar 2024 · 0 repositories · arXiv:2403.13331
-
AUD-TGN: Advancing Action Unit Detection with Temporal Convolution and GPT-2 in Wild Audiovisual Contexts 20 Mar 2024 · 0 repositories · arXiv:2403.13678
-
Incentivizing News Consumption on Social Media Platforms Using Large Language Models and Realistic Bot Accounts 20 Mar 2024 · 1 repository · arXiv:2403.13362
-
Motion Generation from Fine-grained Textual Descriptions 20 Mar 2024 · 1 repository · arXiv:2403.13518
-
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs 20 Mar 2024 · 0 repositories · arXiv:2403.13801
-
PARAMANU-AYN: Pretrain from scratch or Continual Pretraining of LLMs for Legal Domain Adaptation? 20 Mar 2024 · 0 repositories · arXiv:2403.13681
-
Automated Data Curation for Robust Language Model Fine-Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.12776
-
Can AI Outperform Human Experts in Creating Social Media Creatives? 19 Mar 2024 · 0 repositories · arXiv:2404.00018
-
Fine-Tuning Pre-trained Language Models to Detect In-Game Trash Talks 19 Mar 2024 · 0 repositories · arXiv:2403.15458
-
Instructing Large Language Models to Identify and Ignore Irrelevant Conditions 19 Mar 2024 · 1 repository · arXiv:2403.12744
-
Construction of Hyper-Relational Knowledge Graphs Using Pre-Trained Large Language Models 18 Mar 2024 · 0 repositories · arXiv:2403.11786
-
EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.12171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Embedded Named Entity Recognition using Probing Classifiers 18 Mar 2024 · 2 repositories · arXiv:2403.11747Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Embracing the Generative AI Revolution: Advancing Tertiary Education in Cybersecurity with GPT 18 Mar 2024 · 0 repositories · arXiv:2403.11402
-
Ensuring Safe and High-Quality Outputs: A Guideline Library Approach for Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11838