Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 12
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 12 of 38: papers 1,101 to 1,200 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction 2 Jun 2024 · 1 repository · arXiv:2406.00755Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
FOCUS: Forging Originality through Contrastive Use in Self-Plagiarism for Language Models 2 Jun 2024 · 0 repositories · arXiv:2406.00839
-
Pretrained Hybrids with MAD Skills 2 Jun 2024 · 0 repositories · arXiv:2406.00894
-
Domain-specific ReAct for physics-integrated iterative modeling: A case study of LLM agents for gas path analysis of gas turbines 1 Jun 2024 · 0 repositories · arXiv:2406.07572
-
An Evaluation Benchmark for Autoformalization in Lean4 1 Jun 2024 · 0 repositories · arXiv:2406.06555
-
Beyond Metrics: Evaluating LLMs' Effectiveness in Culturally Nuanced, Low-Resource Real-World Scenarios 1 Jun 2024 · 0 repositories · arXiv:2406.00343
-
Generative AI Voting: Fair Collective Choice is Resilient to LLM Biases and Inconsistencies 31 May 2024 · 1 repository · arXiv:2406.11871
-
Hard Cases Detection in Motion Prediction by Vision-Language Foundation Models 31 May 2024 · 1 repository · arXiv:2405.20991
-
Large Language Models are Zero-Shot Next Location Predictors 31 May 2024 · 1 repository · arXiv:2405.20962Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
LOLAMEME: Logic, Language, Memory, Mechanistic Framework 31 May 2024 · 0 repositories · arXiv:2406.02592
-
Multilingual Text Style Transfer: Datasets & Models for Indian Languages 31 May 2024 · 2 repositories · arXiv:2405.20805
-
The Point of View of a Sentiment: Towards Clinician Bias Detection in Psychiatric Notes 31 May 2024 · 0 repositories · arXiv:2405.20582
-
ANAH: Analytical Annotation of Hallucinations in Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20315Syntology official: no sample here; runs from other or unrecorded repositories · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
AutoBreach: Universal and Adaptive Jailbreaking with Efficient Wordplay-Guided Optimization 30 May 2024 · 0 repositories · arXiv:2405.19668
-
Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation 30 May 2024 · 0 repositories · arXiv:2405.20092
-
Knowledge Graph Tuning: Real-time Large Language Model Personalization based on Human Feedback 30 May 2024 · 0 repositories · arXiv:2405.19686
-
LLaMEA: A Large Language Model Evolutionary Algorithm for Automatically Generating Metaheuristics 30 May 2024 · 2 repositories · arXiv:2405.20132Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Phantom: General Trigger Attacks on Retrieval Augmented Language Generation 30 May 2024 · 0 repositories · arXiv:2405.20485
-
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs 30 May 2024 · 0 repositories · arXiv:2405.20179
-
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation 30 May 2024 · 0 repositories · arXiv:2405.19701
-
Towards Ontology-Enhanced Representation Learning for Large Language Models 30 May 2024 · 1 repository · arXiv:2405.20527
-
A Multi-Source Retrieval Question Answering Framework Based on RAG 29 May 2024 · 0 repositories · arXiv:2405.19207
-
Beyond Agreement: Diagnosing the Rationale Alignment of Automated Essay Scoring Methods based on Linguistically-informed Counterfactuals 29 May 2024 · 1 repository · arXiv:2405.19433
-
Can GPT Redefine Medical Understanding? Evaluating GPT on Biomedical Machine Reading Comprehension 29 May 2024 · 0 repositories · arXiv:2405.18682
-
Efficient Model-agnostic Alignment via Bayesian Persuasion 29 May 2024 · 0 repositories · arXiv:2405.18718
-
LMO-DP: Optimizing the Randomization Mechanism for Differentially Private Fine-Tuning (Large) Language Models 29 May 2024 · 0 repositories · arXiv:2405.18776
-
MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series 29 May 2024 · 1 repository · arXiv:2405.19327Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Aligning to Thousands of Preferences via System Message Generalization 28 May 2024 · 1 repository · arXiv:2405.17977Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
An Empirical Analysis on Large Language Models in Debate Evaluation 28 May 2024 · 1 repository · arXiv:2406.00050Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Are PPO-ed Language Models Hackable? 28 May 2024 · 0 repositories · arXiv:2406.02577
-
Edinburgh Clinical NLP at MEDIQA-CORR 2024: Guiding Large Language Models with Hints 28 May 2024 · 0 repositories · arXiv:2405.18028
-
LLMs and Memorization: On Quality and Specificity of Copyright Compliance 28 May 2024 · 1 repository · arXiv:2405.18492
-
Understanding Intrinsic Socioeconomic Biases in Large Language Models 28 May 2024 · 0 repositories · arXiv:2405.18662
-
Assessing LLMs Suitability for Knowledge Graph Completion 27 May 2024 · 1 repository · arXiv:2405.17249
-
InversionView: A General-Purpose Method for Reading Information from Neural Activations 27 May 2024 · 1 repository · arXiv:2405.17653Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified; the one sample that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
REVECA: Adaptive Planning and Trajectory-based Validation in Cooperative Language Agents using Information Relevance and Relative Proximity 27 May 2024 · 0 repositories · arXiv:2405.16751
-
Performance evaluation of Reddit Comments using Machine Learning and Natural Language Processing methods in Sentiment Analysis 27 May 2024 · 0 repositories · arXiv:2405.16810
-
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation 27 May 2024 · 1 repository · arXiv:2405.17057Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
RTL-Repo: A Benchmark for Evaluating LLMs on Large-Scale RTL Design Projects 27 May 2024 · 1 repository · arXiv:2405.17378
-
The Scaling Law in Stellar Light Curves 27 May 2024 · 0 repositories · arXiv:2405.17156
-
THREAD: Thinking Deeper with Recursive Spawning 27 May 2024 · 1 repository · arXiv:2405.17402Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Accelerating Transformers with Spectrum-Preserving Token Merging 25 May 2024 · 1 repository · arXiv:2405.16148Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
AutoManual: Constructing Instruction Manuals by LLM Agents via Interactive Environmental Learning 25 May 2024 · 1 repository · arXiv:2405.16247Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Incremental Comprehension of Garden-Path Sentences by Large Language Models: Semantic Interpretation, Syntactic Re-Analysis, and Attention 25 May 2024 · 0 repositories · arXiv:2405.16042
-
MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time 25 May 2024 · 0 repositories · arXiv:2405.16265
-
An Evaluation of Estimative Uncertainty in Large Language Models 24 May 2024 · 0 repositories · arXiv:2405.15185
-
Benchmarking the Performance of Pre-trained LLMs across Urdu NLP Tasks 24 May 2024 · 0 repositories · arXiv:2405.15453
-
CulturePark: Boosting Cross-cultural Understanding in Large Language Models 24 May 2024 · 1 repository · arXiv:2405.15145Syntology official: harvested, nothing ran · 0 ran · 7 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Evaluating and Safeguarding the Adversarial Robustness of Retrieval-Based In-Context Learning 24 May 2024 · 1 repository · arXiv:2405.15984Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Generalizable and Scalable Multistage Biomedical Concept Normalization Leveraging Large Language Models 24 May 2024 · 1 repository · arXiv:2405.15122
-
GPT is Not an Annotator: The Necessity of Human Annotation in Fairness Benchmark Construction 24 May 2024 · 0 repositories · arXiv:2405.15760
-
Learning the Language of Protein Structure 24 May 2024 · 1 repository · arXiv:2405.15840Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
The Impact and Opportunities of Generative AI in Fact-Checking 24 May 2024 · 0 repositories · arXiv:2405.15985
-
The Buffer Mechanism for Multi-Step Information Reasoning in Language Models 24 May 2024 · 0 repositories · arXiv:2405.15302
-
EditWorld: Simulating World Dynamics for Instruction-Following Image Editing 23 May 2024 · 1 repository · arXiv:2405.14785Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Eliciting Informative Text Evaluations with Large Language Models 23 May 2024 · 1 repository · arXiv:2405.15077Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step 23 May 2024 · 1 repository · arXiv:2405.14838Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Large Language Models Can Self-Correct with Key Condition Verification 23 May 2024 · 0 repositories · arXiv:2405.14092
-
Not All Language Model Features Are Linear 23 May 2024 · 1 repository · arXiv:2405.14860Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
WISE: Rethinking the Knowledge Memory for Lifelong Model Editing of Large Language Models 23 May 2024 · 1 repository · arXiv:2405.14768Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 9 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Automatically Identifying Local and Global Circuits with Linear Computation Graphs 22 May 2024 · 0 repositories · arXiv:2405.13868
-
Evaluating Large Language Models with Human Feedback: Establishing a Swedish Benchmark 22 May 2024 · 1 repository · arXiv:2405.14006
-
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR 22 May 2024 · 0 repositories · arXiv:2406.00014
-
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment 22 May 2024 · 1 repository · arXiv:2405.13911Syntology official (archive's flag): 3 ran · 6 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities 21 May 2024 · 0 repositories · arXiv:2405.12750
-
How Reliable AI Chatbots are for Disease Prediction from Patient Complaints? 21 May 2024 · 0 repositories · arXiv:2405.13219
-
Investigating Persuasion Techniques in Arabic: An Empirical Study Leveraging Large Language Models 21 May 2024 · 0 repositories · arXiv:2405.12884
-
Quantifying Semantic Emergence in Language Models 21 May 2024 · 1 repository · arXiv:2405.12617Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 6 harvested samples)
-
Evaluating and Modeling Social Intelligence: A Comparative Study of Human and AI Capabilities 20 May 2024 · 1 repository · arXiv:2405.11841
-
DaVinci at SemEval-2024 Task 9: Few-shot prompting GPT-3.5 for Unconventional Reasoning 19 May 2024 · 0 repositories · arXiv:2405.11559
-
Human-Centered LLM-Agent User Interface: A Position Paper 19 May 2024 · 1 repository · arXiv:2405.13050
-
Your Transformer is Secretly Linear 19 May 2024 · 1 repository · arXiv:2405.12250Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Zero-Shot Stance Detection using Contextual Data Generation with LLMs 19 May 2024 · 1 repository · arXiv:2405.11637Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Cross-Language Assessment of Mathematical Capability of ChatGPT 18 May 2024 · 0 repositories · arXiv:2405.11264
-
Evaluation of large language model performance on the Biomedical Language Understanding and Reasoning Benchmark 17 May 2024 · 0 repositories
-
Language Models can Exploit Cross-Task In-context Learning for Data-Scarce Novel Tasks 17 May 2024 · 1 repository · arXiv:2405.10548Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
FinTextQA: A Dataset for Long-form Financial Question Answering 16 May 2024 · 0 repositories · arXiv:2405.09980
-
GPT Store Mining and Analysis 16 May 2024 · 0 repositories · arXiv:2405.10210
-
HW-GPT-Bench: Hardware-Aware Architecture Benchmark for Language Models 16 May 2024 · 2 repositories · arXiv:2405.10299
-
Optimization Techniques for Sentiment Analysis Based on LLM (GPT-3) 16 May 2024 · 0 repositories · arXiv:2405.09770
-
Matching domain experts by training from scratch on domain knowledge 15 May 2024 · 0 repositories · arXiv:2405.09395
-
Beyond Scaling Laws: Understanding Transformer Performance with Associative Memory 14 May 2024 · 0 repositories · arXiv:2405.08707
-
Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis 14 May 2024 · 0 repositories · arXiv:2405.08944
-
GPT-3.5 for Grammatical Error Correction 14 May 2024 · 0 repositories · arXiv:2405.08469
-
Refinement of an Epilepsy Dictionary through Human Annotation of Health-related posts on Instagram 14 May 2024 · 0 repositories · arXiv:2405.08784
-
Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control 14 May 2024 · 0 repositories · arXiv:2405.08366
-
When Large Language Models Meet Optical Networks: Paving the Way for Automation 14 May 2024 · 0 repositories · arXiv:2405.17441
-
Can Language Models Explain Their Own Classification Behavior? 13 May 2024 · 1 repository · arXiv:2405.07436
-
Coding historical causes of death data with Large Language Models 13 May 2024 · 1 repository · arXiv:2405.07560
-
FreeVA: Offline MLLM as Training-Free Video Assistant 13 May 2024 · 1 repository · arXiv:2405.07798Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
MacBehaviour: An R package for behavioural experimentation on large language models 13 May 2024 · 1 repository · arXiv:2405.07495
-
Many-Shot Regurgitation (MSR) Prompting 13 May 2024 · 0 repositories · arXiv:2405.08134
-
Open-vocabulary Auditory Neural Decoding Using fMRI-prompted LLM 13 May 2024 · 0 repositories · arXiv:2405.07840
-
Learning Reward for Robot Skills Using Large Language Models via Self-Alignment 12 May 2024 · 0 repositories · arXiv:2405.07162
-
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis 12 May 2024 · 1 repository · arXiv:2405.07248Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Quite Good, but Not Enough: Nationality Bias in Large Language Models -- A Case Study of ChatGPT 11 May 2024 · 1 repository · arXiv:2405.06996
-
RETTA: Retrieval-Enhanced Test-Time Adaptation for Zero-Shot Video Captioning 11 May 2024 · 0 repositories · arXiv:2405.07046
-
An Assessment of Model-On-Model Deception 10 May 2024 · 0 repositories · arXiv:2405.12999
-
ChatGPTest: opportunities and cautionary tales of utilizing AI for questionnaire pretesting 10 May 2024 · 0 repositories · arXiv:2405.06329
-
A Mixture of Experts Approach to 3D Human Motion Prediction 9 May 2024 · 1 repository · arXiv:2405.06088