Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 3
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 3 of 38: papers 201 to 300 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Rethinking Prompt-based Debiasing in Large Language Models 12 Mar 2025 · 0 repositories · arXiv:2503.09219
-
VaxGuard: A Multi-Generator, Multi-Type, and Multi-Role Dataset for Detecting LLM-Generated Vaccine Misinformation 12 Mar 2025 · 0 repositories · arXiv:2503.09103
-
VLog: Video-Language Models by Generative Retrieval of Narration Vocabulary 12 Mar 2025 · 1 repository · arXiv:2503.09402Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Context-aware Biases for Length Extrapolation 11 Mar 2025 · 1 repository · arXiv:2503.08067Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
GPT-PPG: A GPT-based Foundation Model for Photoplethysmography Signals 11 Mar 2025 · 0 repositories · arXiv:2503.08015
-
Interpretable and Robust Dialogue State Tracking via Natural Language Summarization with LLMs 11 Mar 2025 · 0 repositories · arXiv:2503.08857
-
Llms, Virtual Users, and Bias: Predicting Any Survey Question Without Human Data 11 Mar 2025 · 0 repositories · arXiv:2503.16498
-
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings 10 Mar 2025 · 0 repositories · arXiv:2503.06980
-
Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning 10 Mar 2025 · 0 repositories · arXiv:2503.07591
-
Fully Autonomous Programming using Iterative Multi-Agent Debugging with Large Language Models 10 Mar 2025 · 0 repositories · arXiv:2503.07693
-
Implicit Reasoning in Transformers is Reasoning through Shortcuts 10 Mar 2025 · 1 repository · arXiv:2503.07604Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Improving cognitive diagnostics in pathology: a deep learning approach for augmenting perceptional understanding of histopathology images 10 Mar 2025 · 0 repositories · arXiv:2503.06894
-
MapQA: Open-domain Geospatial Question Answering on Map Data 10 Mar 2025 · 0 repositories · arXiv:2503.07871
-
Taking Notes Brings Focus? Towards Multi-Turn Multimodal Dialogue Learning 10 Mar 2025 · 0 repositories · arXiv:2503.07002
-
Effectiveness of Zero-shot-CoT in Japanese Prompts 9 Mar 2025 · 0 repositories · arXiv:2503.06765
-
UniGenX: Unified Generation of Sequence and Structure with Autoregressive Diffusion 9 Mar 2025 · 0 repositories · arXiv:2503.06687
-
LimTopic: LLM-based Topic Modeling and Text Summarization for Analyzing Scientific Articles limitations 8 Mar 2025 · 1 repository · arXiv:2503.10658
-
Evaluating Large Language Models in Code Generation: INFINITE Methodology for Defining the Inference Index 7 Mar 2025 · 0 repositories · arXiv:2503.05852
-
Simulating and Analysing Human Survey Responses with Large Language Models: A Case Study in Energy Stated Preference 7 Mar 2025 · 0 repositories · arXiv:2503.10652
-
Zero-shot Medical Event Prediction Using a Generative Pre-trained Transformer on Electronic Health Records 7 Mar 2025 · 0 repositories · arXiv:2503.05893
-
Compositional Causal Reasoning Evaluation in Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04556
-
HILGEN: Hierarchically-Informed Data Generation for Biomedical NER Using Knowledgebases and Large Language Models 6 Mar 2025 · 0 repositories · arXiv:2503.04930
-
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation 5 Mar 2025 · 1 repository · arXiv:2503.03308
-
Effectively Steer LLM To Follow Preference via Building Confident Directions 4 Mar 2025 · 0 repositories · arXiv:2503.02989
-
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models 4 Mar 2025 · 0 repositories · arXiv:2503.02917
-
Use Me Wisely: AI-Driven Assessment for LLM Prompting Skills Development 4 Mar 2025 · 0 repositories · arXiv:2503.02532
-
Weak-to-Strong Generalization Even in Random Feature Networks, Provably 4 Mar 2025 · 0 repositories · arXiv:2503.02877
-
Zero-Shot Multi-Label Classification of Bangla Documents: Large Decoders Vs. Classic Encoders 4 Mar 2025 · 0 repositories · arXiv:2503.02993
-
Machine Learners Should Acknowledge the Legal Implications of Large Language Models as Personal Data 3 Mar 2025 · 0 repositories · arXiv:2503.01630
-
LLMInit: A Free Lunch from Large Language Models for Selective Initialization of Recommendation 3 Mar 2025 · 0 repositories · arXiv:2503.01814
-
Attention Condensation via Sparsity Induced Regularized Training 3 Mar 2025 · 0 repositories · arXiv:2503.01564
-
Cancer Type, Stage and Prognosis Assessment from Pathology Reports using LLMs 3 Mar 2025 · 1 repository · arXiv:2503.01194
-
Label Ranker: Self-Aware Preference for Classification Label Position in Visual Masked Self-Supervised Pre-Trained Model 3 Mar 2025 · 1 repository
-
Using (Not so) Large Language Models for Generating Simulation Models in a Formal DSL -- A Study on Reaction Networks 3 Mar 2025 · 0 repositories · arXiv:2503.01675
-
SemViQA: A Semantic Question Answering System for Vietnamese Information Fact-Checking 2 Mar 2025 · 1 repository · arXiv:2503.00955
-
Psychological Counseling Ability of Large Language Models 1 Mar 2025 · 0 repositories · arXiv:2503.07627
-
Beyond Words: A Latent Memory Approach to Internal Reasoning in LLMs 28 Feb 2025 · 0 repositories · arXiv:2502.21030
-
CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation 28 Feb 2025 · 1 repository · arXiv:2502.21074Syntology 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
NutriGen: Personalized Meal Plan Generator Leveraging Large Language Models to Enhance Dietary and Nutritional Adherence 28 Feb 2025 · 1 repository · arXiv:2502.20601
-
SeisMoLLM: Advancing Seismic Monitoring via Cross-modal Transfer with Pre-trained Large Language Model 27 Feb 2025 · 1 repository · arXiv:2502.19960Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Cognitive networks highlight differences and similarities in the STEM mindsets of human and LLM-simulated trainees, experts and academics 26 Feb 2025 · 0 repositories · arXiv:2502.19529
-
Deep-Bench: Deep Learning Benchmark Dataset for Code Generation 26 Feb 2025 · 0 repositories · arXiv:2502.18726
-
The Sharpness Disparity Principle in Transformers for Accelerating Language Model Pre-Training 26 Feb 2025 · 0 repositories · arXiv:2502.19002
-
Assessing Large Language Models in Agentic Multilingual National Bias 25 Feb 2025 · 0 repositories · arXiv:2502.17945
-
Independent Mobility GPT (IDM-GPT): A Self-Supervised Multi-Agent Large Language Model Framework for Customized Traffic Mobility Analysis Using Machine Learning Models 25 Feb 2025 · 0 repositories · arXiv:2502.18652
-
Scaling LLM Pre-training with Vocabulary Curriculum 25 Feb 2025 · 0 repositories · arXiv:2502.17910
-
Applying LLMs to Active Learning: Towards Cost-Efficient Cross-Task Text Classification without Manually Labeled Data 24 Feb 2025 · 0 repositories · arXiv:2502.16892
-
Are Large Language Models Good Data Preprocessors? 24 Feb 2025 · 0 repositories · arXiv:2502.16790
-
Reasoning about Affordances: Causal and Compositional Reasoning in LLMs 23 Feb 2025 · 0 repositories · arXiv:2502.16606
-
A Close Look at Decomposition-based XAI-Methods for Transformer Language Models 21 Feb 2025 · 2 repositories · arXiv:2502.15886
-
BP-GPT: Auditory Neural Decoding Using fMRI-prompted LLM 21 Feb 2025 · 1 repository · arXiv:2502.15172Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Single-pass Detection of Jailbreaking Input in Large Language Models 21 Feb 2025 · 0 repositories · arXiv:2502.15435
-
DeepRTL: Bridging Verilog Understanding and Generation with a Unified Representation Model 20 Feb 2025 · 0 repositories · arXiv:2502.15832
-
Entropy-UID: A Method for Optimizing Information Density 20 Feb 2025 · 0 repositories · arXiv:2502.14366
-
Hallucination Detection in Large Language Models with Metamorphic Relations 20 Feb 2025 · 0 repositories · arXiv:2502.15844
-
Giving AI Personalities Leads to More Human-Like Reasoning 19 Feb 2025 · 0 repositories · arXiv:2502.14155
-
Hidden Darkness in LLM-Generated Designs: Exploring Dark Patterns in Ecommerce Web Components Generated by LLMs 19 Feb 2025 · 0 repositories · arXiv:2502.13499
-
PitVQA++: Vector Matrix-Low-Rank Adaptation for Open-Ended Visual Question Answering in Pituitary Surgery 19 Feb 2025 · 1 repository · arXiv:2502.14149
-
RGAR: Recurrence Generation-augmented Retrieval for Factual-aware Medical Question Answering 19 Feb 2025 · 0 repositories · arXiv:2502.13361
-
An LLM-Powered Agent for Physiological Data Analysis: A Case Study on PPG-based Heart Rate Estimation 18 Feb 2025 · 0 repositories · arXiv:2502.12836
-
LMN: A Tool for Generating Machine Enforceable Policies from Natural Language Access Control Rules using LLMs 18 Feb 2025 · 0 repositories · arXiv:2502.12460
-
AdaSplash: Adaptive Sparse Flash Attention 17 Feb 2025 · 1 repository · arXiv:2502.12082Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
AI-generated Text Detection with a GLTR-based Approach 17 Feb 2025 · 0 repositories · arXiv:2502.12064
-
Biases in Edge Language Models: Detection, Analysis, and Mitigation 17 Feb 2025 · 0 repositories · arXiv:2502.11349
-
DiSCo: Device-Server Collaborative LLM-Based Text Streaming Services 17 Feb 2025 · 0 repositories · arXiv:2502.11417
-
SmartLLM: Smart Contract Auditing using Custom Generative AI 17 Feb 2025 · 0 repositories · arXiv:2502.13167
-
Performance Review on LLM for solving leetcode problems 16 Feb 2025 · 0 repositories · arXiv:2502.15770
-
TUMLU: A Unified and Native Language Understanding Benchmark for Turkic Languages 16 Feb 2025 · 1 repository · arXiv:2502.11020
-
Vendi-RAG: Adaptively Trading-Off Diversity And Quality Significantly Improves Retrieval Augmented Generation With LLMs 16 Feb 2025 · 0 repositories · arXiv:2502.11228
-
ControllableGPT: A Ground-Up Designed Controllable GPT for Molecule Optimization 15 Feb 2025 · 0 repositories · arXiv:2502.10631
-
The underlying structures of self-attention: symmetry, directionality, and emergent dynamics in Transformer training 15 Feb 2025 · 1 repository · arXiv:2502.10927
-
Do Large Language Models Reason Causally Like Us? Even Better? 14 Feb 2025 · 0 repositories · arXiv:2502.10215
-
Mechanistic Unveiling of Transformer Circuits: Self-Influence as a Key to Model Reasoning 13 Feb 2025 · 0 repositories · arXiv:2502.09022
-
Can Vision-Language Models Infer Speaker's Ignorance? The Role of Visual and Linguistic Cues 13 Feb 2025 · 0 repositories · arXiv:2502.09120
-
InTAR: Inter-Task Auto-Reconfigurable Accelerator Design for High Data Volume Variation in DNNs 12 Feb 2025 · 1 repository · arXiv:2502.08807
-
YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation 12 Feb 2025 · 0 repositories · arXiv:2502.10467
-
Automated Capability Discovery via Model Self-Exploration 11 Feb 2025 · 2 repositories · arXiv:2502.07577
-
Grammar Control in Dialogue Response Generation for Language Learning Chatbots 11 Feb 2025 · 1 repository · arXiv:2502.07544
-
Tractable Transformers for Flexible Conditional Generation 11 Feb 2025 · 0 repositories · arXiv:2502.07616
-
DebateBench: A Challenging Long Context Reasoning Benchmark For Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06279
-
Find Central Dogma Again: Leveraging Multilingual Transfer in Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.06253
-
LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights 10 Feb 2025 · 0 repositories · arXiv:2502.07049
-
Leveraging GPT-4o Efficiency for Detecting Rework Anomaly in Business Processes 10 Feb 2025 · 0 repositories · arXiv:2502.06918
-
Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models 9 Feb 2025 · 0 repositories · arXiv:2502.06039
-
Emergence of Episodic Memory in Transformers: Characterizing Changes in Temporal Structure of Attention Scores During Training 9 Feb 2025 · 0 repositories · arXiv:2502.06902
-
Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End" 9 Feb 2025 · 0 repositories · arXiv:2502.06898
-
ScaffoldGPT: A Scaffold-based GPT Model for Drug Optimization 9 Feb 2025 · 0 repositories · arXiv:2502.06891
-
Forbidden Science: Dual-Use AI Challenge Benchmark and Scientific Refusal Tests 8 Feb 2025 · 0 repositories · arXiv:2502.06867
-
The Odyssey of the Fittest: Can Agents Survive and Still Be Good? 8 Feb 2025 · 1 repository · arXiv:2502.05442
-
Can Large Language Models Understand Intermediate Representations? 7 Feb 2025 · 0 repositories · arXiv:2502.06854
-
Detection of LLM-Generated Java Code Using Discretized Nested Bigrams 7 Feb 2025 · 0 repositories · arXiv:2502.15740
-
EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification 7 Feb 2025 · 0 repositories · arXiv:2502.06852
-
Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis 6 Feb 2025 · 1 repository · arXiv:2502.04128
-
PsyPlay: Personality-Infused Role-Playing Conversational Agents 6 Feb 2025 · 0 repositories · arXiv:2502.03821
-
Aligning Human and Machine Attention for Enhanced Supervised Learning 4 Feb 2025 · 0 repositories · arXiv:2502.06811
-
CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance 4 Feb 2025 · 1 repository · arXiv:2502.04350
-
Conversation AI Dialog for Medicare powered by Finetuning and Retrieval Augmented Generation 4 Feb 2025 · 0 repositories · arXiv:2502.02249
-
Harmonic Loss Trains Interpretable AI Models 3 Feb 2025 · 1 repository · arXiv:2502.01628
-
Polynomial, trigonometric, and tropical activations 3 Feb 2025 · 1 repository · arXiv:2502.01247
-
DeepGate4: Efficient and Effective Representation Learning for Circuit Design at Scale 2 Feb 2025 · 1 repository · arXiv:2502.01681Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)