Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 2
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 2 of 40: papers 101 to 200 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
GRADA: Graph-based Reranker against Adversarial Documents Attack 12 May 2025 · 1 repository · arXiv:2505.07546
-
No Query, No Access 12 May 2025 · 0 repositories · arXiv:2505.07258
-
Evaluating Reasoning LLMs for Suicide Screening with the Columbia-Suicide Severity Rating Scale 11 May 2025 · 1 repository · arXiv:2505.13480
-
REFINE-AF: A Task-Agnostic Framework to Align Language Models via Self-Generated Instructions using Reinforcement Learning from Automated Feedback 10 May 2025 · 0 repositories · arXiv:2505.06548
-
An empathic GPT-based chatbot to talk about mental disorders with Spanish teenagers 9 May 2025 · 0 repositories · arXiv:2505.05828
-
CellVerse: Do Large Language Models Really Understand Cell Biology? 9 May 2025 · 0 repositories · arXiv:2505.07865
-
Multi-Agent Systems for Robotic Autonomy with LLMs 9 May 2025 · 0 repositories · arXiv:2505.05762
-
What Is Next for LLMs? Next-Generation AI Computing Hardware Using Photonic Chips 9 May 2025 · 0 repositories · arXiv:2505.05794
-
AI Approaches to Qualitative and Quantitative News Analytics on NATO Unity 8 May 2025 · 0 repositories · arXiv:2505.06313
-
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization 8 May 2025 · 0 repositories · arXiv:2505.05070
-
An Empirical Study of OpenAI API Discussions on Stack Overflow 7 May 2025 · 0 repositories · arXiv:2505.04084
-
Bringing legal knowledge to the public by constructing a legal question bank using large-scale pre-trained language model 7 May 2025 · 0 repositories · arXiv:2505.04132
-
LLM-e Guess: Can LLMs Capabilities Advance Without Hardware Progress? 7 May 2025 · 1 repository · arXiv:2505.04075
-
Towards Effectively Leveraging Execution Traces for Program Repair with Code LLMs 7 May 2025 · 0 repositories · arXiv:2505.04441
-
A Comparative Analysis of Ethical and Safety Gaps in LLMs using Relative Danger Coefficient 6 May 2025 · 0 repositories · arXiv:2505.04654
-
Evaluation of LLMs on Long-tail Entity Linking in Historical Documents 6 May 2025 · 0 repositories · arXiv:2505.03473
-
Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights 6 May 2025 · 0 repositories · arXiv:2505.03205
-
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis 5 May 2025 · 0 repositories · arXiv:2505.03019
-
A Domain Adaptation of Large Language Models for Classifying Mechanical Assembly Components 2 May 2025 · 0 repositories · arXiv:2505.01627
-
JaccDiv: A Metric and Benchmark for Quantifying Diversity of Generated Marketing Text in the Music Industry 29 Apr 2025 · 0 repositories · arXiv:2504.20849
-
UniDetox: Universal Detoxification of Large Language Models via Dataset Distillation 29 Apr 2025 · 1 repository · arXiv:2504.20500Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
VeriDebug: A Unified LLM for Verilog Debugging via Contrastive Embedding and Guided Correction 27 Apr 2025 · 0 repositories · arXiv:2504.19099
-
Why you shouldn't fully trust ChatGPT: A synthesis of this AI tool's error rates across disciplines and the software engineering lifecycle 26 Apr 2025 · 0 repositories · arXiv:2504.18858
-
Adversarial Attacks on LLM-as-a-Judge Systems: Insights from Prompt Injections 25 Apr 2025 · 0 repositories · arXiv:2504.18333
-
Optimism, Expectation, or Sarcasm? Multi-Class Hope Speech Detection in Spanish and English 24 Apr 2025 · 0 repositories · arXiv:2504.17974
-
A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms 23 Apr 2025 · 0 repositories · arXiv:2504.16420
-
Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate 23 Apr 2025 · 0 repositories · arXiv:2504.16489
-
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 22 Apr 2025 · 0 repositories · arXiv:2504.16027
-
The 1st EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval 21 Apr 2025 · 0 repositories · arXiv:2504.14788
-
AI with Emotions: Exploring Emotional Expressions in Large Language Models 20 Apr 2025 · 0 repositories · arXiv:2504.14706
-
Less is More: Adaptive Coverage for Synthetic Training Data 20 Apr 2025 · 0 repositories · arXiv:2504.14508
-
LLM Sensitivity Evaluation Framework for Clinical Diagnosis 18 Apr 2025 · 0 repositories · arXiv:2504.13475
-
Validating LLM-Generated Relevance Labels for Educational Resource Search 17 Apr 2025 · 0 repositories · arXiv:2504.12732
-
ZeroSumEval: Scaling LLM Evaluation with Inter-Model Competition 17 Apr 2025 · 1 repository · arXiv:2504.12562
-
Using customized GPT to develop prompting proficiency in architectural AI-generated images 16 Apr 2025 · 0 repositories · arXiv:2504.13948
-
Hallucination-Aware Generative Pretrained Transformer for Cooperative Aerial Mobility Control 15 Apr 2025 · 0 repositories · arXiv:2504.10831
-
VEXP: A Low-Cost RISC-V ISA Extension for Accelerated Softmax Computation in Transformers 15 Apr 2025 · 0 repositories · arXiv:2504.11227
-
Paging Dr. GPT: Extracting Information from Clinical Notes to Enhance Patient Predictions 14 Apr 2025 · 0 repositories · arXiv:2504.12338
-
Lumos: Efficient Performance Modeling and Estimation for Large-scale LLM Training 12 Apr 2025 · 0 repositories · arXiv:2504.09307
-
Examining GPT's Capability to Generate and Map Course Concepts and Their Relationship 11 Apr 2025 · 0 repositories · arXiv:2504.08856
-
Learning from Elders: Making an LLM-powered Chatbot for Retirement Communities more Accessible through User-centered Design 11 Apr 2025 · 0 repositories · arXiv:2504.08985
-
LLM for Comparative Narrative Analysis 11 Apr 2025 · 0 repositories · arXiv:2504.08211
-
SWAN-GPT: An Efficient and Scalable Approach for Long-Context Language Modeling 11 Apr 2025 · 0 repositories · arXiv:2504.08719
-
AI Coding with Few-Shot Prompting for Thematic Analysis 10 Apr 2025 · 0 repositories · arXiv:2504.07408
-
Can Reasoning LLMs Enhance Clinical Document Classification? 10 Apr 2025 · 1 repository · arXiv:2504.08040
-
ConceptFormer: Towards Efficient Use of Knowledge-Graph Embeddings in Large Language Models 10 Apr 2025 · 0 repositories · arXiv:2504.07624
-
MRD-RAG: Enhancing Medical Diagnosis with Multi-Round Retrieval-Augmented Generation 10 Apr 2025 · 1 repository · arXiv:2504.07724
-
PoGO: A Scalable Proof of Useful Work via Quantized Gradient Descent and Merkle Proofs 10 Apr 2025 · 0 repositories · arXiv:2504.07540
-
Face-LLaVA: Facial Expression and Attribute Understanding through Instruction Tuning 9 Apr 2025 · 0 repositories · arXiv:2504.07198
-
Assessing how hyperparameters impact Large Language Models' sarcasm detection performance 8 Apr 2025 · 0 repositories · arXiv:2504.06166
-
Weak-for-Strong: Training Weak Meta-Agent to Harness Strong Executors 7 Apr 2025 · 1 repository · arXiv:2504.04785
-
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models 6 Apr 2025 · 0 repositories · arXiv:2506.10005
-
Do LLM Evaluators Prefer Themselves for a Reason? 4 Apr 2025 · 1 repository · arXiv:2504.03846
-
Efficient Dynamic Clustering-Based Document Compression for Retrieval-Augmented-Generation 4 Apr 2025 · 1 repository · arXiv:2504.03165
-
Metamorphic Testing for Fairness Evaluation in Large Language Models: Identifying Intersectional Bias in LLaMA and GPT 4 Apr 2025 · 0 repositories · arXiv:2504.07982
-
Localized Definitions and Distributed Reasoning: A Proof-of-Concept Mechanistic Interpretability Study via Activation Patching 3 Apr 2025 · 1 repository · arXiv:2504.02976
-
Task as Context Prompting for Accurate Medical Symptom Coding Using Large Language Models 3 Apr 2025 · 1 repository · arXiv:2504.03051
-
Analysis of an Idealized Stochastic Polyak Method and its Application to Black-Box Model Distillation 2 Apr 2025 · 0 repositories · arXiv:2504.01898
-
Efficient Model Selection for Time Series Forecasting via LLMs 2 Apr 2025 · 0 repositories · arXiv:2504.02119
-
GPT Adoption and the Impact of Disclosure Policies 2 Apr 2025 · 0 repositories · arXiv:2504.01566
-
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System 2 Apr 2025 · 0 repositories · arXiv:2504.02894
-
Strategize Globally, Adapt Locally: A Multi-Turn Red Teaming Agent with Dual-Level Learning 2 Apr 2025 · 0 repositories · arXiv:2504.01278
-
Towards Interpretable Soft Prompts 2 Apr 2025 · 1 repository · arXiv:2504.02144
-
A Unified Virtual Mixture-of-Experts Framework:Enhanced Inference and Hallucination Mitigation in Single-Model System 1 Apr 2025 · 0 repositories · arXiv:2504.03739
-
Grade Guard: A Smart System for Short Answer Automated Grading 1 Apr 2025 · 0 repositories · arXiv:2504.01253
-
Accelerating High-Efficiency Organic Photovoltaic Discovery via Pretrained Graph Neural Networks and Generative Reinforcement Learning 31 Mar 2025 · 0 repositories · arXiv:2503.23766
-
Does "Reasoning" with Large Language Models Improve Recognizing, Generating, and Reframing Unhelpful Thoughts? 31 Mar 2025 · 0 repositories · arXiv:2504.00163
-
Rubric Is All You Need: Enhancing LLM-based Code Evaluation With Question-Specific Rubrics 31 Mar 2025 · 0 repositories · arXiv:2503.23989
-
Text Chunking for Document Classification for Urban System Management using Large Language Models 31 Mar 2025 · 1 repository · arXiv:2504.00274
-
Beyond Detection: Designing AI-Resilient Assessments with Automated Feedback Tool to Foster Critical Thinking 30 Mar 2025 · 0 repositories · arXiv:2503.23622
-
Large Language Models Are Better Logical Fallacy Reasoners with Counterargument, Explanation, and Goal-Aware Prompt Formulation 30 Mar 2025 · 1 repository · arXiv:2503.23363
-
LaViC: Adapting Large Vision-Language Models to Visually-Aware Conversational Recommendation 30 Mar 2025 · 1 repository · arXiv:2503.23312
-
SCORE: Story Coherence and Retrieval Enhancement for AI Narratives 30 Mar 2025 · 0 repositories · arXiv:2503.23512
-
A Training-free LLM Framework with Interaction between Contextually Related Subtasks in Solving Complex Tasks 29 Mar 2025 · 0 repositories · arXiv:2503.23053
-
The realization of tones in spontaneous spoken Taiwan Mandarin: a corpus-based survey and theory-driven computational modeling 29 Mar 2025 · 0 repositories · arXiv:2503.23163
-
Integrating Artificial Intelligence with Human Expertise: An In-depth Analysis of ChatGPT's Capabilities in Generating Metamorphic Relations 28 Mar 2025 · 0 repositories · arXiv:2503.22141
-
Leveraging LLMs for Predicting Unknown Diagnoses from Clinical Notes 28 Mar 2025 · 0 repositories · arXiv:2503.22092
-
An evaluation of LLMs and Google Translate for translation of selected Indian languages via sentiment and semantic analyses 27 Mar 2025 · 0 repositories · arXiv:2503.21393
-
Advancements in Natural Language Processing: Exploring Transformer-Based Architectures for Text Understanding 26 Mar 2025 · 0 repositories · arXiv:2503.20227
-
Can We Make Code Green? Understanding Trade-Offs in LLMs vs. Human Code Optimizations 26 Mar 2025 · 0 repositories · arXiv:2503.20126
-
Iterative Prompting with Persuasion Skills in Jailbreaking Large Language Models 26 Mar 2025 · 0 repositories · arXiv:2503.20320
-
Patients Speak, AI Listens: LLM-based Analysis of Online Reviews Uncovers Key Drivers for Urgent Care Satisfaction 26 Mar 2025 · 0 repositories · arXiv:2503.20981
-
GPT Meets Graphs and KAN Splines: Testing Novel Frameworks on Multitask Fine-Tuned GPT-2 with LoRA 25 Mar 2025 · 0 repositories · arXiv:2504.10490
-
How to Capture and Study Conversations Between Research Participants and ChatGPT: GPT for Researchers (g4r.org) 24 Mar 2025 · 0 repositories · arXiv:2503.18303
-
REALM: A Dataset of Real-World LLM Use Cases 24 Mar 2025 · 0 repositories · arXiv:2503.18792
-
Investigating Recent Large Language Models for Vietnamese Machine Reading Comprehension 23 Mar 2025 · 0 repositories · arXiv:2503.18062
-
Event-Based Crossing Dataset (EBCD) 21 Mar 2025 · 1 repository · arXiv:2503.17499
-
DocVideoQA: Towards Comprehensive Understanding of Document-Centric Videos through Question Answering 20 Mar 2025 · 0 repositories · arXiv:2503.15887
-
GraPLUS: Graph-based Placement Using Semantics for Image Composition 20 Mar 2025 · 0 repositories · arXiv:2503.15761
-
Unify and Triumph: Polyglot, Diverse, and Self-Consistent Generation of Unit Tests with LLMs 20 Mar 2025 · 0 repositories · arXiv:2503.16144
-
TROVE: A Challenge for Fine-Grained Text Provenance via Source Sentence Tracing and Relationship Classification 19 Mar 2025 · 0 repositories · arXiv:2503.15289
-
BurTorch: Revisiting Training from First Principles by Coupling Autodiff, Math Optimization, and Systems 18 Mar 2025 · 1 repository · arXiv:2503.13795
-
XOXO: Stealthy Cross-Origin Context Poisoning Attacks against AI Coding Assistants 18 Mar 2025 · 0 repositories · arXiv:2503.14281
-
Are LLMs (Really) Ideological? An IRT-based Analysis and Alignment Tool for Perceived Socio-Economic Bias in LLMs 17 Mar 2025 · 0 repositories · arXiv:2503.13149
-
Can Language Models Follow Multiple Turns of Entangled Instructions? 17 Mar 2025 · 1 repository · arXiv:2503.13222
-
Feature Extraction and Analysis for GPT-Generated Text 17 Mar 2025 · 0 repositories · arXiv:2503.13687
-
Generative AI for Software Architecture. Applications, Trends, Challenges, and Future Directions 17 Mar 2025 · 0 repositories · arXiv:2503.13310
-
Integrating Chain-of-Thought and Retrieval Augmented Generation Enhances Rare Disease Diagnosis from Clinical Notes 15 Mar 2025 · 0 repositories · arXiv:2503.12286
-
Augmenting Image Annotation: A Human-LMM Collaborative Framework for Efficient Object Selection and Label Generation 14 Mar 2025 · 0 repositories · arXiv:2503.11096
-
Combining Causal Models for More Accurate Abstractions of Neural Networks 14 Mar 2025 · 1 repository · arXiv:2503.11429Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)