Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 18
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 18 of 40: papers 1,701 to 1,800 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Exploiting Inter-Layer Expert Affinity for Accelerating Mixture-of-Experts Model Inference 16 Jan 2024 · 1 repository · arXiv:2401.08383
-
RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture 16 Jan 2024 · 0 repositories · arXiv:2401.08406
-
RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning 16 Jan 2024 · 1 repository · arXiv:2401.08326Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 15 harvested samples)
-
Tuning Language Models by Proxy 16 Jan 2024 · 2 repositories · arXiv:2401.08565Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
A Novel Approach for Automatic Program Repair using Round-Trip Translation with Large Language Models 15 Jan 2024 · 1 repository · arXiv:2401.07994
-
Survival Analysis of Young Triple-Negative Breast Cancer Patients 15 Jan 2024 · 0 repositories · arXiv:2401.08712
-
Harnessing Large Language Models Over Transformer Models for Detecting Bengali Depressive Social Media Text: A Comprehensive Study 14 Jan 2024 · 1 repository · arXiv:2401.07310
-
Learning to be Homo Economicus: Can an LLM Learn Preferences from Choice 14 Jan 2024 · 0 repositories · arXiv:2401.07345
-
MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation 14 Jan 2024 · 0 repositories · arXiv:2401.07314
-
Streamlining the Selection Phase of Systematic Literature Reviews (SLRs) Using AI-Enabled GPT-4 Assistant API 14 Jan 2024 · 0 repositories · arXiv:2402.18582
-
A Novel Multi-Stage Prompting Approach for Language Agnostic MCQ Generation using GPT 13 Jan 2024 · 1 repository · arXiv:2401.07098
-
Assessing Large Language Models in Mechanical Engineering Education: A Study on Mechanics-Focused Conceptual Understanding 13 Jan 2024 · 0 repositories · arXiv:2401.12983
-
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation 13 Jan 2024 · 0 repositories · arXiv:2401.08694
-
Comparing GPT-4 and Open-Source Language Models in Misinformation Mitigation 12 Jan 2024 · 0 repositories · arXiv:2401.06920
-
Human-AI Collaborative Essay Scoring: A Dual-Process Framework with LLMs 12 Jan 2024 · 1 repository · arXiv:2401.06431
-
How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs 12 Jan 2024 · 2 repositories · arXiv:2401.06373Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Intention Analysis Makes LLMs A Good Jailbreak Defender 12 Jan 2024 · 1 repository · arXiv:2401.06561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Mission: Impossible Language Models 12 Jan 2024 · 1 repository · arXiv:2401.06416Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 12 harvested samples)
-
PersianMind: A Cross-Lingual Persian-English Large Language Model 12 Jan 2024 · 0 repositories · arXiv:2401.06466
-
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes 12 Jan 2024 · 1 repository · arXiv:2401.06930
-
Investigating Data Contamination for Pre-training Language Models 11 Jan 2024 · 0 repositories · arXiv:2401.06059
-
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs 11 Jan 2024 · 0 repositories · arXiv:2401.05940
-
Prompt-based mental health screening from social media text 11 Jan 2024 · 0 repositories · arXiv:2401.05912
-
The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models 11 Jan 2024 · 1 repository · arXiv:2401.05618Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
YOLO-Former: YOLO Shakes Hand With ViT 11 Jan 2024 · 0 repositories · arXiv:2401.06244
-
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning 10 Jan 2024 · 1 repository · arXiv:2401.05268Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
I am a Strange Dataset: Metalinguistic Tests for Language Models 10 Jan 2024 · 1 repository · arXiv:2401.05300
-
InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks 10 Jan 2024 · 1 repository · arXiv:2401.05507Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Can Active Label Correction Improve LLM-based Modular AI Systems? 10 Jan 2024 · 0 repositories · arXiv:2401.05467
-
Monte Carlo Tree Search for Recipe Generation using GPT-2 10 Jan 2024 · 0 repositories · arXiv:2401.05199
-
Reinforcement Learning for Optimizing RAG for Domain Chatbots 10 Jan 2024 · 0 repositories · arXiv:2401.06800
-
Fighting Fire with Fire: Adversarial Prompting to Generate a Misinformation Detection Dataset 9 Jan 2024 · 0 repositories · arXiv:2401.04481
-
Advancing Spatial Reasoning in Large Language Models: An In-Depth Evaluation and Enhancement Using the StepGame Benchmark 8 Jan 2024 · 1 repository · arXiv:2401.03991
-
Distortions in Judged Spatial Relations in Large Language Models 8 Jan 2024 · 0 repositories · arXiv:2401.04218
-
Advancing bioinformatics with large language models: components, applications and perspectives 8 Jan 2024 · 0 repositories · arXiv:2401.04155
-
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems 8 Jan 2024 · 1 repository · arXiv:2401.05443
-
Mixtral of Experts 8 Jan 2024 · 6 repositories · arXiv:2401.04088Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Exploring Defeasibility in Causal Reasoning 6 Jan 2024 · 0 repositories · arXiv:2401.03183
-
PIXAR: Auto-Regressive Language Modeling in Pixel Space 6 Jan 2024 · 0 repositories · arXiv:2401.03321
-
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors 6 Jan 2024 · 0 repositories · arXiv:2401.03238
-
Can Large Language Models Understand Molecules? 5 Jan 2024 · 2 repositories · arXiv:2402.00024
-
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 5 Jan 2024 · 1 repository · arXiv:2401.02954
-
Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks 5 Jan 2024 · 2 repositories · arXiv:2401.02731
-
Are LLMs Robust for Spoken Dialogues? 4 Jan 2024 · 0 repositories · arXiv:2401.02297
-
HyperSense: Hyperdimensional Intelligent Sensing for Energy-Efficient Sparse Data Processing 4 Jan 2024 · 0 repositories · arXiv:2401.10267
-
Re-evaluating the Memory-balanced Pipeline Parallelism: BPipe 4 Jan 2024 · 0 repositories · arXiv:2401.02088
-
Text2MDT: Extracting Medical Decision Trees from Medical Texts 4 Jan 2024 · 1 repository · arXiv:2401.02034
-
The Internet of Things in the Era of Generative AI: Vision and Challenges 3 Jan 2024 · 0 repositories · arXiv:2401.01923
-
Revisiting Zero-Shot Abstractive Summarization in the Era of Large Language Models from the Perspective of Position Bias 3 Jan 2024 · 1 repository · arXiv:2401.01989
-
Vietnamese Poem Generation & The Prospect Of Cross-Language Poem-To-Poem Translation 2 Jan 2024 · 1 repository · arXiv:2401.01078
-
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models 1 Jan 2024 · 1 repository · arXiv:2401.00757Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Computational Framework for Behavioral Assessment of LLM Therapists 1 Jan 2024 · 1 repository · arXiv:2401.00820
-
Adapt or Perish: Adaptive Sparse Transformer with Attentive Feature Refinement for Image Restoration 1 Jan 2024 · 1 repository
-
Large Language Models aren't all that you need 1 Jan 2024 · 0 repositories · arXiv:2401.00698
-
SEED-Bench: Benchmarking Multimodal Large Language Models 1 Jan 2024 · 1 repository
-
Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation 30 Dec 2023 · 1 repository · arXiv:2401.00280
-
Trace and Edit Relation Associations in GPT 30 Dec 2023 · 0 repositories · arXiv:2401.02976
-
Gemini in Reasoning: Unveiling Commonsense in Multimodal Large Language Models 29 Dec 2023 · 1 repository · arXiv:2312.17661
-
Jatmo: Prompt Injection Defense by Task-Specific Finetuning 29 Dec 2023 · 1 repository · arXiv:2312.17673Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Evaluating the Performance of Large Language Models for Spanish Language in Undergraduate Admissions Exams 28 Dec 2023 · 0 repositories · arXiv:2312.16845
-
LLM4Causal: Democratized Causal Tools for Everyone via Large Language Model 28 Dec 2023 · 0 repositories · arXiv:2312.17122
-
PanGu-π: Enhancing Language Model Architectures via Nonlinearity Compensation 27 Dec 2023 · 0 repositories · arXiv:2312.17276
-
ChartBench: A Benchmark for Complex Visual Reasoning in Charts 26 Dec 2023 · 0 repositories · arXiv:2312.15915
-
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4 26 Dec 2023 · 2 repositories · arXiv:2312.16171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SecQA: A Concise Question-Answering Dataset for Evaluating Large Language Models in Computer Security 26 Dec 2023 · 1 repository · arXiv:2312.15838Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Task Contamination: Language Models May Not Be Few-Shot Anymore 26 Dec 2023 · 0 repositories · arXiv:2312.16337
-
Fairness-Aware Structured Pruning in Transformers 24 Dec 2023 · 1 repository · arXiv:2312.15398Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Do LLM Agents Exhibit Social Behavior? 23 Dec 2023 · 0 repositories · arXiv:2312.15198
-
Understanding the Potential of FPGA-Based Spatial Acceleration for Large Language Model Inference 23 Dec 2023 · 1 repository · arXiv:2312.15159
-
Efficacy of Machine-Generated Instructions 22 Dec 2023 · 0 repositories · arXiv:2312.14423
-
FM-OV3D: Foundation Model-based Cross-modal Knowledge Blending for Open-Vocabulary 3D Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14465
-
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection 22 Dec 2023 · 1 repository · arXiv:2312.15068
-
How Smooth Is Attention? 22 Dec 2023 · 0 repositories · arXiv:2312.14820
-
Argue with Me Tersely: Towards Sentence-Level Counter-Argument Generation 21 Dec 2023 · 1 repository · arXiv:2312.13608
-
ChatGPT as a commenter to the news: can LLMs generate human-like opinions? 21 Dec 2023 · 1 repository · arXiv:2312.13961
-
De novo Drug Design using Reinforcement Learning with Multiple GPT Agents 21 Dec 2023 · 2 repositories · arXiv:2401.06155Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
InfoVisDial: An Informative Visual Dialogue Dataset by Bridging Large Multimodal and Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13503
-
TraceFL: Interpretability-Driven Debugging in Federated Learning via Neuron Provenance 21 Dec 2023 · 2 repositories · arXiv:2312.13632
-
Team Irisapu Project Description for DRC2023 21 Dec 2023 · 0 repositories · arXiv:2312.13765
-
Typhoon: Thai Large Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13951
-
AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation 20 Dec 2023 · 1 repository · arXiv:2312.13010Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Benchmarking and Analyzing In-context Learning, Fine-tuning and Supervised Learning for Biomedical Knowledge Curation: a focused study on chemical entities of biological interest 20 Dec 2023 · 0 repositories · arXiv:2312.12989
-
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks 20 Dec 2023 · 3 repositories · arXiv:2312.13322
-
Can ChatGPT be Your Personal Medical Assistant? 19 Dec 2023 · 0 repositories · arXiv:2312.12006
-
Can Transformers Learn Sequential Function Classes In Context? 19 Dec 2023 · 0 repositories · arXiv:2312.12655
-
Large Language Models in Medical Term Classification and Unexpected Misalignment Between Response and Reasoning 19 Dec 2023 · 0 repositories · arXiv:2312.14184
-
An In-depth Look at Gemini's Language Abilities 18 Dec 2023 · 1 repository · arXiv:2312.11444Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Generative linguistic representation for spoken language identification 18 Dec 2023 · 0 repositories · arXiv:2312.10964
-
Artificial intelligence optical hardware empowers high-resolution hyperspectral video understanding at 1.2 Tb/s 17 Dec 2023 · 0 repositories · arXiv:2312.10639
-
Decoding Concerns: Multi-label Classification of Vaccine Sentiments in Social Media 17 Dec 2023 · 1 repository · arXiv:2312.10626
-
Evaluating AI Vocational Skills Through Professional Testing 17 Dec 2023 · 0 repositories · arXiv:2312.10603
-
HyperPIE: Hyperparameter Information Extraction from Scientific Publications 17 Dec 2023 · 1 repository · arXiv:2312.10638
-
Mixed Distillation Helps Smaller Language Model Better Reasoning 17 Dec 2023 · 0 repositories · arXiv:2312.10730
-
Multi-Label Classification of COVID-Tweets Using Large Language Models 17 Dec 2023 · 1 repository · arXiv:2312.10748
-
T2M-HiFiGPT: Generating High Quality Human Motion from Textual Descriptions with Residual Discrete Representations 17 Dec 2023 · 0 repositories · arXiv:2312.10628
-
A Comparative Analysis of Large Language Models for Code Documentation Generation 16 Dec 2023 · 0 repositories · arXiv:2312.10349
-
A Novel Dataset for Financial Education Text Simplification in Spanish 15 Dec 2023 · 0 repositories · arXiv:2312.09897
-
Distilling Large Language Models for Matching Patients to Clinical Trials 15 Dec 2023 · 0 repositories · arXiv:2312.09958
-
Exploring Multi-Level Threats in Telegram Data with AI-Human Annotation: A Preliminary Study 15 Dec 2023 · 0 repositories
-
Red AI? Inconsistent Responses from GPT3.5 Models on Political Issues in the US and China 15 Dec 2023 · 0 repositories · arXiv:2312.09917