Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 8
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 8 of 40: papers 701 to 800 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Task-oriented Prompt Enhancement via Script Generation 24 Sep 2024 · 0 repositories · arXiv:2409.16418
-
UICE-MIRNet guided image enhancement for underwater object detection 24 Sep 2024 · 0 repositories
-
Advancing Depression Detection on Social Media Platforms Through Fine-Tuned Large Language Models 23 Sep 2024 · 0 repositories · arXiv:2409.14794
-
Chattronics: using GPTs to assist in the design of data acquisition systems 23 Sep 2024 · 0 repositories · arXiv:2409.15183
-
PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs 23 Sep 2024 · 1 repository · arXiv:2409.14866Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
GEM-RAG: Graphical Eigen Memories For Retrieval Augmented Generation 23 Sep 2024 · 0 repositories · arXiv:2409.15566
-
Location is Key: Leveraging Large Language Model for Functional Bug Localization in Verilog 23 Sep 2024 · 0 repositories · arXiv:2409.15186
-
Privacy Policy Analysis through Prompt Engineering for LLMs 23 Sep 2024 · 0 repositories · arXiv:2409.14879
-
Safe Guard: an LLM-agent for Real-time Voice-based Hate Speech Detection in Social Virtual Reality 23 Sep 2024 · 0 repositories · arXiv:2409.15623
-
SDBA: A Stealthy and Long-Lasting Durable Backdoor Attack in Federated Learning 23 Sep 2024 · 1 repository · arXiv:2409.14805
-
Can pre-trained language models generate titles for research papers? 22 Sep 2024 · 1 repository · arXiv:2409.14602
-
Evaluating the Quality of Code Comments Generated by Large Language Models for Novice Programmers 22 Sep 2024 · 0 repositories · arXiv:2409.14368
-
LLMs are One-Shot URL Classifiers and Explainers 22 Sep 2024 · 0 repositories · arXiv:2409.14306
-
Proof Automation with Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14274
-
Can LLMs replace Neil deGrasse Tyson? Evaluating the Reliability of LLMs as Science Communicators 21 Sep 2024 · 1 repository · arXiv:2409.14037
-
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling 21 Sep 2024 · 1 repository · arXiv:2409.14175
-
Drift to Remember 21 Sep 2024 · 0 repositories · arXiv:2409.13997
-
Knowledge in Triples for LLMs: Enhancing Table QA Accuracy with Semantic Extraction 21 Sep 2024 · 0 repositories · arXiv:2409.14192
-
Loop Neural Networks for Parameter Sharing 21 Sep 2024 · 0 repositories · arXiv:2409.14199
-
EMMeTT: Efficient Multimodal Machine Translation Training 20 Sep 2024 · 0 repositories · arXiv:2409.13523
-
FAIR GPT: A virtual consultant for research data management in ChatGPT 20 Sep 2024 · 1 repository · arXiv:2410.07108
-
HUT: A More Computation Efficient Fine-Tuning Method With Hadamard Updated Transformation 20 Sep 2024 · 0 repositories · arXiv:2409.13501
-
Leveraging Knowledge Graphs and LLMs to Support and Monitor Legislative Systems 20 Sep 2024 · 0 repositories · arXiv:2409.13252
-
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks 20 Sep 2024 · 1 repository · arXiv:2409.13203Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
TalkMosaic: Interactive PhotoMosaic with Multi-modal LLM Q&A Interactions 20 Sep 2024 · 0 repositories · arXiv:2409.13941
-
Enhancing TinyBERT for Financial Sentiment Analysis Using GPT-Augmented FinBERT Distillation 19 Sep 2024 · 1 repository · arXiv:2409.18999
-
On the Effectiveness of LLMs for Manual Test Verifications 19 Sep 2024 · 0 repositories · arXiv:2409.12405
-
Retrieval-Augmented Test Generation: How Far Are We? 19 Sep 2024 · 0 repositories · arXiv:2409.12682
-
MAgICoRe: Multi-Agent, Iterative, Coarse-to-Fine Refinement for Reasoning 18 Sep 2024 · 1 repository · arXiv:2409.12147Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Recommendation with Generative Models 18 Sep 2024 · 0 repositories · arXiv:2409.15173
-
TART: An Open-Source Tool-Augmented Framework for Explainable Table-based Reasoning 18 Sep 2024 · 1 repository · arXiv:2409.11724
-
A Unified Framework to Classify Business Activities into International Standard Industrial Classification through Large Language Models for Circular Economy 17 Sep 2024 · 0 repositories · arXiv:2409.18988
-
Small Language Models can Outperform Humans in Short Creative Writing: A Study Comparing SLMs with Humans and LLMs 17 Sep 2024 · 1 repository · arXiv:2409.11547
-
Benchmarking Large Language Model Uncertainty for Prompt Optimization 16 Sep 2024 · 1 repository · arXiv:2409.10044
-
Convergence of Sharpness-Aware Minimization Algorithms using Increasing Batch Size and Decaying Learning Rate 16 Sep 2024 · 0 repositories · arXiv:2409.09984
-
LLM-DER:A Named Entity Recognition Method Based on Large Language Models for Chinese Coal Chemical Domain 16 Sep 2024 · 0 repositories · arXiv:2409.10077
-
SelECT-SQL: Self-correcting ensemble Chain-of-Thought for Text-to-SQL 16 Sep 2024 · 1 repository · arXiv:2409.10007
-
Detection Made Easy: Potentials of Large Language Models for Solidity Vulnerabilities 15 Sep 2024 · 0 repositories · arXiv:2409.10574
-
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation 15 Sep 2024 · 0 repositories · arXiv:2409.09584
-
An empirical evaluation of using ChatGPT to summarize disputes for recommending similar labor and employment cases in Chinese 14 Sep 2024 · 0 repositories · arXiv:2409.09280
-
Optimizing Ingredient Substitution Using Large Language Models to Enhance Phytochemical Content in Recipes 13 Sep 2024 · 0 repositories · arXiv:2409.08792
-
Experimenting with Legal AI Solutions: The Case of Question-Answering for Access to Justice 12 Sep 2024 · 0 repositories · arXiv:2409.07713
-
Stable Language Model Pre-training by Reducing Embedding Variability 12 Sep 2024 · 0 repositories · arXiv:2409.07787
-
A Novel Mathematical Framework for Objective Characterization of Ideas 11 Sep 2024 · 0 repositories · arXiv:2409.07578
-
Towards Fairer Health Recommendations: finding informative unbiased samples via Word Sense Disambiguation 11 Sep 2024 · 0 repositories · arXiv:2409.07424
-
Accelerating Large Language Model Pretraining via LFR Pedagogy: Learn, Focus, and Review 10 Sep 2024 · 0 repositories · arXiv:2409.06131
-
Can Large Language Models Unlock Novel Scientific Research Ideas? 10 Sep 2024 · 1 repository · arXiv:2409.06185Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
Generative AI for Requirements Engineering: A Systematic Literature Review 10 Sep 2024 · 0 repositories · arXiv:2409.06741
-
Classification performance and reproducibility of GPT-4 omni for information extraction from veterinary electronic health records 9 Sep 2024 · 1 repository · arXiv:2409.13727
-
Assessing SPARQL capabilities of Large Language Models 9 Sep 2024 · 2 repositories · arXiv:2409.05925
-
Elsevier Arena: Human Evaluation of Chemistry/Biology/Health Foundational Large Language Models 9 Sep 2024 · 0 repositories · arXiv:2409.05486
-
FairHome: A Fair Housing and Fair Lending Dataset 9 Sep 2024 · 0 repositories · arXiv:2409.05990
-
Harmonic Reasoning in Large Language Models 9 Sep 2024 · 0 repositories · arXiv:2409.05521
-
Identifying the sources of ideological bias in GPT models through linguistic variation in output 9 Sep 2024 · 0 repositories · arXiv:2409.06043
-
Regression with Large Language Models for Materials and Molecular Property Prediction 9 Sep 2024 · 0 repositories · arXiv:2409.06080
-
Vision-fused Attack: Advancing Aggressive and Stealthy Adversarial Text against Neural Machine Translation 8 Sep 2024 · 1 repository · arXiv:2409.05021Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Column Vocabulary Association (CVA): semantic interpretation of dataless tables 6 Sep 2024 · 0 repositories · arXiv:2409.13709
-
Towards Safer Online Spaces: Simulating and Assessing Intervention Strategies for Eating Disorder Discussions 6 Sep 2024 · 0 repositories · arXiv:2409.04043
-
Bypassing DARCY Defense: Indistinguishable Universal Adversarial Triggers 5 Sep 2024 · 0 repositories · arXiv:2409.03183
-
Evaluating Open-Source Sparse Autoencoders on Disentangling Factual Knowledge in GPT-2 Small 5 Sep 2024 · 1 repository · arXiv:2409.04478Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MaterialBENCH: Evaluating College-Level Materials Science Problem-Solving Abilities of Large Language Models 5 Sep 2024 · 0 repositories · arXiv:2409.03161
-
Sketch: A Toolkit for Streamlining LLM Operations 5 Sep 2024 · 0 repositories · arXiv:2409.03346
-
A Comparative Study on Large Language Models for Log Parsing 4 Sep 2024 · 0 repositories · arXiv:2409.02474
-
How Privacy-Savvy Are Large Language Models? A Case Study on Compliance and Privacy Technical Review 4 Sep 2024 · 0 repositories · arXiv:2409.02375
-
Irrelevant Alternatives Bias Large Language Model Hiring Decisions 4 Sep 2024 · 0 repositories · arXiv:2409.15299
-
Large Language Models as Efficient Reward Function Searchers for Custom-Environment Multi-Objective Reinforcement Learning 4 Sep 2024 · 0 repositories · arXiv:2409.02428
-
More is More: Addition Bias in Large Language Models 4 Sep 2024 · 1 repository · arXiv:2409.02569
-
Dialogue You Can Trust: Human and AI Perspectives on Generated Conversations 3 Sep 2024 · 0 repositories · arXiv:2409.01808
-
It is Time to Develop an Auditing Framework to Promote Value Aware Chatbots 3 Sep 2024 · 1 repository · arXiv:2409.01539
-
LifeGPT: Topology-Agnostic Generative Pretrained Transformer Model for Cellular Automata 3 Sep 2024 · 1 repository · arXiv:2409.12182
-
The Era of Foundation Models in Medical Imaging is Approaching : A Scoping Review of the Clinical Value of Large-Scale Generative AI Applications in Radiology 3 Sep 2024 · 0 repositories · arXiv:2409.12973
-
Self-Judge: Selective Instruction Following with Alignment Self-Evaluation 2 Sep 2024 · 1 repository · arXiv:2409.00935Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Deep Knowledge-Infusion For Explainable Depression Detection 1 Sep 2024 · 0 repositories · arXiv:2409.02122
-
An Empirical Study on Information Extraction using Large Language Models 31 Aug 2024 · 0 repositories · arXiv:2409.00369
-
Assessing Generative Language Models in Classification Tasks: Performance and Self-Evaluation Capabilities in the Environmental and Climate Change Domain 30 Aug 2024 · 1 repository · arXiv:2408.17362
-
Can Large Language Models Address Open-Target Stance Detection? 30 Aug 2024 · 0 repositories · arXiv:2409.00222
-
ProGRes: Prompted Generative Rescoring on ASR n-Best 30 Aug 2024 · 1 repository · arXiv:2409.00217
-
Retrieval-Augmented Natural Language Reasoning for Explainable Visual Question Answering 30 Aug 2024 · 0 repositories · arXiv:2408.17006
-
Training Ultra Long Context Language Model with Fully Pipelined Distributed Transformer 30 Aug 2024 · 1 repository · arXiv:2408.16978
-
Assessing Large Language Models for Online Extremism Research: Identification, Explanation, and New Knowledge 29 Aug 2024 · 0 repositories · arXiv:2408.16749
-
LLaVA-Chef: A Multi-modal Generative Model for Food Recipes 29 Aug 2024 · 1 repository · arXiv:2408.16889
-
FRACTURED-SORRY-Bench: Framework for Revealing Attacks in Conversational Turns Undermining Refusal Efficacy and Defenses over SORRY-Bench (Automated Multi-shot Jailbreaks) 28 Aug 2024 · 0 repositories · arXiv:2408.16163
-
Leveraging Large Language Models for Wireless Symbol Detection via In-Context Learning 28 Aug 2024 · 0 repositories · arXiv:2409.00124
-
Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation 28 Aug 2024 · 1 repository · arXiv:2408.15876
-
A Survey of Large Language Models for European Languages 27 Aug 2024 · 0 repositories · arXiv:2408.15040
-
Strategic Optimization and Challenges of Large Language Models in Object-Oriented Programming 27 Aug 2024 · 0 repositories · arXiv:2408.14834
-
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis 27 Aug 2024 · 0 repositories · arXiv:2409.00106
-
CHARTOM: A Visual Theory-of-Mind Benchmark for Multimodal Large Language Models 26 Aug 2024 · 1 repository · arXiv:2408.14419
-
Bidirectional Awareness Induction in Autoregressive Seq2Seq Models 25 Aug 2024 · 0 repositories · arXiv:2408.13959
-
CodeGraph: Enhancing Graph Reasoning of LLMs with Code 25 Aug 2024 · 1 repository · arXiv:2408.13863
-
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models 25 Aug 2024 · 1 repository · arXiv:2409.00084
-
An In-Depth Investigation of Data Collection in LLM App Ecosystems 23 Aug 2024 · 0 repositories · arXiv:2408.13247
-
Enhancing Automated Program Repair with Solution Design 22 Aug 2024 · 0 repositories · arXiv:2408.12056
-
Enhancing Multi-hop Reasoning through Knowledge Erasure in Large Language Model Editing 22 Aug 2024 · 0 repositories · arXiv:2408.12456
-
Optimizing Performance: How Compact Models Match or Exceed GPT's Classification Capabilities through Fine-Tuning 22 Aug 2024 · 0 repositories · arXiv:2409.11408
-
Applying and Evaluating Large Language Models in Mental Health Care: A Scoping Review of Human-Assessed Generative Tasks 21 Aug 2024 · 0 repositories · arXiv:2408.11288
-
D-RMGPT: Robot-assisted collaborative tasks driven by large multimodal models 21 Aug 2024 · 0 repositories · arXiv:2408.11761
-
Mixed Sparsity Training: Achieving 4× FLOP Reduction for Transformer Pretraining 21 Aug 2024 · 0 repositories · arXiv:2408.11746
-
Unlocking Adversarial Suffix Optimization Without Affirmative Phrases: Efficient Black-box Jailbreaking via LLM as Optimizer 21 Aug 2024 · 1 repository · arXiv:2408.11313Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CTP-LLM: Clinical Trial Phase Transition Prediction Using Large Language Models 20 Aug 2024 · 0 repositories · arXiv:2408.10995