Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 19
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 19 of 40: papers 1,801 to 1,900 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning 14 Dec 2023 · 0 repositories · arXiv:2312.08901
-
Dynamic Retrieval-Augmented Generation 14 Dec 2023 · 0 repositories · arXiv:2312.08976
-
Inter-Layer Scheduling Space Exploration for Multi-model Inference on Heterogeneous Chiplets 14 Dec 2023 · 0 repositories · arXiv:2312.09401
-
Motion Flow Matching for Human Motion Synthesis and Editing 14 Dec 2023 · 0 repositories · arXiv:2312.08895
-
Self-Evaluation Improves Selective Generation in Large Language Models 14 Dec 2023 · 0 repositories · arXiv:2312.09300
-
Successor Heads: Recurring, Interpretable Attention Heads In The Wild 14 Dec 2023 · 0 repositories · arXiv:2312.09230
-
TinyGSM: achieving >80% on GSM8k with small language models 14 Dec 2023 · 0 repositories · arXiv:2312.09241
-
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision 14 Dec 2023 · 0 repositories · arXiv:2312.09390
-
Weaving Pathways for Justice with GPT: LLM-driven automated drafting of interactive legal applications 14 Dec 2023 · 1 repository · arXiv:2312.09198
-
Causality Analysis for Evaluating the Security of Large Language Models 13 Dec 2023 · 1 repository · arXiv:2312.07876Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Large Language Models are Complex Table Parsers 13 Dec 2023 · 0 repositories · arXiv:2312.11521
-
Native Language Identification with Large Language Models 13 Dec 2023 · 0 repositories · arXiv:2312.07819
-
AI Control: Improving Safety Despite Intentional Subversion 12 Dec 2023 · 1 repository · arXiv:2312.06942Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Exploring Large Language Models to Facilitate Variable Autonomy for Human-Robot Teaming 12 Dec 2023 · 0 repositories · arXiv:2312.07214
-
Image Content Generation with Causal Reasoning 12 Dec 2023 · 1 repository · arXiv:2312.07132Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Multilingual large language models leak human stereotypes across language boundaries 12 Dec 2023 · 1 repository · arXiv:2312.07141
-
Reducing Energy Bloat in Large Model Training 12 Dec 2023 · 2 repositories · arXiv:2312.06902
-
SM70: A Large Language Model for Medical Devices 12 Dec 2023 · 0 repositories · arXiv:2312.06974
-
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions 11 Dec 2023 · 1 repository · arXiv:2312.12450Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Evaluating ChatGPT as a Question Answering System: A Comprehensive Analysis and Comparison with Existing Models 11 Dec 2023 · 0 repositories · arXiv:2312.07592
-
Generative Large Language Models Are All-purpose Text Analytics Engines: Text-to-text Learning Is All Your Need 11 Dec 2023 · 0 repositories · arXiv:2312.06099
-
Survey on Foundation Models for Prognostics and Health Management in Industrial Cyber-Physical Systems 11 Dec 2023 · 0 repositories · arXiv:2312.06261
-
Early ChatGPT User Portrait through the Lens of Data 10 Dec 2023 · 0 repositories · arXiv:2312.10078
-
Sim-GPT: Text Similarity via GPT Annotated Data 9 Dec 2023 · 1 repository · arXiv:2312.05603
-
Exploring the Limits of ChatGPT in Software Security Applications 8 Dec 2023 · 0 repositories · arXiv:2312.05275
-
LLM Interactive Optimization of Open Source Python Libraries -- Case Studies and Generalization 8 Dec 2023 · 0 repositories · arXiv:2312.14949
-
Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs 8 Dec 2023 · 0 repositories · arXiv:2312.04782
-
Prospective Role of Foundation Models in Advancing Autonomous Vehicles 8 Dec 2023 · 0 repositories · arXiv:2405.02288
-
User-Aware Prefix-Tuning is a Good Learner for Personalized Image Captioning 8 Dec 2023 · 0 repositories · arXiv:2312.04793
-
On Sarcasm Detection with OpenAI GPT-based Models 7 Dec 2023 · 0 repositories · arXiv:2312.04642
-
Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models 7 Dec 2023 · 0 repositories · arXiv:2312.04724
-
Holmes: Towards Distributed Training Across Clusters with Heterogeneous NIC Environment 6 Dec 2023 · 0 repositories · arXiv:2312.03549
-
Exploring the Reversal Curse and Other Deductive Logical Reasoning in BERT and GPT-Based Large Language Models 6 Dec 2023 · 1 repository · arXiv:2312.03633
-
A Hardware Evaluation Framework for Large Language Model Inference 5 Dec 2023 · 0 repositories · arXiv:2312.03134
-
DRAFT: Dense Retrieval Augmented Few-shot Topic classifier Framework 5 Dec 2023 · 1 repository · arXiv:2312.02532
-
GPT vs Human for Scientific Reviews: A Dual Source Review on Applications of ChatGPT in Science 5 Dec 2023 · 0 repositories · arXiv:2312.03769
-
Rank-without-GPT: Building GPT-Independent Listwise Rerankers on Open-Source Large Language Models 5 Dec 2023 · 0 repositories · arXiv:2312.02969
-
Towards More Unified In-context Visual Understanding 5 Dec 2023 · 0 repositories · arXiv:2312.02520
-
A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly 4 Dec 2023 · 0 repositories · arXiv:2312.02003
-
Jellyfish: A Large Language Model for Data Preprocessing 4 Dec 2023 · 0 repositories · arXiv:2312.01678
-
Tree of Attacks: Jailbreaking Black-Box LLMs Automatically 4 Dec 2023 · 2 repositories · arXiv:2312.02119Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
NLEBench+NorGLM: A Comprehensive Empirical Analysis and Benchmark Dataset for Generative Language Models in Norwegian 3 Dec 2023 · 1 repository · arXiv:2312.01314Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
A ripple in time: a discontinuity in American history 2 Dec 2023 · 1 repository · arXiv:2312.01185
-
Harnessing the Power of Prompt-based Techniques for Generating School-Level Questions using Large Language Models 2 Dec 2023 · 1 repository · arXiv:2312.01032
-
Large Language Models Are Zero-Shot Text Classifiers 2 Dec 2023 · 1 repository · arXiv:2312.01044
-
Toward Improving Robustness of Object Detectors Against Domain Shift 2 Dec 2023 · 1 repository · arXiv:2403.12049
-
Generative Parameter-Efficient Fine-Tuning 1 Dec 2023 · 1 repository · arXiv:2312.00700Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples) · 10 pointer-only (licence)
-
Applying Large Language Models and Chain-of-Thought for Automatic Scoring 30 Nov 2023 · 0 repositories · arXiv:2312.03748
-
IAG: Induction-Augmented Generation Framework for Answering Reasoning Questions 30 Nov 2023 · 0 repositories · arXiv:2311.18397
-
Robust Concept Erasure via Kernelized Rate-Distortion Maximization 30 Nov 2023 · 1 repository · arXiv:2312.00194Syntology official (archive's flag): 20 ran · 20 ran (of which 0 constructed an object rather than computing a result; 18 with no instrument failure: 0 honoured, 0 violated, 18 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 21 harvested samples) · 2 pointer-only (licence)
-
Biomedical knowledge graph-optimized prompt generation for large language models 29 Nov 2023 · 1 repository · arXiv:2311.17330Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Improving the Robustness of Transformer-based Large Language Models with Dynamic Attention 29 Nov 2023 · 0 repositories · arXiv:2311.17400
-
TimelyGPT: Extrapolatable Transformer Pre-training for Long-term Time-Series Forecasting in Healthcare 29 Nov 2023 · 0 repositories · arXiv:2312.00817
-
CharacterGLM: Customizing Chinese Conversational AI Characters with Large Language Models 28 Nov 2023 · 1 repository · arXiv:2311.16832Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ChatGPT's One-year Anniversary: Are Open-Source Large Language Models Catching up? 28 Nov 2023 · 1 repository · arXiv:2311.16989
-
COLE: A Hierarchical Generation Framework for Multi-Layered and Editable Graphic Design 28 Nov 2023 · 0 repositories · arXiv:2311.16974
-
Comparing Generative Chatbots Based on Process Requirements 28 Nov 2023 · 0 repositories · arXiv:2312.03741
-
SEED-Bench-2: Benchmarking Multimodal Large Language Models 28 Nov 2023 · 2 repositories · arXiv:2311.17092
-
BERT Goes Off-Topic: Investigating the Domain Transfer Challenge using Genre Classification 27 Nov 2023 · 1 repository · arXiv:2311.16083
-
Decoding Logic Errors: A Comparative Study on Bug Detection by Students and Large Language Models 27 Nov 2023 · 0 repositories · arXiv:2311.16017
-
MEDITRON-70B: Scaling Medical Pretraining for Large Language Models 27 Nov 2023 · 1 repository · arXiv:2311.16079Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples)
-
Real Customization or Just Marketing: Are Customized Versions of Chat GPT Useful? 27 Nov 2023 · 0 repositories · arXiv:2312.03728
-
Leveraging AI-derived Data for Carbon Accounting: Information Extraction from Alternative Sources 26 Nov 2023 · 0 repositories · arXiv:2312.03722
-
Machine-Generated Text Detection using Deep Learning 26 Nov 2023 · 1 repository · arXiv:2311.15425
-
UHGEval: Benchmarking the Hallucination of Chinese Large Language Models via Unconstrained Generation 26 Nov 2023 · 1 repository · arXiv:2311.15296Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
CMed-GPT: Prompt Tuning for Entity-Aware Chinese Medical Dialogue Generation 24 Nov 2023 · 0 repositories · arXiv:2311.14539
-
Data-to-Text Bilingual Generation 24 Nov 2023 · 2 repositories · arXiv:2311.14808
-
GPT Struct Me: Probing GPT Models on Narrative Entity Extraction 24 Nov 2023 · 1 repository · arXiv:2311.14583
-
Large Language Models as Automated Aligners for benchmarking Vision-Language Models 24 Nov 2023 · 0 repositories · arXiv:2311.14580
-
Machine Translation for Ge'ez Language 24 Nov 2023 · 0 repositories · arXiv:2311.14530
-
A Cross Attention Approach to Diagnostic Explainability using Clinical Practice Guidelines for Depression 23 Nov 2023 · 1 repository · arXiv:2311.13852
-
Hardware Resilience Properties of Text-Guided Image Classifiers 23 Nov 2023 · 1 repository · arXiv:2311.14062Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Minimizing Factual Inconsistency and Hallucination in Large Language Models 23 Nov 2023 · 0 repositories · arXiv:2311.13878
-
Towards Auditing Large Language Models: Improving Text-based Stereotype Detection 23 Nov 2023 · 0 repositories · arXiv:2311.14126
-
Comparison of pipeline, sequence-to-sequence, and GPT models for end-to-end relation extraction: experiments with the rare disease use-case 22 Nov 2023 · 1 repository · arXiv:2311.13729
-
Drilling Down into the Discourse Structure with LLMs for Long Document Question Answering 22 Nov 2023 · 0 repositories · arXiv:2311.13565
-
Generation of Explanations for Logic Reasoning 22 Nov 2023 · 0 repositories · arXiv:2311.13455
-
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning 22 Nov 2023 · 0 repositories · arXiv:2311.13721
-
PG-Video-LLaVA: Pixel Grounding Large Video-Language Models 22 Nov 2023 · 1 repository · arXiv:2311.13435Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
AlignedCoT: Prompting Large Language Models via Native-Speaking Demonstrations 22 Nov 2023 · 1 repository · arXiv:2311.13538
-
@ve: A Chatbot for Latin 22 Nov 2023 · 0 repositories · arXiv:2311.14741
-
A Survey on Large Language Models for Personalized and Explainable Recommendations 21 Nov 2023 · 0 repositories · arXiv:2311.12338
-
AcademicGPT: Empowering Academic Research 21 Nov 2023 · 0 repositories · arXiv:2311.12315
-
ALPHA: AnomaLous Physiological Health Assessment Using Large Language Models 21 Nov 2023 · 1 repository · arXiv:2311.12524
-
Descriptor and Word Soups: Overcoming the Parameter Efficiency Accuracy Tradeoff for Out-of-Distribution Few-shot Learning 21 Nov 2023 · 1 repository · arXiv:2311.13612
-
Extracting Definienda in Mathematical Scholarly Articles with Transformers 21 Nov 2023 · 2 repositories · arXiv:2311.12448
-
GPT4Motion: Scripting Physical Motions in Text-to-Video Generation via Blender-Oriented GPT Planning 21 Nov 2023 · 0 repositories · arXiv:2311.12631
-
InterPrompt: Interpretable Prompting for Interrelated Interpersonal Risk Factors in Reddit Posts 21 Nov 2023 · 0 repositories · arXiv:2311.12404
-
Assessing Prompt Injection Risks in 200+ Custom GPTs 20 Nov 2023 · 1 repository · arXiv:2311.11538
-
Evil Geniuses: Delving into the Safety of LLM-based Agents 20 Nov 2023 · 1 repository · arXiv:2311.11855Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Towards Human-Level Text Coding with LLMs: The Case of Fatherhood Roles in Public Policy Documents 20 Nov 2023 · 1 repository · arXiv:2311.11844
-
MemoryCompanion: A Smart Healthcare Solution to Empower Efficient Alzheimer's Care Via Unleashing Generative AI 20 Nov 2023 · 0 repositories · arXiv:2311.14730
-
Refactoring Programs Using Large Language Models with Few-Shot Examples 20 Nov 2023 · 0 repositories · arXiv:2311.11690
-
Spot the Bot: Distinguishing Human-Written and Bot-Generated Texts Using Clustering and Information Theory Techniques 19 Nov 2023 · 0 repositories · arXiv:2311.11441
-
Behavior Optimized Image Generation 18 Nov 2023 · 0 repositories · arXiv:2311.10995
-
Advancements in Generative AI: A Comprehensive Review of GANs, GPT, Autoencoders, Diffusion Model, and Transformers 17 Nov 2023 · 0 repositories · arXiv:2311.10242
-
Bias A-head? Analyzing Bias in Transformer-Based Language Model Attention Heads 17 Nov 2023 · 0 repositories · arXiv:2311.10395
-
Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2 17 Nov 2023 · 3 repositories · arXiv:2311.10702
-
DynaPipe: Optimizing Multi-task Training through Dynamic Pipelines 17 Nov 2023 · 2 repositories · arXiv:2311.10418
-
Event Causality Is Key to Computational Story Understanding 16 Nov 2023 · 1 repository · arXiv:2311.09648Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)