Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 22
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 22 of 40: papers 2,101 to 2,200 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LLM4VV: Developing LLM-Driven Testsuite for Compiler Validation 8 Oct 2023 · 1 repository · arXiv:2310.04963
-
Zero-Shot Detection of Machine-Generated Codes 8 Oct 2023 · 1 repository · arXiv:2310.05103
-
Do self-supervised speech and language models extract similar representations as human brain? 7 Oct 2023 · 0 repositories · arXiv:2310.04645
-
Large Language Models Only Pass Primary School Exams in Indonesia: A Comprehensive Test on IndoMMLU 7 Oct 2023 · 1 repository · arXiv:2310.04928Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT 7 Oct 2023 · 2 repositories · arXiv:2310.04673Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Question-focused Summarization by Decomposing Articles into Facts and Opinions and Retrieving Entities 7 Oct 2023 · 0 repositories · arXiv:2310.04880
-
Copy Suppression: Comprehensively Understanding an Attention Head 6 Oct 2023 · 1 repository · arXiv:2310.04625Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Keyword Augmented Retrieval: Novel framework for Information Retrieval integrated with speech interface 6 Oct 2023 · 0 repositories · arXiv:2310.04205
-
Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models 6 Oct 2023 · 2 repositories · arXiv:2310.04406Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Agent Instructs Large Language Models to be General Zero-Shot Reasoners 5 Oct 2023 · 1 repository · arXiv:2310.03710Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Automating Human Tutor-Style Programming Feedback: Leveraging GPT-4 Tutor Model for Hint Generation and GPT-3.5 Student Model for Hint Validation 5 Oct 2023 · 2 repositories · arXiv:2310.03780Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines 5 Oct 2023 · 3 repositories · arXiv:2310.03714Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To! 5 Oct 2023 · 1 repository · arXiv:2310.03693Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Parking Spot Classification based on surround view camera system 5 Oct 2023 · 0 repositories · arXiv:2310.12997
-
SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks 5 Oct 2023 · 1 repository · arXiv:2310.03684
-
A Survey of GPT-3 Family Large Language Models Including ChatGPT and GPT-4 4 Oct 2023 · 0 repositories · arXiv:2310.12321
-
Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning 4 Oct 2023 · 1 repository · arXiv:2310.03094
-
Memoria: Resolving Fateful Forgetting Problem through Human-Inspired Memory Architecture 4 Oct 2023 · 1 repository · arXiv:2310.03052Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 10 harvested samples)
-
NOLA: Compressing LoRA using Linear Combination of Random Basis 4 Oct 2023 · 1 repository · arXiv:2310.02556Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Retrieval meets Long Context Large Language Models 4 Oct 2023 · 0 repositories · arXiv:2310.03025
-
Instances Need More Care: Rewriting Prompts for Instances with LLMs in the Loop Yields Better Zero-Shot Performance 3 Oct 2023 · 1 repository · arXiv:2310.02107Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
GPT-Driver: Learning to Drive with GPT 2 Oct 2023 · 1 repository · arXiv:2310.01415Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
LLM Lies: Hallucinations are not Bugs, but Features as Adversarial Examples 2 Oct 2023 · 1 repository · arXiv:2310.01469Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
PolySketchFormer: Fast Transformers via Sketching Polynomial Kernels 2 Oct 2023 · 0 repositories · arXiv:2310.01655
-
Analyzing and Mitigating Object Hallucination in Large Vision-Language Models 1 Oct 2023 · 1 repository · arXiv:2310.00754Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
BooookScore: A systematic exploration of book-length summarization in the era of LLMs 1 Oct 2023 · 2 repositories · arXiv:2310.00785Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models 1 Oct 2023 · 2 repositories · arXiv:2310.00746
-
Gaze-Driven Sentence Simplification for Language Learners: Enhancing Comprehension and Readability 30 Sep 2023 · 0 repositories · arXiv:2310.00355
-
A Large Language Model Approach to Educational Survey Feedback Analysis 29 Sep 2023 · 0 repositories · arXiv:2309.17447
-
An evaluation of GPT models for phenotype concept recognition 29 Sep 2023 · 0 repositories · arXiv:2309.17169
-
Benchmarking the Abilities of Large Language Models for RDF Knowledge Graph Creation and Comprehension: How Well Do LLMs Speak Turtle? 29 Sep 2023 · 3 repositories · arXiv:2309.17122
-
DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks 29 Sep 2023 · 1 repository · arXiv:2309.17167
-
Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation 29 Sep 2023 · 2 repositories · arXiv:2309.17234Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Memory Gym: Towards Endless Tasks to Benchmark Memory Capabilities of Agents 29 Sep 2023 · 1 repository · arXiv:2309.17207Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Revolutionizing Mobile Interaction: Enabling a 3 Billion Parameter GPT LLM on Mobile 29 Sep 2023 · 0 repositories · arXiv:2310.01434
-
Split and Merge: Aligning Position Biases in LLM-based Evaluators 29 Sep 2023 · 0 repositories · arXiv:2310.01432
-
Training and inference of large language models using 8-bit floating point 29 Sep 2023 · 0 repositories · arXiv:2309.17224
-
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events 28 Sep 2023 · 0 repositories · arXiv:2309.16150
-
GPT-Fathom: Benchmarking Large Language Models to Decipher the Evolutionary Path towards GPT-4 and Beyond 28 Sep 2023 · 1 repository · arXiv:2309.16583Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Large Language Model Soft Ideologization via AI-Self-Consciousness 28 Sep 2023 · 0 repositories · arXiv:2309.16167
-
Stress Testing Chain-of-Thought Prompting for Large Language Models 28 Sep 2023 · 0 repositories · arXiv:2309.16621
-
MindGPT: Interpreting What You See with Non-invasive Brain Recordings 27 Sep 2023 · 1 repository · arXiv:2309.15729Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
NLPBench: Evaluating Large Language Models on Solving NLP Problems 27 Sep 2023 · 1 repository · arXiv:2309.15630
-
Legal Question-Answering in the Indian Context: Efficacy, Challenges, and Potential of Modern AI Models 26 Sep 2023 · 0 repositories · arXiv:2309.14735
-
Exploring Small Language Models with Prompt-Learning Paradigm for Efficient Domain-Specific Text Classification 26 Sep 2023 · 0 repositories · arXiv:2309.14779
-
How to Catch an AI Liar: Lie Detection in Black-Box LLMs by Asking Unrelated Questions 26 Sep 2023 · 1 repository · arXiv:2309.15840Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models 26 Sep 2023 · 3 repositories · arXiv:2309.15088
-
Supersonic: Learning to Generate Source Code Optimizations in C/C++ 26 Sep 2023 · 1 repository · arXiv:2309.14846
-
Evaluating Cognitive Maps and Planning in Large Language Models with CogEval 25 Sep 2023 · 0 repositories · arXiv:2309.15129
-
LogGPT: Log Anomaly Detection via GPT 25 Sep 2023 · 1 repository · arXiv:2309.14482
-
Watch Your Language: Investigating Content Moderation with Large Language Models 25 Sep 2023 · 0 repositories · arXiv:2309.14517
-
Does the "most sinfully decadent cake ever" taste good? Answering Yes/No Questions from Figurative Contexts 24 Sep 2023 · 0 repositories · arXiv:2309.13748
-
Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models 24 Sep 2023 · 1 repository · arXiv:2309.13579
-
A Chat About Boring Problems: Studying GPT-based text normalization 23 Sep 2023 · 0 repositories · arXiv:2309.13426
-
Probing the Moral Development of Large Language Models through Defining Issues Test 23 Sep 2023 · 0 repositories · arXiv:2309.13356
-
Randomize to Generalize: Domain Randomization for Runway FOD Detection 23 Sep 2023 · 0 repositories · arXiv:2309.13264
-
AMPLIFY:Attention-based Mixup for Performance Improvement and Label Smoothing in Transformer 22 Sep 2023 · 1 repository · arXiv:2309.12689
-
BenLLMEval: A Comprehensive Evaluation into the Potentials and Pitfalls of Large Language Models on Bengali NLP 22 Sep 2023 · 0 repositories · arXiv:2309.13173
-
Contextual Emotion Estimation from Image Captions 22 Sep 2023 · 0 repositories · arXiv:2309.13136
-
Investigating Large Language Models and Control Mechanisms to Improve Text Readability of Biomedical Abstracts 22 Sep 2023 · 1 repository · arXiv:2309.13202
-
Large Language Models Are Also Good Prototypical Commonsense Reasoners 22 Sep 2023 · 0 repositories · arXiv:2309.13165
-
SPION: Layer-Wise Sparse Training of Transformer via Convolutional Flood Filling 22 Sep 2023 · 0 repositories · arXiv:2309.12578
-
Goal-Oriented Prompt Attack and Safety Evaluation for LLMs 21 Sep 2023 · 2 repositories · arXiv:2309.11830
-
Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection 21 Sep 2023 · 1 repository · arXiv:2309.12247
-
Constraints First: A New MDD-based Model to Generate Sentences Under Constraints 21 Sep 2023 · 0 repositories · arXiv:2309.12415
-
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models 21 Sep 2023 · 1 repository · arXiv:2309.12284Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 14 where Syntology's instrument failed) · 7 unverified (of 22 harvested samples)
-
Random-Access Infinite Context Length for Transformers 21 Sep 2023 · 1 repository
-
TART: A plug-and-play Transformer module for task-agnostic reasoning 21 Sep 2023 · 1 repository
-
The Cambridge Law Corpus: A Dataset for Legal AI Research 21 Sep 2023 · 0 repositories · arXiv:2309.12269
-
The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A" 21 Sep 2023 · 2 repositories · arXiv:2309.12288Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
TOA: Task-oriented Active VQA 21 Sep 2023 · 0 repositories
-
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11674Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Controlled Generation with Prompt Insertion for Natural Language Explanations in Grammatical Error Correction 20 Sep 2023 · 1 repository · arXiv:2309.11439
-
Design of Chain-of-Thought in Math Problem Solving 20 Sep 2023 · 1 repository · arXiv:2309.11054
-
Fictional Worlds, Real Connections: Developing Community Storytelling Social Chatbots through LLMs 20 Sep 2023 · 0 repositories · arXiv:2309.11478
-
Generative AI in Mafia-like Game Simulation 20 Sep 2023 · 0 repositories · arXiv:2309.11672
-
Safurai 001: New Qualitative Approach for Code LLM Evaluation 20 Sep 2023 · 1 repository · arXiv:2309.11385
-
Sequence-to-Sequence Spanish Pre-trained Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11259
-
The Languini Kitchen: Enabling Language Modelling Research at Different Scales of Compute 20 Sep 2023 · 1 repository · arXiv:2309.11197
-
Language as the Medium: Multimodal Video Classification through text only 19 Sep 2023 · 0 repositories · arXiv:2309.10783
-
Rigorously Assessing Natural Language Explanations of Neurons 19 Sep 2023 · 0 repositories · arXiv:2309.10312
-
Writer-Defined AI Personas for On-Demand Feedback Generation 19 Sep 2023 · 0 repositories · arXiv:2309.10433
-
Evaluation of GPT-3 for Anti-Cancer Drug Sensitivity Prediction 18 Sep 2023 · 0 repositories · arXiv:2309.10016
-
RECAP: Retrieval-Augmented Audio Captioning 18 Sep 2023 · 1 repository · arXiv:2309.09836Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Towards Ontology Construction with Language Models 18 Sep 2023 · 0 repositories · arXiv:2309.09898
-
Contrastive Decoding Improves Reasoning in Large Language Models 17 Sep 2023 · 0 repositories · arXiv:2309.09117
-
Do Large GPT Models Discover Moral Dimensions in Language Representations? A Topological Study Of Sentence Embeddings 17 Sep 2023 · 0 repositories · arXiv:2309.09397
-
From Cooking Recipes to Robot Task Trees -- Improving Planning Correctness and Task Efficiency by Leveraging LLMs with a Knowledge Network 17 Sep 2023 · 0 repositories · arXiv:2309.09181
-
Decoder-only Architecture for Speech Recognition with CTC Prompts and Text Data Augmentation 16 Sep 2023 · 0 repositories · arXiv:2309.08876
-
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 16 Sep 2023 · 1 repository · arXiv:2309.08963
-
A Modern Turkish Poet: Fine-Tuned GPT-2 15 Sep 2023 · 1 repository
-
Advancing the Evaluation of Traditional Chinese Language Models: Towards a Comprehensive Benchmark Suite 15 Sep 2023 · 1 repository · arXiv:2309.08448
-
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models 15 Sep 2023 · 1 repository · arXiv:2309.08573
-
EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers 15 Sep 2023 · 2 repositories · arXiv:2309.08532Syntology 11 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending 15 Sep 2023 · 1 repository · arXiv:2309.08646
-
GPT-Lab: Next Generation Of Optimal Chemistry Discovery By GPT Driven Robotic Lab 15 Sep 2023 · 0 repositories · arXiv:2309.16721
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
Large Language Models for Failure Mode Classification: An Investigation 15 Sep 2023 · 1 repository · arXiv:2309.08181
-
An Empirical Evaluation of Prompting Strategies for Large Language Models in Zero-Shot Clinical Natural Language Processing 14 Sep 2023 · 0 repositories · arXiv:2309.08008
-
Assessing the nature of large language models: A caution against anthropocentrism 14 Sep 2023 · 0 repositories · arXiv:2309.07683