Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 26
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 26 of 40: papers 2,501 to 2,600 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Harnessing Explanations: LLM-to-LM Interpreter for Enhanced Text-Attributed Graph Representation Learning 31 May 2023 · 3 repositories · arXiv:2305.19523Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Knowledge Base Question Answering for Space Debris Queries 31 May 2023 · 1 repository · arXiv:2305.19734
-
Does Conceptual Representation Require Embodiment? Insights From Large Language Models 30 May 2023 · 0 repositories · arXiv:2305.19103
-
Generate then Select: Open-ended Visual Question Answering Guided by World Knowledge 30 May 2023 · 0 repositories · arXiv:2305.18842
-
GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation 30 May 2023 · 0 repositories · arXiv:2305.18997
-
Seeing Seeds Beyond Weeds: Green Teaming Generative AI for Beneficial Uses 30 May 2023 · 0 repositories · arXiv:2306.03097
-
Check-COVID: Fact-Checking COVID-19 News Claims with Scientific Evidence 29 May 2023 · 1 repository · arXiv:2305.18265
-
Coeditor: Leveraging Contextual Changes for Multi-round Code Auto-editing 29 May 2023 · 0 repositories · arXiv:2305.18584
-
Do Large Language Models Know What They Don't Know? 29 May 2023 · 1 repository · arXiv:2305.18153Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Exploring Effectiveness of GPT-3 in Grammatical Error Correction: A Study on Performance and Controllability in Prompt-Based Methods 29 May 2023 · 0 repositories · arXiv:2305.18156
-
LM-CPPF: Paraphrasing-Guided Data Augmentation for Contrastive Prompt-Based Few-Shot Fine-Tuning 29 May 2023 · 1 repository · arXiv:2305.18169
-
Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models 29 May 2023 · 1 repository · arXiv:2305.18189Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ProcessGPT: Transforming Business Process Management with Generative Artificial Intelligence 29 May 2023 · 0 repositories · arXiv:2306.01771
-
Syntax and Semantics Meet in the "Middle": Probing the Syntax-Semantics Interface of LMs Through Agentivity 29 May 2023 · 1 repository · arXiv:2305.18185
-
Test-Time Training on Nearest Neighbors for Large Language Models 29 May 2023 · 1 repository · arXiv:2305.18466Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Transformer Language Models Handle Word Frequency in Prediction Head 29 May 2023 · 0 repositories · arXiv:2305.18294
-
Bridging the Language Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs 28 May 2023 · 0 repositories · arXiv:2305.17740
-
Evaluating GPT-3 Generated Explanations for Hateful Content Moderation 28 May 2023 · 1 repository · arXiv:2305.17680
-
Generating EDU Extracts for Plan-Guided Summary Re-Ranking 28 May 2023 · 1 repository · arXiv:2305.17779Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Knowledge-Augmented Reasoning Distillation for Small Language Models in Knowledge-Intensive Tasks 28 May 2023 · 1 repository · arXiv:2305.18395Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model Application 28 May 2023 · 1 repository · arXiv:2305.17701
-
Mitigating Label Biases for In-context Learning 28 May 2023 · 1 repository · arXiv:2305.19148Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SQuARe: A Large-Scale Dataset of Sensitive Questions and Acceptable Responses Created Through Human-Machine Collaboration 28 May 2023 · 1 repository · arXiv:2305.17696
-
Transfer Learning for Power Outage Detection Task with Limited Training Data 28 May 2023 · 0 repositories · arXiv:2305.17817
-
Complementary and Integrative Health Lexicon (CIHLex) and Entity Recognition in the Literature 27 May 2023 · 0 repositories · arXiv:2305.17353
-
DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text 27 May 2023 · 1 repository · arXiv:2305.17359Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
The Curse of Recursion: Training on Generated Data Makes Models Forget 27 May 2023 · 1 repository · arXiv:2305.17493
-
Towards Explainable Conversational Recommender Systems 27 May 2023 · 1 repository · arXiv:2305.18363
-
What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks 27 May 2023 · 1 repository · arXiv:2305.18365
-
Backpack Language Models 26 May 2023 · 1 repository · arXiv:2305.16765Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance 26 May 2023 · 1 repository · arXiv:2305.17306Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
ChatGPT: A Study on its Utility for Ubiquitous Software Engineering Tasks 26 May 2023 · 0 repositories · arXiv:2305.16837
-
Counterfactual reasoning: Testing language models' understanding of hypothetical scenarios 26 May 2023 · 1 repository · arXiv:2305.16572
-
DeepSeaNet: Improving Underwater Object Detection using EfficientDet 26 May 2023 · 0 repositories · arXiv:2306.06075
-
Distinguishing Human Generated Text From ChatGPT Generated Text Using Machine Learning 26 May 2023 · 0 repositories · arXiv:2306.01761
-
Do GPTs Produce Less Literal Translations? 26 May 2023 · 1 repository · arXiv:2305.16806
-
Evaluation of Question Generation Needs More References 26 May 2023 · 0 repositories · arXiv:2305.16626
-
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing 26 May 2023 · 0 repositories · arXiv:2305.16635
-
Improving accuracy of GPT-3/4 results on biomedical data using a retrieval-augmented language model 26 May 2023 · 0 repositories · arXiv:2305.17116
-
Large Language Models as Tool Makers 26 May 2023 · 1 repository · arXiv:2305.17126
-
Learning and Leveraging Verifiers to Improve Planning Capabilities of Pre-trained Language Models 26 May 2023 · 0 repositories · arXiv:2305.17077
-
LLMs and the Abstraction and Reasoning Corpus: Successes, Failures, and the Importance of Object-based Representations 26 May 2023 · 1 repository · arXiv:2305.18354
-
NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models 26 May 2023 · 2 repositories · arXiv:2305.16986Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Playing repeated games with Large Language Models 26 May 2023 · 0 repositories · arXiv:2305.16867
-
A Survey on ChatGPT: AI-Generated Contents, Challenges, and Solutions 25 May 2023 · 0 repositories · arXiv:2305.18339
-
Landmark Attention: Random-Access Infinite Context Length for Transformers 25 May 2023 · 2 repositories · arXiv:2305.16300Syntology official (archive's flag): 1 ran · 11 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 6 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
Linguistic Properties of Truthful Response 25 May 2023 · 1 repository · arXiv:2305.15875
-
Not wacky vs. definitely wacky: A study of scalar adverbs in pretrained language models 25 May 2023 · 0 repositories · arXiv:2305.16426
-
A Causal View of Entity Bias in (Large) Language Models 24 May 2023 · 1 repository · arXiv:2305.14695Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LAraBench: Benchmarking Arabic AI with Large Language Models 24 May 2023 · 0 repositories · arXiv:2305.14982
-
Chain-of-Questions Training with Latent Answers for Robust Multistep Question Answering 24 May 2023 · 0 repositories · arXiv:2305.14901
-
ChatAgri: Exploring Potentials of ChatGPT on Cross-linguistic Agricultural Text Classification 24 May 2023 · 1 repository · arXiv:2305.15024
-
Don't Take This Out of Context! On the Need for Contextual Models and Evaluations for Stylistic Rewriting 24 May 2023 · 0 repositories · arXiv:2305.14755
-
Don't Trust ChatGPT when Your Question is not in English: A Study of Multilingual Abilities and Types of LLMs 24 May 2023 · 0 repositories · arXiv:2305.16339
-
Editing Common Sense in Transformers 24 May 2023 · 1 repository · arXiv:2305.14956Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
ExpertPrompting: Instructing Large Language Models to be Distinguished Experts 24 May 2023 · 2 repositories · arXiv:2305.14688Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Harnessing the Power of Large Language Models for Natural Language to First-Order Logic Translation 24 May 2023 · 1 repository · arXiv:2305.15541Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Have LLMs Advanced Enough? A Challenging Problem Solving Benchmark For Large Language Models 24 May 2023 · 1 repository · arXiv:2305.15074Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Psychological Metrics for Dialog System Evaluation 24 May 2023 · 0 repositories · arXiv:2305.14757
-
I Spy a Metaphor: Large Language Models and Diffusion Models Co-Create Visual Metaphors 24 May 2023 · 1 repository · arXiv:2305.14724
-
Inference-Time Policy Adapters (IPA): Tailoring Extreme-Scale LMs without Fine-tuning 24 May 2023 · 1 repository · arXiv:2305.15065Syntology official (archive's flag): 9 ran · 9 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
Enabling and Analyzing How to Efficiently Extract Information from Hybrid Long Documents with LLMs 24 May 2023 · 0 repositories · arXiv:2305.16344
-
LLMDet: A Third Party Large Language Models Generated Text Detection Tool 24 May 2023 · 1 repository · arXiv:2305.15004Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Mastering the ABCDs of Complex Questions: Answer-Based Claim Decomposition for Fine-grained Self-Evaluation 24 May 2023 · 0 repositories · arXiv:2305.14750
-
Peek Across: Improving Multi-Document Modeling via Cross-Document Question-Answering 24 May 2023 · 1 repository · arXiv:2305.15387Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 18 harvested samples)
-
Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models 24 May 2023 · 1 repository · arXiv:2305.14623Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Testing Causal Models of Word Meaning in GPT-3 and -4 24 May 2023 · 1 repository · arXiv:2305.14630
-
ToMChallenges: A Principle-Guided Dataset and Diverse Evaluation Tasks for Exploring Theory of Mind 24 May 2023 · 1 repository · arXiv:2305.15068
-
Tricking LLMs into Disobedience: Formalizing, Analyzing, and Detecting Jailbreaks 24 May 2023 · 1 repository · arXiv:2305.14965Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Trusting Your Evidence: Hallucinate Less with Context-aware Decoding 24 May 2023 · 3 repositories · arXiv:2305.14739Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
A Trip Towards Fairness: Bias and De-Biasing in Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.13862
-
Active Learning Principles for In-Context Learning with Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.14264
-
Fine-tuned LLMs Know More, Hallucinate Less with Few-Shot Sequence-to-Sequence Semantic Parsing over Wikidata 23 May 2023 · 1 repository · arXiv:2305.14202
-
Dancing Between Success and Failure: Edit-level Simplification Evaluation using SALSA 23 May 2023 · 0 repositories · arXiv:2305.14458
-
Deduction under Perturbed Evidence: Probing Student Simulation Capabilities of Large Language Models 23 May 2023 · 0 repositories · arXiv:2305.14507
-
Dynosaur: A Dynamic Growth Paradigm for Instruction-Tuning Data Curation 23 May 2023 · 1 repository · arXiv:2305.14327Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
Few-Shot Data Synthesis for Open Domain Multi-Hop Question Answering 23 May 2023 · 0 repositories · arXiv:2305.13691
-
Enhancing Black-Box Few-Shot Text Classification with Prompt-Based Data Augmentation 23 May 2023 · 0 repositories · arXiv:2305.13785
-
Advancing Precise Outline-Conditioned Text Generation with Task Duality and Explicit Outline Control 23 May 2023 · 0 repositories · arXiv:2305.14459
-
Evaluating Factual Consistency of Summaries with Large Language Models 23 May 2023 · 2 repositories · arXiv:2305.14069
-
HumBEL: A Human-in-the-Loop Approach for Evaluating Demographic Factors of Language Models in Human-Machine Conversations 23 May 2023 · 1 repository · arXiv:2305.14195
-
IfQA: A Dataset for Open-domain Question Answering under Counterfactual Presuppositions 23 May 2023 · 0 repositories · arXiv:2305.14010
-
Images in Language Space: Exploring the Suitability of Large Language Models for Vision & Language Tasks 23 May 2023 · 1 repository · arXiv:2305.13782
-
INSTRUCTSCORE: Explainable Text Generation Evaluation with Finegrained Feedback 23 May 2023 · 2 repositories · arXiv:2305.14282
-
Let's Think Frame by Frame with VIP: A Video Infilling and Prediction Dataset for Evaluating Video Chain-of-Thought 23 May 2023 · 1 repository · arXiv:2305.13903
-
MathDial: A Dialogue Tutoring Dataset with Rich Pedagogical Properties Grounded in Math Reasoning Problems 23 May 2023 · 1 repository · arXiv:2305.14536Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
NAIL: Lexical Retrieval Indices with Efficient Non-Autoregressive Decoders 23 May 2023 · 0 repositories · arXiv:2305.14499
-
NarrativeXL: A Large-scale Dataset For Long-Term Memory Models 23 May 2023 · 1 repository · arXiv:2305.13877
-
On Robustness of Finetuned Transformer-based NLP Models 23 May 2023 · 1 repository · arXiv:2305.14453
-
Physics of Language Models: Part 1, Learning Hierarchical Language Structures 23 May 2023 · 0 repositories · arXiv:2305.13673
-
Probing Brain Context-Sensitivity with Masked-Attention Generation 23 May 2023 · 0 repositories · arXiv:2305.13863
-
Sophia: A Scalable Stochastic Second-order Optimizer for Language Model Pre-training 23 May 2023 · 7 repositories · arXiv:2305.14342Syntology 13 ran (of which 4 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 19 harvested samples)
-
Sources of Hallucination by Large Language Models on Inference Tasks 23 May 2023 · 1 repository · arXiv:2305.14552
-
Training Transitive and Commutative Multimodal Transformers with LoReTTa 23 May 2023 · 0 repositories · arXiv:2305.14243
-
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs 23 May 2023 · 0 repositories · arXiv:2305.14279
-
WikiChat: Stopping the Hallucination of Large Language Model Chatbots by Few-Shot Grounding on Wikipedia 23 May 2023 · 1 repository · arXiv:2305.14292Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Let GPT be a Math Tutor: Teaching Math Word Problem Solvers with Customized Exercise Generation 22 May 2023 · 0 repositories · arXiv:2305.14386
-
A Study of Generative Large Language Model for Medical Research and Healthcare 22 May 2023 · 1 repository · arXiv:2305.13523
-
Can Large Language Models emulate an inductive Thematic Analysis of semi-structured interviews? An exploration and provocation on the limits of the approach and the model 22 May 2023 · 1 repository · arXiv:2305.13014
-
Can LLMs facilitate interpretation of pre-trained language models? 22 May 2023 · 0 repositories · arXiv:2305.13386