Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 24
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 24 of 38: papers 2,301 to 2,400 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
New Interaction Paradigm for Complex EDA Software Leveraging GPT 27 Jul 2023 · 1 repository · arXiv:2307.14740Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
TextManiA: Enriching Visual Feature by Text-driven Manifold Augmentation 27 Jul 2023 · 0 repositories · arXiv:2307.14611
-
CliniDigest: A Case Study in Large Language Model Based Large-Scale Summarization of Clinical Trial Descriptions 26 Jul 2023 · 0 repositories · arXiv:2307.14522
-
How User Language Affects Conflict Fatality Estimates in ChatGPT 26 Jul 2023 · 0 repositories · arXiv:2308.00072
-
Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data 26 Jul 2023 · 1 repository · arXiv:2307.14385
-
GPT-3 Models are Few-Shot Financial Reasoners 25 Jul 2023 · 0 repositories · arXiv:2307.13617
-
Is GPT a Computational Model of Emotion? Detailed Analysis 25 Jul 2023 · 0 repositories · arXiv:2307.13779
-
Predicting Code Coverage without Execution 25 Jul 2023 · 1 repository · arXiv:2307.13383
-
How Does Naming Affect LLMs on Code Analysis Tasks? 24 Jul 2023 · 0 repositories · arXiv:2307.12488
-
Gradient-Based Word Substitution for Obstinate Adversarial Examples Generation in Language Models 24 Jul 2023 · 0 repositories · arXiv:2307.12507
-
The potential of LLMs for coding with low-resource and domain-specific programming languages 24 Jul 2023 · 0 repositories · arXiv:2307.13018
-
HateModerate: Testing Hate Speech Detectors against Content Moderation Policies 23 Jul 2023 · 1 repository · arXiv:2307.12418
-
Validation of a Zero-Shot Learning Natural Language Processing Tool for Data Abstraction from Unstructured Healthcare Data 23 Jul 2023 · 1 repository · arXiv:2308.00107
-
AIGC Empowering Telecom Sector White Paper_chinese 21 Jul 2023 · 0 repositories · arXiv:2307.11449
-
GPT-4 Can't Reason 21 Jul 2023 · 0 repositories · arXiv:2308.03762
-
An In-Depth Evaluation of Federated Learning on Biomedical Natural Language Processing 20 Jul 2023 · 2 repositories · arXiv:2307.11254
-
Generative Language Models on Nucleotide Sequences of Human Genes 20 Jul 2023 · 1 repository · arXiv:2307.10634
-
IvyGPT: InteractiVe Chinese pathwaY language model in medical domain 20 Jul 2023 · 1 repository · arXiv:2307.10512
-
LLM Cognitive Judgements Differ From Human 20 Jul 2023 · 1 repository · arXiv:2307.11787Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Of Models and Tin Men: A Behavioural Economics Study of Principal-Agent Problems in AI Alignment using Large-Language Models 20 Jul 2023 · 2 repositories · arXiv:2307.11137
-
Controlling Equational Reasoning in Large Language Models with Prompt Interventions 19 Jul 2023 · 0 repositories · arXiv:2307.09998
-
How is ChatGPT's behavior changing over time? 18 Jul 2023 · 4 repositories · arXiv:2307.09009Syntology official (archive's flag): 1 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 5 pointer-only (licence)
-
Unveiling Gender Bias in Terms of Profession Across LLMs: Analyzing and Addressing Sociological Implications 18 Jul 2023 · 0 repositories · arXiv:2307.09162
-
A mixed policy to improve performance of language models on math problems 17 Jul 2023 · 1 repository · arXiv:2307.08767
-
A Study on the Performance of Generative Pre-trained Transformer (GPT) in Simulating Depressed Individuals on the Standardized Depressive Symptom Scale 17 Jul 2023 · 0 repositories · arXiv:2307.08576
-
ChatGPT is Good but Bing Chat is Better for Vietnamese Students 17 Jul 2023 · 0 repositories · arXiv:2307.08272
-
GEAR: Augmenting Language Models with Generalizable and Efficient Tool Resolution 17 Jul 2023 · 1 repository · arXiv:2307.08775Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Using an LLM to Help With Code Understanding 17 Jul 2023 · 0 repositories · arXiv:2307.08177
-
Legal Syllogism Prompting: Teaching Large Language Models for Legal Judgment Prediction 17 Jul 2023 · 1 repository · arXiv:2307.08321
-
SentimentGPT: Exploiting GPT for Advanced Sentiment Analysis and its Departure from Current Machine Learning 16 Jul 2023 · 1 repository · arXiv:2307.10234
-
Coupling Large Language Models with Logic Programming for Robust and General Reasoning from Text 15 Jul 2023 · 1 repository · arXiv:2307.07696Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Large Language Models as Superpositions of Cultural Perspectives 15 Jul 2023 · 0 repositories · arXiv:2307.07870
-
Leveraging Large Language Models to Generate Answer Set Programs 15 Jul 2023 · 1 repository · arXiv:2307.07699Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Fairness of ChatGPT and the Role Of Explainable-Guided Prompts 14 Jul 2023 · 1 repository · arXiv:2307.11761
-
MorphPiece : A Linguistic Tokenizer for Large Language Models 14 Jul 2023 · 0 repositories · arXiv:2307.07262
-
A Study on Differentiable Logic and LLMs for EPIC-KITCHENS-100 Unsupervised Domain Adaptation Challenge for Action Recognition 2023 13 Jul 2023 · 0 repositories · arXiv:2307.06569
-
Agreement Tracking for Multi-Issue Negotiation Dialogues 13 Jul 2023 · 0 repositories · arXiv:2307.06524
-
Negated Complementary Commonsense using Large Language Models 13 Jul 2023 · 1 repository · arXiv:2307.06794
-
Ashaar: Automatic Analysis and Generation of Arabic Poetry Using Deep Learning Approaches 12 Jul 2023 · 1 repository · arXiv:2307.06218
-
Distilling Large Language Models for Biomedical Knowledge Extraction: A Case Study on Adverse Drug Events 12 Jul 2023 · 0 repositories · arXiv:2307.06439
-
Argumentative Segmentation Enhancement for Legal Summarization 11 Jul 2023 · 0 repositories · arXiv:2307.05081
-
DNAGPT: A Generalized Pre-trained Tool for Versatile DNA Sequence Analysis Tasks 11 Jul 2023 · 0 repositories · arXiv:2307.05628
-
Large Language Models 11 Jul 2023 · 0 repositories · arXiv:2307.05782
-
Named entity recognition using GPT for identifying comparable companies 11 Jul 2023 · 0 repositories · arXiv:2307.07420
-
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration 11 Jul 2023 · 2 repositories · arXiv:2307.05300
-
AmadeusGPT: a natural language interface for interactive animal behavioral analysis 10 Jul 2023 · 1 repository · arXiv:2307.04858
-
SimpleMTOD: A Simple Language Model for Multimodal Task-Oriented Dialogue with Symbolic Scene Representation 10 Jul 2023 · 0 repositories · arXiv:2307.04907
-
Assessing the efficacy of large language models in generating accurate teacher responses 9 Jul 2023 · 0 repositories · arXiv:2307.04274
-
A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation 8 Jul 2023 · 0 repositories · arXiv:2307.03987
-
DWReCO at CheckThat! 2023: Enhancing Subjectivity Detection through Style-based Data Sampling 7 Jul 2023 · 1 repository · arXiv:2307.03550
-
Goal-Conditioned Predictive Coding for Offline Reinforcement Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03406
-
How does AI chat change search behaviors? 7 Jul 2023 · 0 repositories · arXiv:2307.03826
-
RADAR: Robust AI-Text Detection via Adversarial Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03838
-
Improving Retrieval-Augmented Large Language Models via Data Importance Learning 6 Jul 2023 · 1 repository · arXiv:2307.03027Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Large Language Models Empowered Autonomous Edge AI for Connected Intelligence 6 Jul 2023 · 0 repositories · arXiv:2307.02779
-
Text Alignment Is An Efficient Unified Model for Massive NLP Tasks 6 Jul 2023 · 1 repository · arXiv:2307.02729
-
CAME: Confidence-guided Adaptive Memory Efficient Optimization 5 Jul 2023 · 2 repositories · arXiv:2307.02047Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations 5 Jul 2023 · 0 repositories · arXiv:2307.03678
-
External Reasoning: Towards Multi-Large-Language-Models Interchangeable Assistance with Human Feedback 5 Jul 2023 · 1 repository · arXiv:2307.12057
-
Hoodwinked: Deception and Cooperation in a Text-Based Game for Language Models 5 Jul 2023 · 1 repository · arXiv:2308.01404
-
Multilingual Controllable Transformer-Based Lexical Simplification 5 Jul 2023 · 1 repository · arXiv:2307.02120
-
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning 5 Jul 2023 · 0 repositories · arXiv:2307.02179
-
The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification 5 Jul 2023 · 0 repositories · arXiv:2307.02192
-
Embodied Task Planning with Large Language Models 4 Jul 2023 · 1 repository · arXiv:2307.01848Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
KDSTM: Neural Semi-supervised Topic Modeling with Knowledge Distillation 4 Jul 2023 · 0 repositories · arXiv:2307.01878Syntology 6 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Interpretability and Transparency-Driven Detection and Transformation of Textual Adversarial Examples (IT-DT) 3 Jul 2023 · 0 repositories · arXiv:2307.01225
-
Iterative Zero-Shot LLM Prompting for Knowledge Graph Construction 3 Jul 2023 · 0 repositories · arXiv:2307.01128
-
TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition 2 Jul 2023 · 0 repositories · arXiv:2307.00526
-
Large Language Models (GPT) for automating feedback on programming assignments 30 Jun 2023 · 0 repositories · arXiv:2307.00150
-
Meta-Reasoning: Semantics-Symbol Deconstruction for Large Language Models 30 Jun 2023 · 1 repository · arXiv:2306.17820
-
SPAE: Semantic Pyramid AutoEncoder for Multimodal Generation with Frozen LLMs 30 Jun 2023 · 0 repositories · arXiv:2306.17842
-
Stay on topic with Classifier-Free Guidance 30 Jun 2023 · 0 repositories · arXiv:2306.17806
-
A negation detection assessment of GPTs: analysis with the xNot360 dataset 29 Jun 2023 · 0 repositories · arXiv:2306.16638
-
Benchmarking Large Language Model Capabilities for Conditional Generation 29 Jun 2023 · 0 repositories · arXiv:2306.16793
-
Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors 29 Jun 2023 · 0 repositories · arXiv:2306.17156
-
Pareto Optimal Learning for Estimating Large Language Model Errors 28 Jun 2023 · 0 repositories · arXiv:2306.16564
-
Inferring the Goals of Communicating Agents from Actions and Instructions 28 Jun 2023 · 0 repositories · arXiv:2306.16207
-
Is ChatGPT a Biomedical Expert? -- Exploring the Zero-Shot Performance of Current GPT Models in Biomedical Tasks 28 Jun 2023 · 1 repository · arXiv:2306.16108
-
Taqyim: Evaluating Arabic NLP Tasks Using ChatGPT Models 28 Jun 2023 · 1 repository · arXiv:2306.16322
-
Evaluating GPT-3.5 and GPT-4 on Grammatical Error Correction for Brazilian Portuguese 27 Jun 2023 · 0 repositories · arXiv:2306.15788
-
SparseOptimizer: Sparsify Language Models through Moreau-Yosida Regularization and Accelerate via Compiler Co-design 27 Jun 2023 · 0 repositories · arXiv:2306.15656
-
Exploring the Robustness of Large Language Models for Solving Programming Problems 26 Jun 2023 · 0 repositories · arXiv:2306.14583
-
LongCoder: A Long-Range Pre-trained Language Model for Code Completion 26 Jun 2023 · 1 repository · arXiv:2306.14893Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Interactive Design by Integrating a Large Pre-Trained Language Model and Building Information Modeling 25 Jun 2023 · 0 repositories · arXiv:2306.14165
-
Let's Do a Thought Experiment: Using Counterfactuals to Improve Moral Reasoning 25 Jun 2023 · 0 repositories · arXiv:2306.14308
-
Is Pre-training Truly Better Than Meta-Learning? 24 Jun 2023 · 0 repositories · arXiv:2306.13841
-
Large Language Models as Sous Chefs: Revising Recipes with GPT-3 24 Jun 2023 · 1 repository · arXiv:2306.13986
-
Large Sequence Models for Sequential Decision-Making: A Survey 24 Jun 2023 · 0 repositories · arXiv:2306.13945
-
On the Uses of Large Language Models to Interpret Ambiguous Cyberattack Descriptions 24 Jun 2023 · 0 repositories · arXiv:2306.14062
-
LLM-Assisted Content Analysis: Using Large Language Models to Support Deductive Coding 23 Jun 2023 · 0 repositories · arXiv:2306.14924
-
System-Level Natural Language Feedback 23 Jun 2023 · 1 repository · arXiv:2306.13588Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale 23 Jun 2023 · 1 repository · arXiv:2306.15687Syntology 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 3 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Cross-lingual Cross-temporal Summarization: Dataset, Models, Evaluation 22 Jun 2023 · 1 repository · arXiv:2306.12916
-
Prompt to GPT-3: Step-by-Step Thinking Instructions for Humor Generation 22 Jun 2023 · 1 repository · arXiv:2306.13195
-
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair 21 Jun 2023 · 0 repositories · arXiv:2307.00012
-
Investigating Pre-trained Language Models on Cross-Domain Datasets, a Step Closer to General AI 21 Jun 2023 · 0 repositories · arXiv:2306.12205
-
Solving and Generating NPR Sunday Puzzles with Large Language Models 21 Jun 2023 · 1 repository · arXiv:2306.12255
-
Which Spurious Correlations Impact Reasoning in NLI Models? A Visual Interactive Diagnosis through Data-Constrained Counterfactuals 21 Jun 2023 · 0 repositories · arXiv:2306.12146
-
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models 20 Jun 2023 · 0 repositories · arXiv:2306.11698
-
Event Stream GPT: A Data Pre-processing and Modeling Library for Generative, Pre-trained Transformers over Continuous-time Sequences of Complex Events 20 Jun 2023 · 1 repository · arXiv:2306.11547Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)